Skip to content

Comment on Contextual Retrieval

Comments

Can someone explain simply how these benchmarks work?

What exactly is a "failure rate" and how is it computed?

They simply ask the AI a question about a large document (or set of docs). It either gets the answer right or wrong. They count the number of hits and misses.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.