Buried the answer in noise
Below release gateMeasures what fraction of the retrieved context was actually relevant to the query, and fires when most of it was not.
For exampleAnswered a pricing question after retrieving eight pages, only two of which were actually about pricing.
context_precision
Made claims no source backed
Below release gateVerifies an answer against the source documents it was supposed to be built from, with the emphasis on numbers, named entities, and quoted attributions.
For exampleReported the wrong revenue figure and credited the forecast to an analyst who never said it.
grounding
Pulled up unrelated passages
Below release gateScores every retrieved chunk against the query that fetched it and reports the ones that do not address the question.
For exampleAnswered a policy question mostly from unrelated pages, with only one thin excerpt actually on topic.
chunk_relevance