RAG Interview Questions, issue 16, Jul 20, 2026

The Green Dashboard Paradox

Staff ML Engineer interview at Anthropic, and the interviewer asks:

Your Corrective-RAG system has an AI evaluator scoring its own retrievals. Six months in, answer quality is degrading, but your dashboards are green. What’s happening, and how would you have prevented it on day one?

Don’t say: The evaluator threshold needs retuning.

The hidden danger of AI vetting AI: why your evaluator gets more confident as your output tanks, and the architectural fix needed to anchor your system back to reality.

The full answer, with the mechanism and the arithmetic, is for paid subscribers on Substack.

Read it on Substack

Get the next one

Free on Substack. Unsubscribe in one click.