RAG Interview Questions, issue 16, Jul 20, 2026
The Green Dashboard Paradox
Staff ML Engineer interview at Anthropic, and the interviewer asks:
“Your Corrective-RAG system has an AI evaluator scoring its own retrievals. Six months in, answer quality is degrading, but your dashboards are green. What’s happening, and how would you have prevented it on day one?”
Don’t say: “The evaluator threshold needs retuning.”
The hidden danger of AI vetting AI: why your evaluator gets more confident as your output tanks, and the architectural fix needed to anchor your system back to reality.
The full answer, with the mechanism and the arithmetic, is for paid subscribers on Substack.