Retrieval misses, lost-in-the-middle, stale indexes, chunk fragmentation — a triage guide.
Skip to solutionKEEP THE
hardAI Engineering
What are common RAG failure modes and how do you debug them?
562 views
01
Understand the problem
ragdebuggingfailure-modesretrieval
02
Attempt it yourself
Sketch your approach before reading the solution — that's what interviews test.
Nudge consolestandby
Stuck? Beam a request up — the console returns a conceptual nudge that guides your logic without spoiling the implementation.
03
Study the solution
Frequent failures: the answer chunk is not retrieved (vocabulary mismatch, bad chunking, wrong k), retrieved but ignored (buried mid-context, conflicting chunks), stale or duplicated index content, and questions that need synthesis across many chunks (retrieval alone cannot). Debug by logging the retrieved set per quer
Solution ready — 2 min read
Classified // press E to declassify
04
Read the code
Per-request retrieval logging for triage
const chunks = await retrieve(query, { k: 8 });
const reply = await generate(query, chunks);
await log.rag({
query,
retrieved: chunks.map((c) => ({ id: c.id, score: c.score, doc: c.docId })),
answer: reply.text,
feedback: null, // filled by thumbs up/down later
});
// weekly: cluster thumbs-down rows by "gold chunk retrieved?" → fix the bigger pile05
Join the discussion
Discussion (0)
Sign in to join the discussion.
No responses yet. Be the first to share what you think.
Transmission complete // awaiting log
KEEP THE
STREAK ALIVE.
Dossier 70 of 80 decoded in the AI Engineering track. One more won't hurt.