Most RAG failures are not model failures.

The model did not forget how to read. The prompt is not necessarily bad. The embedding model is not always the culprit.

The answer was often lost earlier, in the part of the system nobody demos:

how documents were parsed,

how chunks were created,