Why these kits exist. Before building, I cataloged 21 RAG explainers, courses and videos and scored each with a decision model on four questions: does it teach evaluation mechanics (golden set, retrieval hit vs citation hit, the judge, the gate), does it teach retrieval mechanics with numbers (threshold, BM25 + cosine fusion, top-k), is it usable from zero, are the numbers from a real run.
No external resource scored above 0.5 on both mechanics columns. The two that teach evaluation are code courses for people who already know the words. The free RAG visualizers all stop where retrieval ends.
The method, the full table (theirs and mine, same questions) and the raw files are public: niktechai.com/blog/rag-explainers-scored
If you find something I missed, tell me and I will add it to the table.