The most useful part wasn't building another RAG pipeline, it was learning how to actually evaluate one. The RAG Triad (answer relevance, context relevance, groundedness) gives you a way to catch exactly where a system is failing: bad retrieval vs. the model making things up vs. just answering the wrong question.