Ragas is the standard framework for measurable RAG evaluation: hallucination, faithfulness, relevancy
mdancho84 · x · 2026-09-19
The thread's #6 pick is Ragas, described as the standard framework for evaluating RAG pipelines — hallucination, faithfulness and relevancy are all measurable, and it can generate production-aligned test sets when you don't have one. The adjacent pick is Qdrant, a production vector database written in Rust (34.7k stars) that JDs list alongside Pinecone, Chroma, and FAISS.
Related event: Matt Dancho Lists 10 GitHub Repos to Become an AI Engineer in 90 Days(11 posts)→
More from coding & agent
- 8 components of harness engineering: why the system around the model, not the model, makes agents reliable — blaizedsouza · 2026-09-19
- Google's 12-page Agentic Engineering guide lays out a 5-stage pipeline for building agent teams — blaizedsouza · 2026-09-19
- AgentRun: a purpose-built harness for high-frequency repetitive knowledge work — blaizedsouza · 2026-09-19
- Qwen 3.8 27B on a single RTX 5090 builds a full animation using only code — Acceptable-Object390 · 2026-09-19
- Compiler Explorer runs 92M compilations a year — here's how it works — blaizedsouza · 2026-09-19
- Netflix builds module-first, agent-friendly Java tooling on OpenJDK foundations — blaizedsouza · 2026-09-19