RAG on Sparse Database Seeking Help
Desperate-Vast-4899 · reddit · 2026-07-10
The author is building a RAG pipeline with Ollama and Qwen2.5, aiming to let the model access a risk register database via SQL queries and semantic retrieval to answer questions about risks, incidents, and mitigation measures. The current problem is that the database is too sparse, with many empty tables and columns, leading to retrieval results lacking context and low answer efficiency. The author has tried chunking by rows and embedding, but hasn't used more advanced methods like hybrid search or RRF, and is seeking improvement advice.
Related event: Help Needed: Building RAG Pipelines on Sparse SQL Databases(2 posts)→
More from coding & agent
- The author says Codex reached 20x and is now debugging spec decoding on a hybrid parallel setup — TheZachMueller · 2026-07-21
- Axcess adds an MCP connector for WCAG accessibility checks that scanners miss — modelcontextprotocol · 2026-07-21
- X post asks whether Cursor Composer, built on Kimi models, would also be banned — max_paperclips · 2026-07-21
- A developer’s Codex usage is draining pooled enterprise credits at a small company — Distinct_Relation_62 · 2026-07-21
- Qwen Code ships cua-driver-rs 0.7.3 with relative coordinates and MCP filtering — github-actions[bot] · 2026-07-21
- Matt Pocock says every new codebase turns legacy within days — mattpocockuk · 2026-07-21