Embeddings 1400x Cheaper Than LLMs, Hybrid Search Remains Top Choice
jeremyphoward · x · 2026-08-22
Addressing whether LLMs can replace embedding models, the new paper "Embedder's Dilemma" finds that while LLMs now outperform specialized embedding models, they cost 1400x more. The author advocates sticking with embeddings (dense, sparse, multi-vector) combined with BM25 and rerankers. Listwise cross-encoders are also noted as an interesting option.
Related event: Harvard-Stanford Paper: LLMs Match Embedding Models but Cost ~1431x More(6 posts)→
More from Research
- Retriever: A Framework for Asynchronous, Closed-Loop Robot Agents — ZeYanjie · 2026-08-24
- Converting GMMs ↔ PEFs for fast KLD approximation — FrnkNlsn · 2026-08-24
- Netflix details its production LLM judge: hundreds of thousands of recommendations scored weekly — omarsar0 · 2026-08-24
- Nature Comment: Provenance, not interpretability, grounds trust in autonomous science — gabepgomes · 2026-08-24
- New Architecture RHEA: Train 1B Model on 8GB VRAM — zemondza · 2026-08-24
- Trained two 16M-param models to do generative CAD with real physics — debreuil · 2026-08-24