llama.cpp ships Day-0 support for Google's EmbeddingGemma 2
ggerganov · x · 2026-10-07
llama.cpp creator Georgi Gerganov announced Day-0 support for Google's EmbeddingGemma 2, meaning the new embedding model can be used locally in llama.cpp from release day — handy for on-device embeddings and retrieval workloads.
More from Models
- User Reports ChatGPT Mobile Bug Exposing Hidden Instructions in Voice Output — VoidStateKate · 2026-10-07
- Paradigm releases tech report for Limite 1B-Violetto, a from-scratch math model — tensorqt · 2026-10-07
- ChatGPT Reportedly Removes Free-Tier Text Chat Limits and Upgrades Default Model — hey_abusiddik · 2026-10-07
- Qwen3.8 Flash Next IQ1_M hits 55 tok/s on a 5060 Ti 16GB and still codes well — bobaburger · 2026-10-07
- EmbeddingGemma 2 runs offline multimodal RAG on phones in ~191MB–567MB of RAM — Saboo_Shubham_ · 2026-10-07
- Paradigm evals its math model across 7 hard benchmarks, releases full eval suite — tensorqt · 2026-10-07