llama.app launches: no-code local AI with one-click Gemma 4, MCP support
mervenoyann · x · 2026-09-10
llama.cpp (123.9K GitHub stars) launched llama.app, wrapping llama.cpp in a clean no-code UI: one-click model downloads, clear memory estimates, and support for Gemma 4, Qwen 3.8, GPT-OSS and Gemma 3. Demos show Gemma 4 parsing receipt tables, streaming reasoning logs, and connecting to MCP for web search. Workflow: llama serve, install the pi-llama plugin, and Pi coding agent auto-discovers your local model — no API keys, data stays on-machine. One binary runs from Apple Silicon and RTX 5090 to H100 and DGX Spark.
Related event: llama.cpp Launches llama.app Portal for One-Click Local LLMs(2 posts)→
More from Infra
- NASA chief backs orbital AI compute as SpaceX targets first space data center in 2027 — rohanpaul_ai · 2026-09-10
- NVIDIA ships open-source PAIR: turns idle home PCs into a local AI cluster, ~51% faster in demo — solyarisoftware · 2026-09-10
- Massachusetts moves to require community agreements before data center permits — Polymarket · 2026-09-10
- At Scale, KV Cache Becomes a Storage System: How LLMs Serve GBs of Cached State — blaizedsouza · 2026-09-10
- Don't let FOMO win: you can learn more about local LLMs with a tiny model than a 5090 — sn2006gy · 2026-09-10
- Stealth Chip Startup Kepler Emerges With EUV-Free 3D-Stacked HBM, $468M Raised — pstAsiatech · 2026-09-10