llama.cpp adds support for Moonshot Kimi-K3 text model
pmttyji · reddit · 2026-08-15
A Pull Request proposes adding the Kimi-K3 text model to the ggml-org/llama.cpp project. This allows developers to run Moonshot's Kimi-K3 model locally on their devices, further enhancing llama.cpp's support for major Chinese LLMs.
More from Infra
- LLMRouter open-sources 16+ implementations with xRouteBench for LLM routing — Justgototheeffinmoon · 2026-08-16
- User Benchmarks Qwen3.8 27b on M5 Max: 8t/s (bf16), 17t/s (8bit) — julianharris · 2026-08-16
- The accountability gap in LLM inference: Proving which model weights actually ran — Some_parts_Bi · 2026-08-16
- WeeLLM: Run FLUX.1-dev on 4GB VRAM Without Quantization — AlarmingPhrase8174 · 2026-08-16
- Enterprise AI buyers push Lenovo to record quarter: revenue up 43%, services margin 3x PC — shashib · 2026-08-16
- Google floats many TPU RFPs, takes more wafers directly to TSMC each generation — BenBajarin · 2026-08-16