Ternary MoE Model Scion-35B-A3B Released on Hugging Face With llama.cpp Support
pmttyji · reddit · 2026-10-06
SkyIsNotGreen/Scion-35B-A3B has landed on Hugging Face: a ternary-weight MoE model with 3B active parameters out of 35B total.
- Ships with a llama.cpp fork (prism-ml-llama.cpp) offering prebuilt binaries for instant testing.
- Also released the companion Scion-35B-A3B-mtp-drafter draft model for speculative decoding.
More from Models
- Anthropic reviewers alerted police to a Claude chat threatening a sheriff's office, leading to an arrest — rohanpaul_ai · 2026-10-06
- flow-1: RL-trained model matches GPT-6-sol at trace debugging while 23x cheaper — kalyan_kpl · 2026-10-06
- First large-scale 3B/8B continuous diffusion LMs match pass@1 and beat pass@k vs masked dLMs — ArashVahdat · 2026-10-06
- Claude nitpicks, Codex says LGTM: what happens when AI rivals review each other's PRs — _lewtun · 2026-10-06
- OpenAI boosts default GPT-6 Astra and GPT-6.1 Sol speed ~50% to 50 tokens/sec — kimmonismus · 2026-10-06
- 8GB local image gen reality check: 2/10 prompt accuracy vs 9/10 on cloud APIs — Tricky-Brother-7 · 2026-10-06