20B Model Maple-Preview Runs at 200+ tokens/s on Mac Mini, Solves IMO Math

tylerbruno05 · x · 2026-08-05

DeepGrove has introduced Maple-Preview, an open-source 20B-A1B ternary-weight reasoning LLM, claiming SOTA performance in its weight class.

The model showcases extreme edge-inference efficiency:

Original post →

More from Infra

Infra channel →