Limite 1B runs locally on M4 Pro at ~57 token/s, scoring 4/8 on AIME sampling

A mathematician ran Paradigma's Limite 1B (Violetto) locally on M4 Pro via an experimental MLX port at about 57 tokens/s, scoring 4/8 on sampled AIME 2025 problems. An account believed to be official noted the model averages 50k output tokens on AIME and recommends allowing longer generation time.

2026-09-23 ~ 2026-09-23 · 2 related posts