Inco Splash hits 144 tok/s running Qwen3.8-27B on an M5 Max, 3x faster than Ollama

TheMoonMidas · x · 2026-09-19

Inco released Splash, an open-source inference engine built specifically around Apple silicon, with the DFlash 2 draft model pre-integrated and runnable in two commands.

Related event: Inco Splash Open-Source Engine Runs Qwen3.8-27B at 144 tok/s on M5 Max(2 posts)→

Original post →

More from Infra

Infra channel →