Phonon-2 hits 606x real-time on a MacBook Air: one hour of speech in 6 seconds

julianweisser · x · 2026-10-07

Manan announces a Core ML package for the speech model Phonon-2, reaching 606x real time on a base M5 MacBook Air — up from 174x at launch — transcribing an hour of speech in about 6 seconds with no quality loss. It's now the default engine in the app Detta, with peak memory cut from 3.1GB to 1.1GB and idle memory from 1.8GB to 600MB even with Gluon loaded, matching WisprFlow's footprint while running fully locally.

Original post →

More from Infra

Infra channel →