Underdog's Husky Inference Engine Claims 4.5x Speedup Over MLX, 730 tok/s on MacBook

jimmykoppel · x · 2026-09-22

Underdog launched Husky, a Model-Specific Inference (MSI) engine claiming up to 4.5x speedups over Apple's MLX, with its Pareto-frontier local model hitting up to 730 tokens/sec on a MacBook.

Husky powers Underdog's invite-only, 100% local and private AI assistant (under 4GB) that uses the browser to book flights, reserve tables, order groceries and cancel subscriptions with user approval, connects directly to Gmail/Outlook for on-device email, and does fully local meeting transcription and notes — no audio or data ever leaves the machine.

Original post →

More from Infra

Infra channel →