Gemma 4 26B A4B inference on Mac is now 2x faster
GlennCameronjr · x · 2026-09-02
Google announced that Gemma 4 26B A4B is now running 2x faster on Macs, thanks to the developer community pushing Apple Silicon to its limits on the leaderboards.
More from Models
- Scale AI Releases Muse Voice: SOTA Streaming Speech-to-Text Model — alexandr_wang · 2026-09-02
- PINNACLE Benchmark Scores Cost per Correct Task: Models Are Good Enough, Who Decides? — ryanshrout · 2026-09-02
- Meta Releases Muse Voice Transcribe: Real-Time Streaming Speech Model with Diarization — bowenc0221 · 2026-09-02
- Gemini Adds Agentic Video Understanding, Cuts Token Usage by 88% — GoogleDeepMind · 2026-09-02
- Atlas runs on one unified multimodal autoregressive diffusion transformer, pretrained from scratch — theworldlabs · 2026-09-02
- World Labs unveils Atlas, an omni world model for spatial intelligence, early access soon — theworldlabs · 2026-09-02