Running gemma4:31b-mlx locally with Ollama feels indistinguishable from paid tiers
walkingriver · x · 2026-09-06
A developer ran gemma4:31b-mlx locally on a MacBook via Ollama and found the answers essentially indistinguishable from his paid subscription, with 100% local inference — the only cost being a spinning fan.
More from Infra
- Musk: 3D printing enables integrated flow paths but is too slow and costly for volume production — i_bioloid · 2026-09-06
- Nvidia guides 70% revenue growth next year, supply sets the ceiling — BenBajarin · 2026-09-06
- Unreal Engine pipeline yields 8.7K+ hours of action-conditioned video for world-model pretraining — udmrzn · 2026-09-06
- PyPI's Recent Download Corrections Sharply Cut Some Packages' Stats — dbreunig · 2026-09-06
- Redditor seeks a lightweight OpenAI-compatible API client, complains OpenWebUI is 30GB — MelodicRecognition7 · 2026-09-06
- SK Hynix weighs Intel Foundry for part of HBM4e base die output as TSMC costs run 3-4x higher — Beth_Kindig · 2026-09-06