Nativ brings LiquidAI LFM2.5 to Mac locally: 82 tok/s decode, <8.5GB RAM for 128K context
JosephJacks_ · x · 2026-08-06
Open-source local AI tool Nativ has announced Day 0 support for the newly released LFM2.5-2.6B model by LiquidAI.
Running on Apple Silicon Macs, the model operates without quantization (Full BF16) and delivers impressive performance:
- Prefill speed: 11,000+ tokens/s
- Decode speed: 82 tokens/s
- Memory usage: Peak memory under 8.5GB (even at full 128K context)
- Concurrency: Scales to 476 tokens/s aggregate across 16 concurrent requests
Nativ is a free, open-source macOS app optimized via MLX-VLM, supporting text, vision, video, code, and audio. It focuses on a fully local, account-free, and subscription-free deployment experience.
Related event: LiquidAI LFM2.5 Runs Locally on Mac with Strong Performance(3 posts)→
More from Infra
- Samsung to Lock 60-70% of Production in Long-Term Deals, Tech Giants as Key Clients — Beth_Kindig · 2026-08-06
- SanDisk Executives Assert: Over 80% Gross Margin is a 'Fair Return' — firstadopter · 2026-08-06
- Google's AI token processing surges 330x in two years, signaling booming inference demand — Beth_Kindig · 2026-08-06
- Long Contexts Multiply Speculative Decoding Gains, Acceptance Rate Nears 100% — DjCanalex · 2026-08-06
- The Enterprise AI Question: Where Does Your AI Actually Run? — DavidLinthicum · 2026-08-06
- AI Inference Demand Growing 10x Yearly Will Make Compute Scarcity the Default — TansuYegen · 2026-08-06