Perplexity Hybrid details: Supports Qwen and Gemma locally
testingcatalog · x · 2026-08-31
Details on Perplexity Mac Hybrid mode reveal a cloud orchestrator delegating tasks to local models to keep data on-device. Downloadable options include a 5.6GB Gemma model (16GB RAM), a 17.4GB Qwen 32B model (32GB RAM), and a 19GB proprietary Perplexity model (32GB RAM). The feature aims to reduce cloud credit usage and enhance privacy.
Related event: Perplexity Brings Hybrid Mode to Mac With Local Models(2 posts)→
More from Infra
- Using an iPhone 13 mini as a Reinforcement Learning Environment — willcb · 2026-09-01
- MTP released for Qwen3.8-Flash-Next GGUF, promising big local TPS gains — vini542reddit · 2026-09-01
- Google Cloud Monitoring MCP Connector Released — modelcontextprotocol · 2026-09-01
- Distributed.systems Launches Auditable Agent Infrastructure — arthurcolle · 2026-09-01
- Does enabling ChatGPT Memory or history reference increase token usage? — ssunki · 2026-09-01
- Engineer fixes ROCm inference crash on MI350X, uncovers 9 bugs in deep dive — AnushElangovan · 2026-09-01