Perplexity open-sources Lily, its Apple Silicon inference engine for Qwen3.6-35B-A3B
inductionheads · x · 2026-09-03
Perplexity has open-sourced Lily, the local inference engine powering its new hybrid compute feature in the Mac app. Lily is specialized for Qwen3.6-35B-A3B on Apple silicon, built so on-device compute doesn't bottleneck Computer tasks, splitting work between local and cloud.
Related event: Perplexity Open-Sources Lily, a Qwen3.6 Inference Engine for Apple Silicon(8 posts)→
More from Infra
- DeepSeek-V4-Pro ships with 1.6T-param MoE; open-source eval harness steals the show — DeepLearningAI · 2026-09-03
- Broadcom CEO: Anthropic and OpenAI to become its No.1 and No.2 custom AI chip clients — dinabass · 2026-09-03
- Hock Tan says power availability dictates the exact timing of compute capacity deployment — BenBajarin · 2026-09-03
- Perplexity open-sources its Mac inference server optimized for Qwen 3.6 on Apple Silicon — Specter_Origin · 2026-09-03
- GLM-5.3-Flash hits 1,005 tok/s locally on dual RTX PRO 6000 Blackwell cards — BanghuaZ · 2026-09-03
- Nvidia now ~8% of S&P 500 market cap, worth 16.3% of US GDP at $5.3T — ivan_bezdomny · 2026-09-03