Qwen3.8-Flash-Next runs fully locally on a Xiaomi 14T Pro phone CPU
Tall_Abrocoma_3533 · reddit · 2026-09-05
A user reports running Qwen3.8-Flash-Next-UD-IQ3XXS entirely locally on a Xiaomi 14T Pro using the BigMoeOnEdge app—no GPU needed, showing small quantized models now run on phone CPUs.
More from Infra
- Extropic unveils Z1T sparse models claiming up to 140x energy efficiency gains over GPUs — beffjezos · 2026-09-05
- DeepSeek to deploy 160,000 Huawei next-gen AI chips in Inner Mongolia data center — Polymarket · 2026-09-05
- DeepSeek to deploy at least 160,000 next-gen Huawei AI chips at massive Inner Mongolia data center — Polymarket · 2026-09-05
- Nvidia guides FY28 to ~$691B: non-hyperscaler AI customers now half of business, growing 100% a year — Beth_Kindig · 2026-09-05
- Coatue in talks to form multibillion-dollar JV with chip startup MatX to finance die purchases and foundry capacity — steph_palazzolo · 2026-09-05
- Running an Opus-level coding agent locally at 2x speed for free: a 15-page report — julianharris · 2026-09-05