How Long Until Local ~30B A3B Models Match GLM 5.3 Flash Quality?

Aggravating-Push-207 · reddit · 2026-09-26

A Reddit user asks how long until local 30B A3B models on a 16GB RAM + 8GB VRAM machine match GLM 5.3 Flash quality, noting the jump from Qwen3 Coder 30B A3B to Qwen 3.6 35B A3B took only 6 months. The thread debates the pace of small MoE model progress and whether cloud-free local AI could arrive within a year.

Original post →

More from Infra

Infra channel →