GLM 5.3 Flash (380B, local 8-bit) praised for one-shot note app beating frontier models

9r4n4y · reddit · 2026-09-17

A user reports building their long-wanted note app in one shot with GLM 5.3 Flash running locally (8-bit quant, vLLM, 380B parameters), claiming the output looked better than Gemini 3.8's attempt at the same app and even beat Opus 4.6 Adaptive Max, with very few bugs. They also used Qwen 3.6 35B A3B for online research. Subjective hands-on impressions only — no screenshots or benchmarks — so treat with skepticism.

Original post →

More from Models

Models channel →