Users rank Sonnet 5 above Grok 4.5 and Gemini 3.6 Flash on coding tasks
firstadopter · x · 2026-07-21
A user says Gemini 3.6 Flash trails Grok 4.5 and Sonnet 5 on coding tasks
The post quotes a comparison claiming that Gemini 3.6 Flash performs worse than Grok 4.5 on coding tasks, and ranks the models as:
- Sonnet 5
- Grok 4.5
- GPT 5.6 Luna
- Gemini 3.6 Flash
The author adds that they once praised Google, but now want the company to “be serious” again, suggesting frustration with the model gap in coding quality.
More from Models
- Rumor claims Kimi faked performance by serving Claude; DeepSeek new model surprises in evals — realsohamparekh · 2026-09-11
- GPT-5.6 writes well but is instantly forgettable, user complains — BasedRaddka · 2026-09-11
- Opus Refuses Protein Research Codebase Over 'Safety' Concerns, Dev Considers Rolling His Own — josephdviviano · 2026-09-11
- User Hails Unconfirmed 'DeepSeek 4.1 Flash' as an Inflection Point in LLMs — himanshustwts · 2026-09-11
- Terminal Bench v4: GLM-5.3 Leads at 41.9%, Kimi-K3 Underwhelms at 12.6% — Ok_Warning2146 · 2026-09-11
- GPT-6 Astra beats Factorio with enemies in 44 in-game hours at ~$4,500 API cost — liminal_bardo · 2026-09-11