20-year engineer benchmarks Gemma4-31B vs Qwen3.8-27B locally; GPT-6.1-Sol is still another tier

therealjerseytom · reddit · 2026-10-11

A software engineer with 20 years of experience tested Gemma4-31B vs Qwen3.8-27B (same Q4 quant) on work-like tasks, with GPT-6.1-Sol as reference:

Takeaway: both open-weight models are workable on consumer hardware, but clearly not in the same conversation as frontier models. Next up: the true developer hell of legacy codebases like Doom or Command & Conquer source.

Original post →

More from coding & agent

coding & agent channel →