In real C++ debugging, open-source Qwen 3.8 27B showed better engineering judgment than Gemini 3.7 Flash

GravyPoo · reddit · 2026-08-20

The author compared Gemini 3.7 Flash (High) vs open-source Qwen 3.8 27B as coding agents on a heavily modified OrcaSlicer C++ fork (TBB threading, slicing geometry, multi-tool scheduling, G-code generation, regression tests).

Key finding: the gap wasn't code generation but engineering judgment.

Conclusion: not universally smarter, but more trustworthy engineering judgment for long-running repo-level debugging. Qwen ran at FP8 quantization, 262K native context, reasoningeffort=xhigh.

Original post →

More from coding & agent

coding & agent channel →