Qwen Code Releases SWE-bench Verified Run: Quarantined with 0 Resolved
DennisYu07 · ghdev · 2026-08-13
QwenLM/qwen-code released a non-production SWE-bench Verified E2E validation run (dsw-eas-full-20260813-r1). The run status is QUARANTINED, with 500/500 completed but 0 resolved, 0 unresolved, 0 execution errors, and 0 infrastructure failures, so no score was published. The model used was qwen3.7-plus with Qwen Code v0.21.11.
More from coding & agent
- Tencent Releases UI-Mate-27B, a Desktop GUI Agent Model — tencent · 2026-08-24
- Comparing AI Subscriptions: DeepSeek API vs. Claude Pro vs. Local LLMs — Unlikely_Bluejay5392 · 2026-08-24
- Claude Code introduces 'Remote Control' feature to boost coding efficiency — rohanpaul_ai · 2026-08-24
- rauchg lays out fx extension philosophy: MCP, Skills, Plugins and Unix composition — AccBalanced · 2026-08-24
- Netflix details its production LLM judge: hundreds of thousands of recommendations scored weekly — omarsar0 · 2026-08-24
- smolvm passes Simon Willison's Fable 5 agent test as a secure sandbox — yawnxyz · 2026-08-24