Qwen3.8-27B tested against models 100x its size in first impressions video
arena · x · 2026-08-22
Arena released a first impressions video of the Qwen3.8-27B model with Peter Gostev. They tested it against models up to 100x larger, including DeepSeek v4, Qwen 3.8 Max, Kimi K3, GLM 5.3, Grok 4.6, GPT-5.6, and Fable. The benchmarks covered identical one-shot generation and agentic tasks, with full scores coming soon.
More from Models
- Anthropic opens Mythos 5 access for Claude Enterprise security beta — AccBalanced · 2026-08-22
- Rumor suggests Ox Alpha stealth is Cursor Composer based on GLM 5.2 — teortaxesTex · 2026-08-22
- OpenAI Launches Ultrafast GPT-5.6 Sol with 14x Speed Boost — gajesh · 2026-08-22
- Users Report Degradation in OpenAI Sol High Chat/Codex — Illustrious-Bet-1368 · 2026-08-22
- Discussion: Which Public LLM Benchmarks Do You Actually Trust? — ThomasAger · 2026-08-22
- Ornith 1.5 32B tested: Fast performance in Three.js demo — draginol · 2026-08-22