Qwen 3.8 27B Beats Claude Opus 4.6 in Three.js Coding Test for Free
testingcatalog · x · 2026-08-15
AI/ML API conducted a benchmark task: generating a single self-contained three.js file to render a sunken Atlantis scene (temple, columns, fish, kelp, caustics) with a scripted 20-second camera flythrough.
Results show that the open-weights Qwen 3.8 27B (self-hosted) outperformed Claude Opus 4.6 in code density and scene richness:
- Qwen generated a full temple colonnade, domed roof, and a centered glowing portal, with layered elements like fish and kelp.
- Opus produced a long row of columns fading into fog, with a small, off-center portal and a thinner scene.
Cost-wise, Qwen ran for free (self-hosted), while Opus cost $0.21 per run.
More from coding & agent
- Ex-Meta Scientist: Agents should use the web like humans via pixels and clicks — DhruvBatra_ · 2026-08-15
- reBot Arm Control Stack Integrates Agentic AI, VLM, and LLM for Robotics — kamathsblog · 2026-08-15
- Dev's Cloud Agent experience: escaping pesky permission prompts — davidcrawshaw · 2026-08-15
- Tested 3 models to spec a local AI-brain install: one cited real files, one got macOS compat backwards — schwentker · 2026-08-15
- Claude Code CLI 2.1.233: Added Sandbox, GitLab MR Support — ClaudeCodeLog · 2026-08-15
- Claude Code 2.1.233 changelog: per-user spend headers, Bash memory cgroups, MCP fix — ClaudeCodeLog · 2026-08-15