Grok 4.6 tops VISTA benchmark, turning Figma designs into web apps at $2.38 per task
XFreeze · x · 2026-08-17
Grok 4.6 has topped the VISTA benchmark, which tests AI coding agents' ability to understand Figma designs and build functional web apps. It outperformed Claude Fable 5, Opus 5, and GPT-5.6 Sol, with a cost of about $2.38 per task. This demonstrates the full design-to-code capability, which is practically relevant for developers.
More from coding & agent
- Researcher: Agents that can work in a sandbox shouldn't have outbound access at all — moniquejmorrow · 2026-10-02
- llama.cpp adds decision models: typed questions, per-option probabilities in one forward pass — ngxson · 2026-10-02
- Agents aren't GPU-bound: tool execution, memory bandwidth and sandbox overhead are the real bottleneck — ai · 2026-10-02
- Ruff creator: 'no one writes code anymore' isn't the point — does anyone read it? — charliermarsh · 2026-10-02
- 8 Practical Upgrades to Turn Fragile Chat Loops Into Production-Grade Agents — alexcovo_eth · 2026-10-02
- Event-Driven Agents Are Right for Real Work, But AWS Stack Makes Humans the Bottleneck — TansuYegen · 2026-10-02