Reviewing Grok 4.5's Website Building Performance
vista8 · x · 2026-07-09
The poster is formally reviewing Grok 4.5, having it develop a website for AI model evaluation cases. Current observation: its skill call accuracy is good; next is to see the final website development outcome.
Related event: Testing Grok 4.5 in Web Dev: Fast Speed and Solid Frontend(4 posts)→
More from coding & agent
- Soft Clamp cuts tool-call overuse in multi-teacher distillation, from 13.7% to 9.0% — antgroup · 2026-07-21
- Agent harness memory loss and compaction are still a major usability problem — adityaag · 2026-07-21
- SpecJudge runs locally on Ollama to pick the right-sized AI model for your project — jokiruiz · 2026-07-21
- A developer maps out six design rules for CLIs that humans and AI agents can both use — yujiezha · 2026-07-21
- A coding-agent skill that forces ADHD-friendly, answer-first output — ayghri · 2026-07-21
- A set of agent skills for CAD, robotics, and hardware design — earthtojake · 2026-07-21