Stanford, Yale Win First Live AI Agent Competition at Data + AI Summit
CShorten30 · x · 2026-08-19
Databricks hosted a live competition at Data + AI Summit challenging AI researchers to adapt their agents in real-time to a brand-new benchmark, OfficeQA Pro V2. The event aimed to test whether benchmark performance improvements generalize to new tasks. Teams from Stanford, UMass Amherst, and Yale won the competition. A blog post detailing the winning strategies is available.
Related event: Stanford-Led Team Wins Databricks Grounded Reasoning Cup at 63.3%(3 posts)→
More from coding & agent
- Gemini Image Generation Silently Fails From Hetzner IPs — Network Origin Was the Culprit — dota2dinall · 2026-08-19
- Vercel open sources fx: a tiny, fast native coding agent — Rasmic · 2026-08-19
- GitHub Repo Open-Sources Author Style Mimicry Prompts Featuring Ottessa Moshfegh — TuhinChakr · 2026-08-19
- Open source file upload service Byteship built with Grok released — jasonkneen · 2026-08-19
- ClawGym II paper: Improving agents via mixed-harness training — omarsar0 · 2026-08-19
- MacStories' Codex Automation Guide: Process Notes, Save Emails, Auto-Tag Read-Later — Dimillian · 2026-08-19