AI agents spend 4 weeks and $3k in GPUs to produce a new math result
giannis_daras · x · 2026-09-15
Dimitris Papail previewed a research result produced by a 4+ week collaboration among AI agents (Astra, Sol & Fable), where his only role was asking questions and paying $3k in GPU costs. A draft plus 100GB of verification certificates will be shared.
Key details from the thread: the GPU budget went to parallelizing lower/upper bound computations, with agents exploring many strategies—different input sources, bounding approaches, and deeper complexity relaxations. The setup used vanilla Codex and Claude Code with no harness optimization, just models talking to each other with handoffs; the process was messy and hard to replicate, though the result itself is reproducible.
More from coding & agent
- How Much Code Is Throwaway Now? Engineers Rethink Project Lifespans in the AI Era — rakyll · 2026-09-15
- Translation MCP team: the hard part wasn't the server, it was getting the model to call it — Pure-Ad7786 · 2026-09-15
- Developer runs Claude inside Siri on macOS 27 via Apple's hidden model-provider support, open-sourced — Teknium · 2026-09-15
- OpenResearch tops GitHub trending, turns Claude Code and Codex into research agents — TheMoonMidas · 2026-09-15
- CoreSpeed Launches: One MCP Endpoint for Tools, Memory and Agent Budgets — Kind-Atmosphere9655 · 2026-09-15
- Track anything cheaply: Astra + SAM 3 pipeline cuts tokens 26x — deepakns · 2026-09-15