Dev burns 5B tokens a day on open models as multi-agent swarms emerge as a new scaling axis

xeophon · x · 2026-09-17

Developer Florian Brand reports his daily token usage exploded from under 100M to over 5B tokens (open models only), driven by more reliable long-running agents, cheap GPU access, and multi-agent swarms. Key points: single agents are hitting wall-clock-time limits, making parallel swarms an underexplored scaling axis; swarms excel at data work and broad research via agent-to-agent communication; SOTA open models can maintain persistent fleets of sub-subagents that decompose, delegate, and relay findings. He also teases a section on GPT-6's training. Even with visible inefficiencies at every step, "point a swarm at a problem and waste a bazillion tokens" is working better than expected.

Original post →

More from coding & agent

coding & agent channel →