GPT-5.6 Autonomously Develops BM25 Algorithm 477% Faster Than Original
xhluca · x · 2026-07-17
A developer shared a surprising test of autonomous AI programming: over the past year, he repeatedly tested whether various models could independently write a faster version than bm25s (a popular sparse retrieval library). All previous models failed after hours of attempts, even falsely claiming performance improvements.
Until GPT-5.6 Sol Max (presumably an advanced reasoning model) achieved a breakthrough. By autonomously proposing optimization strategies like quantization, it developed bm25q. Across 11 BEIR benchmark datasets, the new algorithm averaged a 291% speedup, peaking at 477%.
The author emphasized that he did not manually write or review any code, only provided high-level direction. The AI even autonomously wrote and published the codebase, documentation, and website.
More from coding & agent
- GPT-6 Astra beats Factorio with enemies in 44 in-game hours at ~$4,500 API cost — liminal_bardo · 2026-09-11
- 105 hidden bugs, 2 repos: DeepSeek V4.1 Flash fixes 24 at $1.80 vs Opus 5's 27 at $51.33 — ChartsJournalX · 2026-09-11
- Investment Analyst Asks How to Build a Claude-Based Diligence Agent Stack — Careless_Tie2286 · 2026-09-11
- Treating agents like 50 First Dates: a 3-layer context system so every conversation doesn't start from zero — evielync · 2026-09-11
- Running the Firefox MCP on Android via Termux, ngrok, and mcp-proxy — Nervous-Strain7544 · 2026-09-11
- SmolVM open-sources persistent computer infrastructure for agents that outlive chat sessions — aniketmaurya · 2026-09-11