Dev argues prompting alone can't get AI to solve natural science problems
felpix_ · x · 2026-09-22
Responding to OpenAI's claim that its model solved "100+ long-standing open problems" in math within two weeks of the NS announcement, developer felpix argues on X that no average person—or even a 90th-percentile user—could prompt the model to solve natural science problems. Beyond cost and compute, you'd need a team of PhD experts who know the state of the literature and which ideas have promise—which, he notes, is exactly the team OpenAI assembled.
More from Models
- Anthropic investigates elevated errors across Claude Mythos 5.1, Fable 5.1 and Opus 5 — ClaudeAI-mod-bot · 2026-09-22
- Dev Swaps Opus for Mimo-v2.6 in Cline on Client Projects: 'It's a Beast' — MicahBerkley · 2026-09-22
- Kev refactored onto Qwen3.5: open-source decision models now at 0.8B, 4B and 9B — alexcovo_eth · 2026-09-22
- Why OpenAI bets on math: it's the most verifiable domain for reinforcement learning — burny_tech · 2026-09-22
- Open-source decision model Laya ported to Core ML: 99.5% ops on ANE, 3.7ms per decision — alexcovo_eth · 2026-09-22
- Fireworks: routing 18 models per task hits 97.6% solve rate at $1.88 vs best single model's 74.1% at $6.52 — sophiamyang · 2026-09-22