AI Discovers Counterexamples in Statistics
teortaxesTex · x · 2026-07-15
- This post recounts a case of AI making a new discovery in statistics: by combining known concepts, a model generated a non-standard yet valid counterexample.
- The original text emphasizes that the output distribution of LLMs is broader than human thinking, making them better at generating such counterexamples. For statisticians, these are open problems that genuinely warrant manual construction.
- The quote also mentions that GPT-5.6 has already solved this conjecture. The author feels that while the AI did the job, they would have much preferred it to be solved by a human.
Related event: GPT-5.6 Reportedly Found a BH Counterexample(8 posts)→
More from Models
- Same Echo Maze prompt, three frontier models: all passed visually but shipped the same hidden bug — eyishazyer · 2026-09-11
- Benchmark scores drop from 89% to 19% on new evals — how benchmaxxing breaks leaderboard trust — airesearch12 · 2026-09-11
- ChatGPT tells user their question is too hard and to 'accept dumber answers' — phido3000 · 2026-09-11
- Claude is no longer available for minors as Anthropic rolls out age assurance — Muhammad523 · 2026-09-11
- Developer Building a Unified Leaderboard of All Model Benchmark Scores — airesearch12 · 2026-09-11
- Rumor claims Kimi faked performance by serving Claude; DeepSeek new model surprises in evals — realsohamparekh · 2026-09-11