FULL STORY

OpenAI's Math Claims Draw Skepticism

OpenAI claimed its newly trained model solved over a hundred open math problems, but a leak about an NS breakthrough and doubts about prompt-level reproducibility sparked controversy.

2026-09-22 ~ 2026-09-22 · 3 episodes · 49 posts

Episode 1 · OpenAI Says New Model Solved 100+ Open Math Problems, Forms Advisory Group (2026-09-22, 44 posts)

On September 22, OpenAI announced it is partnering with an independent advisory group of mathematicians to guide how it responsibly shares progress at the intersection of AI and mathematics. The group operates independently of OpenAI, its members receive no compensation from the company, and it may proactively offer advice the company didn't request, comment on OpenAI's impact on the math community, and make its recommendations public.

Confirmed

  • The advisory group's responsibilities include: how to evaluate and disseminate new mathematical results, maintaining academic and professional standards, and building tools that support mathematical research and learning
  • OpenAI said the goal is to place mathematicians at the center, shaping how AI supports mathematical understanding and benefits the broader community
  • The group was formed against the backdrop of OpenAI's AI solving more than 100 open math problems, which raised concerns in the math community (as relayed by @RexDouglass)

Why it matters

  • AI producing mathematical results at scale could disrupt existing norms for academic publishing and authorship; the independent advisory group is an attempt at industry self-regulation
  • The group's independence (unpaid, free to publish advice proactively) is key to its credibility, and is seen externally as a test case for how open OpenAI really is
  • @RexDouglass noted some online predicted the group would disband within a year, reflecting skepticism among parts of the public about its effectiveness

24 more related posts →

Episode 2 · Rumor: OpenAI's New Model Solved NS Math Problem 11 Days After Training Start (2026-09-22, 2 posts)

An unverified leak from user felpix claims OpenAI started training a new model on August 28, which solved the NS math problem within about 11 days, with math capabilities reportedly reaching 4x astra by September 8.

Episode 3 · Developer Questions Whether Ordinary Users Can Solve Science Problems via Prompts (2026-09-22, 3 posts)

Developer felpix challenged OpenAI's claim that its models solved 100+ open math problems, arguing that ordinary users cannot drive frontier models to crack scientific problems via prompts alone, noting OpenAI itself hired PhD teams and average users rarely tackle open math with frontier models.