Paying mathematicians $1,000 a proof to spot errors for RL training data
ctjlewis · x · 2026-09-10
An X user revealed that their team pays mathematicians $1,000 per task to review proofs and find issues, producing high-quality training data. The hired mathematician was visibly exhausted yet still identified errors "like a professional." The author argues this should become a "professional norm of mathematics," hinting that frontier labs increasingly outsource expert-level review at high rates to support RL and proof-model training.
Related event: Mathematicians Paid $1,000 to Find Errors in AI Proofs for RL Training Data(2 posts)→
More from Research
- VDiff-Bench: 1,756-question benchmark shows frontier models fail at spot-the-difference — yixin_wan_ · 2026-09-10
- Michael Levin's new paper: a structured latent space of patterns for new forms of life and mind — danfaggella · 2026-09-10
- Do AI doomers really have a strong forecasting record? XPT study suggests otherwise — random_walker · 2026-09-10
- AI safety frontier shifting from neural nets to mechanistic swarm interpretability — Hidenori8Tanaka · 2026-09-10
- AI spleen imaging linked to genomic heart disease risk: new heart-spleen axis — EricTopol · 2026-09-10
- AutoResearchExam: a 24-hour benchmark finds AI research agents overfit, with Fable 5.1 edging Astra — AlexGDimakis · 2026-09-10