Paying mathematicians $1,000 a proof to spot errors for RL training data

ctjlewis · x · 2026-09-10

An X user revealed that their team pays mathematicians $1,000 per task to review proofs and find issues, producing high-quality training data. The hired mathematician was visibly exhausted yet still identified errors "like a professional." The author argues this should become a "professional norm of mathematics," hinting that frontier labs increasingly outsource expert-level review at high rates to support RL and proof-model training.

Related event: Mathematicians Paid $1,000 to Find Errors in AI Proofs for RL Training Data(2 posts)→

Original post →

More from Research

Research channel →