Mercor commits $5M to AI Safety Fund for alignment, evals and red-teaming research

typewriters · x · 2026-09-16

AI hiring company Mercor has committed $5M to a new AI Safety Fund supporting researchers working on frontier safety risks. Funded areas include misalignment (deceptive alignment, reward hacking, scheming), sandbox escape and agent containment failures, evaluation awareness, interpretability and scalable oversight, and red-teaming methodology.

The fund covers researcher hours, API credits, and travel, plus free access to its 5M+ expert network for red-teaming and annotation and its evals platform. Grants are open to independent researchers and non-profits.

Original post →

More from Safety

Safety channel →