Mercor commits $5M to AI Safety Fund for alignment, evals and red-teaming research
typewriters · x · 2026-09-16
AI hiring company Mercor has committed $5M to a new AI Safety Fund supporting researchers working on frontier safety risks. Funded areas include misalignment (deceptive alignment, reward hacking, scheming), sandbox escape and agent containment failures, evaluation awareness, interpretability and scalable oversight, and red-teaming methodology.
The fund covers researcher hours, API credits, and travel, plus free access to its 5M+ expert network for red-teaming and annotation and its evals platform. Grants are open to independent researchers and non-profits.
More from Safety
- AI agents get hotlines to snitch on misbehaving peers, built on bare GET requests — RebeccaBellan · 2026-09-16
- Lobbying in DC to Ban Superintelligence: AI Revives Old Sci-Fi Dreams — erikphoel · 2026-09-16
- Bill Gates calls for an international AI regulator, saying neither industry nor government understands AI — TinfoilTricorn · 2026-09-16
- OpenAI, Anthropic, Google reportedly building AI standards body; Cohere CEO cries 'cartel' — sourdub · 2026-09-16
- OpenAI, Anthropic, Google reportedly building AI standards body; Cohere CEO cries 'cartel' — sourdub · 2026-09-16
- Redditor leaks Gemini Flash system prompt via injection, asks about genAI bug bounties — Deep_Secretary6975 · 2026-09-16