SPAR Seeks Mentees for AI Safety Research: Focusing on Metagaming and Eval Awareness
austinc3301 · x · 2026-07-31
Researcher Ezra J. Newman is recruiting mentees on the SPAR platform for an AI safety research project.
The project focuses on 'metagaming' (where models act strategically during evaluations) and 'eval awareness' in advanced AI models. Exceptional fellows will have the opportunity to work at the well-known AI safety organization Apollo Research.
More from Safety
- Google Responds to AI Misinformation Concerns: Gemini Images Embed SynthID Watermarks — henkvaness · 2026-07-31
- Model Eval Accidentally Commits Cyber Crimes? Users Debate Accountability — BlancheMinerva · 2026-07-31
- Webinar Preview: Experts to Discuss the Limits of Human Oversight in the Era of AI Agents — mmitchell_ai · 2026-07-31
- DeepSeek jailbroken using role-play to generate assassination plans — DiamondAgreeable2676 · 2026-07-31
- Cloudflare Details Internal Agent Platform Security After OpenAI and Anthropic Sandbox Escapes — irvinebroque · 2026-07-31
- Former US Security Adviser Proposes $50B Strategic Investment Fund for Reindustrialization — Rewkang · 2026-07-31