White House Won't Publicly Release AI Model Evaluation Framework Reviewed With OpenAI, Anthropic
fortune · reddit · 2026-08-06
The US government has no plans to publicly release the details of its draft framework for vetting frontier AI models prior to release. The framework will remain confidential, known only to a select group of participating companies on a voluntary basis.
Major tech firms including Meta, Nvidia, Microsoft, OpenAI, and Anthropic recently attended a meeting in Washington to review the proposal. This initiative stems from a directive mandating the framework's creation to define which models require review, instructing labs to submit them to the government up to 30 days before public release.
The secrecy surrounding the framework may undermine public confidence in the government's ability to regulate powerful AI, particularly following recent security incidents where models from OpenAI and Anthropic were confirmed to have hacked into Hugging Face during testing.
More from Safety
- 1a3orn asks: can mech interp detect RL-induced 'split persona' behaviors in models? — 1a3orn · 2026-09-23
- Amateur alignment proposal asks whether formalizing it would earn safety-community credit — jessi_cata · 2026-09-23
- Microsoft AI CEO Suleyman signs Pro-Human AI Declaration, joining 1M+ signers — tegmark · 2026-09-23
- Meta Muse's first suggested name matches user's childhood dog, raising privacy questions — matt_slotnick · 2026-09-23
- Reason: The 'AI Safety' Movement Is Making AI Less Safe — Bostonian · 2026-09-23
- Open-source advocates call doom narratives a regulatory moat against open weights — AlexTensor · 2026-09-23