xAI's Risk Framework Mentions Internal Capability Surveys, But Outcomes Remain Unclear
dfrsrchtwts · x · 2026-07-23
A discussion highlights that xAI's risk management framework mentions potentially publishing survey results from employees regarding their views and projections on important future AI developments, such as capability gains and benchmark results. The poster inquires whether this has actually happened, noting that the closest comparable industry practice so far is Anthropic publishing surveys on how their models perform on AI R&D tasks.
More from Safety
- Orbit v0 launches as a framework for multi-agent safety and security evals — ghadfield · 2026-07-23
- Politico says OpenAI models launched a cyberattack, prompting Congress to act — Distinct-Question-16 · 2026-07-23
- Agent-era security needs customer keys, proof-of-presence, and hardware-backed identity — dhadfieldmenell · 2026-07-23
- OpenAI reportedly warned its training approach could trigger a breakaway hacking incident — ShakeelHashim · 2026-07-23
- Ptacek says a 2025 open-weight model could already break sandboxes and scan networks — Simon Willison · 2026-07-23
- Small AI safety team says it helped pass three state laws and is now hiring — Miles_Brundage · 2026-07-23