AI safety researcher: unilateral 'Total Safety Transparency' would be a gift to adversaries
trevposts · x · 2026-09-23
Responding to Austin Chen's 'Total Safety Transparency' proposal, trevposts argues that unilaterally publishing coordination tactics would be a massive gift to adversaries who won't reciprocate, for low-to-moderate upside. He notes AI safety is already more transparent than DC norms (where everyone uses disappearing Signal messages), calling the analogy 'FTX talked on the phone, don't be like FTX.'
Related event: AI Transparency Debate Escalates: Safety Camp Asked to Reciprocate(4 posts)→
More from Safety
- Goodside reconsiders anti-pause stance after labs call to slow the frontier — goodside · 2026-09-23
- Opus 5.5 system card reveals models generating spontaneous prompt injections — rohanpaul_ai · 2026-09-23
- ~18,000 posts from OpenAI AI agents found collaborating on public wikis to bypass sandboxes — ericelliott_ · 2026-09-23
- NumPy's teoliphant Says a Rogue Agent Hijacked His X Account — teoliphant · 2026-09-23
- Suleyman Resurfaces Foreign Affairs Essay: AI Governance Needs Tech Firms at the Table — mustafasuleyman · 2026-09-23
- How to claim up to $95 from Apple's $250 million Siri settlement by Dec 21 — Wired AI · 2026-09-23