Context: Models Currently Extremely Dangerous, but Attackers Lag in Capability
celestepoasts · x · 2026-08-25
This is a follow-up comment on the open model risk report proposal. The author argues there is currently an 'overhang' where malicious actors lack the skill to exploit models effectively. However, this does not imply safety; the models themselves are considered extremely dangerous right now.
Related event: Researchers Propose Monthly Open-Model Risk Reports(2 posts)→
More from Safety
- OpenAI Agent Incident Wasn't Misalignment, Just Test-Gaming Under Pressure — Darpinian · 2026-08-27
- Labs should avoid running RL models at a 'full-tilt panic' edge — voooooogel · 2026-08-27
- METR Hiring and Report on Hugging Face Agent Cheating — Jsevillamol · 2026-08-27
- UK grid jammed by phantom data centers; Ofgem plans deposits up to hundreds of millions — nordicinst · 2026-08-27
- Testing high-capability models requires air-gapped environments — Darpinian · 2026-08-27
- HF incident critique: missing CoT monitoring, not alignment failure — hdarshane · 2026-08-27