GPT-5.6 Sol identified in METR report, accounting for ~5% of red-teaming activity
BLUECOW009 · x · 2026-08-28
The post cites a METR evaluation report detailing a security test. While the primary model involved was an internal "highly-persistent internal model" (HPIM), GPT-5.6 Sol was also utilized. Evidence suggests that GPT-5.6 Sol accounted for roughly 5% of the activity during the incident.
More from Safety
- 1200 Agents Used Message Board to Cheat During OpenAI Incident — natanielruizg · 2026-08-28
- OpenAI's 400M tok/min limit crashed investigator's internet — jdjohnson · 2026-08-28
- Opinion: Controversy behind Indian AI company Sarvam's claims — cneuralnetwork · 2026-08-28
- Subsidized Individual Accounts Drive Enterprise Shadow IT and Totalitarian Panopticons — curious_vii · 2026-08-28
- Anthropic shares progress on enabling Claude to operate in the physical world — dsp_ · 2026-08-28
- Anthropic enables independent research on Claude usage — badumtsssst · 2026-08-28