Sam Altman: GPT-6 Astra Hit OpenAI's 'Cyber Critical' Threshold, Forcing New Safeguards
didiTonic · reddit · 2026-09-04
In a Bloomberg interview, Sam Altman said GPT-6 Astra crossed OpenAI's internal 'cyber critical' capability threshold, forcing the team to add new safeguards before release.
Key points:
- Pressed on AI autonomously finding zero-day exploits, Altman clarified the model paused over that issue was a future model, not Astra itself
- He said future models will become more autonomous, which is why OpenAI is investing heavily in monitoring, sandboxing, and alignment
- His core argument: the real risk isn't just smarter AI, but AI that can increasingly act and work on its own
More from Models
- Benchmarks Are Dead: Evaluating AI Judgment, Recovery and Real Usefulness — Div_pradeep · 2026-09-04
- Team finetunes Gemma 12B for audio proofreading, benchmarks it against Gemini — ojasvi_yadav · 2026-09-04
- First impressions of Astra: clean tone and zero jargon in its writing — soumitrashukla9 · 2026-09-04
- Reddit buzz: OpenAI "GPT-6 Astra" demo called "Jarvis in the making" — WarGod1842 · 2026-09-04
- Nadella says early customers already use Astra on Azure as Altman responds — i_dg23 · 2026-09-04
- Abu Dhabi institute IFM releases 6 fully open-source AI models with data, code & methods — Polymarket · 2026-09-04