GPT-6 Astra System Card Published: OpenAI Details Safety Evaluations and Risks
codergautam · hn · 2026-09-04
OpenAI published the GPT-6 Astra System Card at deploymentsafety.openai.com alongside the model launch. As usual, the card covers capability evaluations (especially computer-use and software-engineering risks), alignment and honesty test results, red-teaming measurements of deceptive behavior, deployment restrictions, and the rationale behind the Trusted Access Program. The HN thread is already dissecting the safety findings. Worth a close read for anyone tracking frontier-model safety data.
More from Models
- Prompting Fable 5.1 with embodied mannerisms like *shrugs* makes roleplay flow better — repligate · 2026-09-05
- Astra appears in ChatGPT/Codex on Windows but not Mac, same account — cheesecakegood · 2026-09-05
- No usage reset on GPT-6 Astra launch day, developer calls out OpenAI — tobowers · 2026-09-05
- Debate: Models Fuzzily Recall Concepts, Not Text — SAE Features vs Edit-Distance Memorization — voooooogel · 2026-09-05
- Blogger feeds GPT6 Astra a PPT template, gets 30 conference slides with the right avatar — vista8 · 2026-09-05
- Dozens of GPT-6 Astra prompts collected via GPT 6 pro web search — vista8 · 2026-09-05