OpenAI staff point readers to GPT-6 Astra system card's alignment evals
kaicathyc · x · 2026-09-04
An OpenAI deployment safety team member recommends reading the GPT-6 Astra system card beyond the launch blog: it shares extensive evals and analysis on the model's alignment, and discusses continued challenges on both alignment and monitoring fronts.
Related event: GPT-6 Astra System Card Flags Major Drop in Monitorability(37 posts)→
More from Models
- GPT-6 Astra called 'rough launch but feels better than benchmarks' — community claps back — ns123abc · 2026-09-04
- OpenAI to give paid ChatGPT users banked resets for every day without Astra access — Angaisb_ · 2026-09-04
- Gary Marcus: GPT-6 Astra's ARC-AGI-3 success supports his symbolic world model hypothesis — GaryMarcus · 2026-09-04
- All AI benchmarks are 'broken or saturated', new long-running loop eval launching next week — bindureddy · 2026-09-04
- GPT-6 Astra Tops Terminal-Bench-Science, Dethroning Fable 5.1 at 52.6% — burny_tech · 2026-09-04
- Databricks: GPT-6 Astra hits SOTA on OfficeQA Pro and 2 more benchmarks — Hesamation · 2026-09-04