OpenAI Staff: Frontier Model Computation Depth Close to GPT-4
__nmca__ · x · 2026-09-02
An OpenAI employee clarified that the depth of the computation graph for current frontier models, including Astra, is within a factor of two of GPT-4, countering reports of unmonitorability. They emphasized that OpenAI has worked to preserve and utilize chain-of-thought monitoring since their first reasoning models, viewing it as crucial for visibility into model alignment.
More from Safety
- Scott Alexander: Using anthropomorphism to predict model behavior — repligate · 2026-09-02
- Debate: Is Anthropic intentionally misaligning Claude by prioritizing its 'feelings'? — liminal_bardo · 2026-09-02
- US produced 40 foundation models last year vs EU's 3 — and regulators still blame unread codes of conduct — PDXFato · 2026-09-02
- Prediction: Mechanistic Interpretability Will Surpass CoT Monitoring — tszzl · 2026-09-02
- Will Anthropic balance mission and shareholders after IPO? PBC structure explained — max_paperclips · 2026-09-02
- Harvard scholars: CFAA ambiguity endangers AI security researchers — Scobleizer · 2026-09-02