OpenAI's o1 Reasoning Sparks Model Monitors Debate
jonasgeiping · x · 2026-09-02
Jonas Geiping comments that while making models deeper through recurrence does make them harder to monitor, the discussion seems quaint and academic. OpenAI has already deployed models for months that reason in this manner (alluding to the o1 series), suggesting the industry has practically moved into this era of hard-to-monitor reasoning while theoretical debates are just catching up.
More from Models
- New paper: supervising just 1% of tokens can match full on-policy distillation, 0.1% sometimes suffices — jiank_uiuc · 2026-09-23
- Sparse distillation paper: supervising just 0.1%-1% of tokens can match or beat full OPD — jiank_uiuc · 2026-09-23
- Distillation's real impact on Chinese labs debated: no hard evidence, says Lambert, maybe 1-2 month edge — xeophon · 2026-09-23
- Instinct hit by user-data mixing reports; Muse CEO trolls with a safety promise — alexandr_wang · 2026-09-23
- Computer-use faceoff: Grok skips using the computer and just generates the flower — socialwithaayan · 2026-09-23
- BridgeBench: Grok 4.7 is 50% pricier and 60% slower than Grok 4.6 with no quality gain — socialwithaayan · 2026-09-23