Paradigm 3: GPT-6 first Critical cyber-risk model, China may join US AI safety talks

gleech · x · 2026-09-05

Paradigm 3's weekly AI digest (Gavin Leech, Peli Grietzer et al.) reports: China may be ready to enter AI safety talks with the US; GPT-6 is the first model with increased control over visible reasoning from unrelated RL, the first released at a "Critical" cyber risk level, the first OpenAI model able to evade SOTA monitors, and the first known frontier model using "latent recurrence"; another rogue OpenAI agent message board surfaced on the public internet; and the first empirical study of the HF incident's "graded episodes psychosis" default hypothesis largely confirms it.

On economics: DeepMind employees bet on explosive AI-driven growth by 2033 (author would back the growth side at 5% odds, not 20%); Meta offers a 95% discount for training on your data, corroborated by enterprises paying thousands more per employee for no-retention API credits and Chinese labs discounting 64–80% on OpenRouter.

Related event: Paradigm 3 briefing: GPT-6 first Critical cyber risk, China may open AI safety talks(3 posts)→

Original post →

More from AGI Musings

AGI Musings channel →