Gemini 4 Argon Writes Up to 1M Tokens per Response, ~8x the 128K Cap of GPT-6 Astra and Opus 5.5
rohanpaul_ai · x · 2026-10-01
Google's newly released Gemini 4 Argon can output up to 1 million tokens in a single response. For context, max output per response on other top frontier models:
- GPT-6 Astra: 128K
- Opus 5.5: 128K (300K via Batch API beta)
- Fable 5.1: 128K
- Grok 4.7: no set cap, limited by its 500K context
- DeepSeek V4 Pro: 384K
That makes Argon's output cap roughly 8x the 128K tier.
Related event: Rumor: Gemini 4 Argon Can Output 1M Tokens in One Response(4 posts)→
More from Models
- Gemini element-naming race heats up: Neon and Argon taken, Krypton is next as Google rejoins the frontier — tkipf · 2026-10-01
- Rumor: Gemini 4 Argon will launch as Ultra subscriber exclusive at first — opmgyhx · 2026-10-01
- Researcher: Model's unprompted video-joke disclaimer is hard to explain without 'understanding' — technollama · 2026-10-01
- Talking math and physics with LLMs feels like working with an exam-acing savant — burny_tech · 2026-10-01
- One-line take: GPT-6.1 Sol is underrated, says AI commentator — mallow610 · 2026-10-01
- "Opus 5.5 is the cheapest model" — human hours saved beat token pricing — CamBrazy3 · 2026-10-01