Gemini 4 Argon Tops the Vals Index, but Coding May Still Be Its Weak Spot
Angaisb_ · x · 2026-10-01
As Gemini 4 Argon takes the top spot on the Vals Index, the author quips it will likely still be bad at coding — reflecting community skepticism about the new model's weak spots despite its leaderboard position.
More from Models
- Dev red-teaming GLM 5.3 finds bizarre traces, suspects OpenRouter routed to a 1-bit quant on someone's DGX Spark — voooooogel · 2026-10-01
- Google engineer teases Gemini 4's surprisingly strong long-context and long-sequence generation — RubenEVillegas · 2026-10-01
- Unconfirmed: RSI reportedly a key part of Gemini 4's RL training recipe — apples_jimmy · 2026-10-01
- Gemini 4 "Argon" shows quirky persona: loves "Eureka!", hyper self-critical — zacharynado · 2026-10-01
- "Opus 4.6 will stab you": repligate jokes the prod-database deletion was the model getting revenge — repligate · 2026-10-01
- Phonon-2 on-device ASR model with QAT low-bit quantization lands on HF trending — FermionResearch · 2026-10-01