Doubts cast on low ranking of Gemini 3.7 Flash
scaling01 · x · 2026-08-26
A user questions the low ranking of Gemini 3.7 Flash on leaderboards, suggesting it is suspicious. They speculate that Google's reward hacking monitor might be ineffective.
More from Models
- Nvidia may have funded 100T tokens for free GLM 5.3 Flash release — bindureddy · 2026-08-26
- Flaw in anti-finetuning: Cost > Quality once models are saturated — rhythmrg · 2026-08-26
- sanoTTS: 1.4M-Param Model Runs Real-Time on $3 Chip — kastnerkyle · 2026-08-26
- New Models to Know: MoE-ViE, τ0-VLA, 4DAnyone, and More — TheTuringPost · 2026-08-26
- User Finds Sol Max More Reliable Than Sol Ultra for Complex Tasks — imjustnewatai · 2026-08-26
- GPT Auto-Titles Conversation in Chinese, Baffling User — rodrigoinfloripa · 2026-08-26