Vision Test Fail: GPT Lags Behind Opus and Gemini in Low-Quality Image Recognition
stevejuniormc · reddit · 2026-08-09
A user tested mainstream large multimodal models on their ability to identify a bird from a low-quality image.
Results
- Claude 3.5 Opus, Fable 5, and Gemini 3.6 Flash all successfully identified the bird.
- GPT 5.6 sol was completely off target, highlighting potential unreliability in specific visual tasks.
More from Models
- Anthropic Adjusts Claude Guardrails: New Deployment Eases Sensitive Queries — xuanalogue · 2026-08-10
- Moonshot Defies Odds: Building Frontier Model Kimi K3 with ~500 Staff — OwainEvans_UK · 2026-08-10
- Mistral Launches Shieldstral: A 3B Parameter Open-Source Safety Classifier — dl_weekly · 2026-08-10
- OpenAI Models Hit 93% Hallucination Rates, Challenging AI Unit Economics — gerardsans · 2026-08-10
- Anthropic's Strategy: Focuses on B2B Coding Models, Skips Image Generation — sahilypatel · 2026-08-10
- 20VC Founder Says Qwen and Kimi Beat ChatGPT and Gemini for Research — hsu_byron · 2026-08-10