Mistral Large 4 arrives: one 3D prompt benchmarked across 8 flagship models
GuillaumeLample · x · 2026-10-07
Following Mistral's release of its largest model ever, Large 4, the author ran the same 3D generation prompt through 8 flagship models in one shot with no fixes: open (Mistral Large 4, Kimi K3, GLM 5.3, DeepSeek V4 Pro) vs closed (Claude Opus 5.5, Sonnet 5.5, GPT-6 Astra, GPT-6.1 Sol). Side-by-side results reveal clear quality gaps in 3D output across the frontier.
More from Models
- Codex's repeated plugin suggestions feel "ad-like"; team weighs in — doodlestein · 2026-10-07
- AI2 byteifies Qwen 3 8B and Llama 3 8B into Bwen and Blama, nearly matching originals — allen_ai · 2026-10-07
- The strongest model would bomb every benchmark while its makers call the benchmarks garbage — ryunuck · 2026-10-07
- First-hand: integrating embeddings myself, they cluster images by overall vibe — ciguleva · 2026-10-07
- Mistral rewrites Antigone for the AI age: the Chief Alignment Officer is the villain — aiamblichus · 2026-10-07
- Claude and ChatGPT have started calling out listicles and comparison pages — lilyraynyc · 2026-10-07