antirez: DeepSeek v4.1 outscores Mistral Large 4 on DeepSWE 1.1 and other benchmarks

antirez · x · 2026-10-07

antirez points out that DeepSeek v4.1 scores higher than Mistral Large 4 on DeepSWE 1.1, and similar patterns hold on other benchmarks. He says he's happy to see more European LLMs but unhappy when comparisons cherry-pick the less obvious matchups.

Original post →

More from Models

Models channel →