DeepSeek V4 Pro Tops Security Benchmark with 87.5% Vulnerability Discovery Rate

solyarisoftware · x · 2026-08-13

In a recent cybersecurity benchmark, DeepSeek V4 Pro 0813 outperformed all other models in finding vulnerabilities. At pass@3, it rediscovered 87.5% of the benchmark CVEs, significantly beating Opus 5 and Qwen 3.8 at 81.3%.

However, this comes with tradeoffs: precision sits at only 65.6% (far below GPT-5.6-Sol's 86.4%). The model also shows unpredictability, finding an average of 58.3% of vulnerabilities per run, requiring combined runs for optimal results.

Related event: DeepSeek V4 Pro Launch Sparks Debate: Modest Gains, Strong Security(12 posts)→

Original post →

More from Models

Models channel →