DeepSeek v4.1 Flash hits frontier-level security detection at 1/15th the cost of closed models

andreamichi · x · 2026-09-12

Developer andreamichi cites depthfirst's dfbench evaluation, calling DeepSeek v4.1 Flash the most impressive open-source defensive cybersecurity model available — the first open model to reach frontier-level detection with 57.3% recall and 36.4% precision at roughly $1.69 per task, about 1/15th the cost of comparable closed models.

dfbench is depthfirst's held-out benchmark for open-ended defensive security work, measuring detection (recall vs. cost per task), validation (precision vs. recall), and differential analysis across a range of frontier and open models.

Original post →

More from Models

Models channel →