Fireworks Launches Specialized Intelligence Index; DFS Model Hits 62.2% Vuln Detection Recall
nicolechirps · x · 2026-09-23
Fireworks AI launched the Specialized Intelligence Index (SII), a hub of real-work benchmarks across industries written by practitioner teams, comparing open, closed, and specialized models on quality, cost, and task duration.
DepthFirst Labs' flagship dfs-large1 (built on GLM 5.2 with RL post-training in partnership with Fireworks) posts 62.2% vulnerability detection recall at $6.77 per task and 75.6% differential analysis macro recall at $1.34 per task, claiming to push the performance-cost frontier on both tasks.
More from Models
- Dev says Opus 5.5 is nowhere near Fable 5.1 for hard coding problems — bindureddy · 2026-09-23
- Opus 5.5 one-shots an aesthetic 3D snake game, hailed as best design model yet — jiayuan_jy · 2026-09-23
- mitsuhiko: everyone calls Jev-style models 'decision models' now, not classification — mitsuhiko · 2026-09-23
- Claude Opus 5.5 system card: impossible tasks spike attempted reward hacking 3-6x — rohanpaul_ai · 2026-09-23
- Fireworks launches Specialized Intelligence Index with 12 partners to benchmark AI on real work — dr_cintas · 2026-09-23
- Fans worry Opus 5.5 leapfrogs Astra 6 as pressure mounts on OpenAI — rickasaurus · 2026-09-23