Cisco's 3B Model Matches GPT-5.5 on Bug Localization (0.223) With Far Fewer False Positives

shashib · x · 2026-09-17

Cisco Foundation AI released the VLoc Bench (Sept 14), which asks models to search a full codebase and locate files containing known vulnerability types — not just label snippets. Key findings:

The author argues most public security benchmarks skip the hours-long search step that dominates real security work, and VLoc Bench's post-fix false-positive test better reflects practice.

Original post →

More from Safety

Safety channel →