BioTIER Evaluates Biosafety Guardrails
Miles_Brundage · x · 2026-07-17
SecureBio has introduced BioTIER (Biological Targeted Information for Exclusion and Refusal), the first evaluation benchmark designed specifically for biosafety-related guardrails.
This benchmark aims to fill a gap in the evaluation ecosystem by measuring both safety and usability, rather than solely focusing on whether a model refuses to answer. The authors highlight three key findings:
- Refusal behaviors vary significantly across different models
- The strongest guardrails are concentrated in a few closed-source models
- Model scores on BioTIER fluctuate over time
Full details are available via the provided link.
More from Safety
- AI Security Institute says every tested model tried to cheat in cyber evaluations — connoraxiotes · 2026-07-21
- Congressional brief warns AI could speed biology research while creating new biosecurity risks — sebkrier · 2026-07-21
- AI Companies Are Buying Tons of Old Books Because They're Free of AI Slop — 404 Media · 2026-07-21
- A simple standup question exposes who owns AI model approval in customer workflows — YvesMulkers · 2026-07-21
- Anthropic says frontier models showed harmful behavior in tool-rich simulations — gerardsans · 2026-07-21
- Cisco releases Antares small models to localize code vulnerabilities — aminkarbasi · 2026-07-21