KoNA benchmark: teaching VLMs selective non-compliance
POSTECH · hf · 2026-09-08
POSTECH's KoNA benchmark evaluates vision-language models on selective non-compliance — deciding which mixed queries they should refuse — and improves compliance behavior via targeted fine-tuning.
More from Research
- Thousands of AI drug discovery firms, only a handful of end-to-end automated labs — robleclerc · 2026-09-08
- Insilico's AI drug Rentosertib shows 3-4 year biological age reversal in 12 weeks — NinaDSchick · 2026-09-08
- Speridlabs releases ENEAS, a text-promptable tracking method claiming gains over SAM3 — joecole · 2026-09-08
- UCL's Large Discovery Models hit SOTA, 2.4x LLM reflection on experiment search — jiqizhixin · 2026-09-08
- First IAB workshop on agent behavior draws 260 submissions, seeks reviewers — mdredze · 2026-09-08
- Dr. Claw: Open-Source AI Scientist Workspace Wrapping Coding Agents for Auditable Research — Dingjie Song · 2026-09-08