EMBL-EBI's saezlab open-sources Karenina, a framework for multi-dimensional biomedical AI evaluation
anshulkundaje · x · 2026-10-02
saezlab (EMBL-EBI, Open Targets, Heidelberg University) released Karenina, an open-source framework for evaluating LLMs and agents, alongside a bioRxiv preprint titled "Turning Domain Expertise into Multi-Dimensional Evaluation of Biomedical AI with Karenina."
- Karenina turns domain expertise into multi-dimensional evaluation pipelines for biomedical AI, covering both LLMs and agents.
- Authors include Francesco Carli, Fabio Petroni, and Julio Saez-Rodriguez; demo video and code are publicly available.
More from Research
- Local Support Learning: 7B LLMs learn new tasks at full capacity without forgetting or old data — CatAstro_Piyush · 2026-10-02
- MIT researchers unveil interface exposing LLM internals during chatbot personality design — patpat_mit · 2026-10-02
- Fourth UK AI Conference Proceedings Now Live on PMLR as Volume 348 — lawrennd · 2026-10-02
- New paper: Training-time internal signals can improve alignment without hurting white-box monitoring — jonasgeiping · 2026-10-02
- Reka Open-Sources RIDM, Extracting Camera and Motor Commands from Raw Video — RekaAILabs · 2026-10-02
- YC Paper Club Explores AI Compute Beyond GPUs: Optical, Neuromorphic and Biological Computing — ycombinator · 2026-10-02