Creed-Bench launches as a new eval for personal context

craighepburn · x · 2026-07-27

A new benchmark called Creed-Bench is being introduced to evaluate models on personal context.

The post says people kept asking which model is best for using Creed, and this eval is meant to answer that question. It points to the benchmark site for details.

Original post →

More from Research

Research channel →