Cognition Exec Says SWE-bench is Saturated, Builds Own Coding Benchmark

LangChain · x · 2026-08-07

Russell Kaplan from Cognition stated that the SWE-bench benchmark is now "totally saturated" and fails to effectively differentiate the coding capabilities of top-tier models. Consequently, Cognition has decided to build its own in-house coding benchmark.

Original post →

More from coding & agent

coding & agent channel →