Grok 4.7 and Claude Opus 5.5 both score zero on CancerBench

iScienceLuvr · x · 2026-09-29

iScienceLuvr added the two newly released frontier models — Grok 4.7 and Claude Opus 5.5 — to his CancerBench benchmark, and they tied for first and last place with a score of zero. He stresses he isn't bearish on AI curing cancer; quite the opposite, he thinks far too little effort is going into applying frontier AI to cancer and hopes these posts draw attention so labs start saturating the benchmark soon.

Original post →

More from Fun

Fun channel →