Frontier Model Council Study: Models Fail Together

soumitrashukla9 · x · 2026-07-14

The authors re-ran frontier model experiments using the same initial benchmarks, including GPT 5.5, Gemini 3.1 Pro, Claude Opus 4.8, Grok 4.5. They developed an "error correlation" metric to measure the probability of two models answering incorrectly simultaneously.

They found:

The author believes this supports the "consensus machines" argument: sharing training data, distillation targets, and RLHF habits causes models to converge on the same wrong conclusions. The text also notes that 6 pairs of vendors exhibited similar phenomena on MMLU-Pro Math.

Original post →

More from Research

Research channel →