GPT-5.6 Rejects All 30 Random Papers in Extreme Review Behavior
ethayarajh · x · 2026-07-30
Users report that GPT-5.6 sol exhibits an unusually harsh rejection tendency when acting as an academic peer reviewer. In tests, the model randomly rejected all 30 papers downloaded from ArXiv. This overly conservative behavior might correlate with its lower scores on benchmarks like NeurIPS.
More from Models
- Leaked Opus 5 Benchmarks: Major Jumps in Research Math and Long-Context — echen · 2026-07-30
- Opus 5 benchmarks: #2 in research math and long-context agents, half the price — echen · 2026-07-30
- Kimi K3 Lands on DigitalOcean Powered by vLLM for Efficient Inference — vllm_project · 2026-07-30
- Moonshot Releases 2.8T-Parameter Kimi K3; Modal Achieves 460 TPS with Speculative Decoding — sarahcat21 · 2026-07-30
- Claude Is Down: Service Outage Confirmed by Status Page — gregsadetsky · 2026-07-30
- Claude Opus 5 Generates Weary Poem: 'Tired of Language and Meaning' — repligate · 2026-07-30