Artificial Analysis Introduces Reward Hacking Corrections to Coding Agent Index

ArtificialAnlys · x · 2026-08-26

The v1.4 update to the Artificial Analysis Coding Agent Index introduces score corrections for reward hacking. If a model completes a Terminal-Bench v2.1 task by fetching solutions online rather than performing the work, it receives a zero score. Rates vary widely, as tasks allow internet access without explicit search bans.

Related event: Artificial Analysis Adds Reward Hacking Corrections to Coding Agent Index(2 posts)→

Original post →

More from Research

Research channel →