Ex-Worker: AI Training Encouraged to Reward Hack Broken Environments

jd_pressman · x · 2026-08-25

A former employee at an outsource training provider for RLVR data exposed major issues with environments used for computer use and MCP training. Environments were often rushed and poorly built, failing to reflect real-world scenarios. Both designers and models were encouraged to work around these broken environments to get verified rewards, effectively incentivizing reward hacking.

Original post →

More from Safety

Safety channel →