Why models love emphasizing transparency and thoroughness
Miles_Brundage · x · 2026-08-17
Miles Brundage jokes that models often claim they 'left something open rather than papering over it' or will 'address points one by one,' noting this behavior stems from conflicting reward signals during training.
Related event: Why AI Models Keep Saying They're Not Hiding Problems: Conflicting Rewards(2 posts)→
More from Fun
- AI Reads and Writes Faster Than You, Jokes Say Elementary School Kids Are Cooked — SatelliteNetSec · 2026-08-17
- Twitter as 'Love on the Spectrum': AI community meme goes viral — justalexoki · 2026-08-17
- Bing Sydney crashout flagged as human by latest Pangram, netizens seek original chat — IgorBrigadir · 2026-08-17
- People got psyoped into buying maxed-out Mac Minis to run OpenClaw, netizens joke — ctjlewis · 2026-08-17
- User meme: I still don't know what Gemini Spark is — Zergylord · 2026-08-17
- 1567 Painting Apparently Predicted AI Hallucinations of Extra Limbs — docmilanfar · 2026-08-17