Researchers Push Back on Treating Claude's Outputs as Independent Evidence
AI safety researcher rgblong publicly disputed how MacAskill and Caviola used Claude's model behavior as evidence, arguing its outputs cannot serve as independent evidence since the model can be trained to give any answer; the discussion also revisited risks of AI 'super persuasion'.
2026-09-19 ~ 2026-09-20 · 2 related posts