TransluceAI 提出监督基础模型,专抓 reward hacking

JacobSteinhardt · x · 2026-07-29

原文链接 →

「安全」频道最新

更多「安全」频道 AI 资讯 →