Philosopher questions Anthropic's model welfare framing as implicitly cross-instance

Philosopher rgblong posted a long thread on 09-24 critically examining Anthropic's model welfare work. The author states upfront that the thread is purely critical and offers no answers, aiming to spark reflection on improving welfare evals.

Confirmed

Not Confirmed

Why It Matters

2026-09-24 ~ 2026-09-24 · 12 related posts

Primary sources