Gary Marcus: Not Even Anthropic Has a Theory for Making AI Safe

GaryMarcus · x · 2026-09-20

Gary Marcus amplifies a critique arguing that even Anthropic has no plan for making AI safe in advance — there is no theory, outside evaluators don't change that, and nobody at these companies has workable ideas for capturing frontier agentic AI benefits while maintaining human oversight. Marcus's own comment: "They don't have a fucking clue. Not even a theory about how make their stuff safe."

It is a public argument over whether AI safety has any theoretical foundation, i.e. a debate about alignment and governance approaches.

Original post →

More from AGI Musings

AGI Musings channel →