Anthropic CEO proposes ASI as a model committee with a separately trained ethicist model

robleclerc · x · 2026-09-16

Rob LeClerc proposes that ASI need not be a single model or MoE — it could be a committee of models including a separately trained conscience/ethicist model that sets the range of permissible directions and checks outputs for consistency. He invokes Alasdair MacIntyre's "first principles are argued from, not to" to explain how such an ethics model would ground the reasoning models.

Original post →

More from AGI Musings

AGI Musings channel →