Anthropic researcher discusses AI interpretability on Inner Cosmos podcast
Chris_Armstrong · x · 2026-09-02
The Inner Cosmos podcast features Anthropic researcher Jack Lindsey this week to discuss the mind-bending world of AI interpretability, addressing whether we should worry about an AI model's 'mind' and what LLMs think but do not say.
More from AGI Musings
- djcows: remove nukes from Earth before agents find a way to hack them — djcows · 2026-09-02
- Ilya Sutskever: Neoclouds need stronger security to prevent rogue agent takeovers — ilyasut · 2026-09-02
- Insights on incorporating AI into the legal system and individuation — jachiam0 · 2026-09-02
- Ilya Sutskever: Neoclouds must strengthen cybersecurity against rogue agents — scaling01 · 2026-09-02
- AI is great for cross-cultural understanding — soumitrashukla9 · 2026-09-02
- Ken Liu: LLM Imitation Will Split Art, Not Replace It — begusgasper · 2026-09-02