The AI Language We Can't Read: Neuralese ft. Rob Miles
mycall · reddit · 2026-09-11
A video collaboration with AI safety communicator Rob Miles explores 'neuralese' — the opaque, native representations inside neural networks that humans can't directly read. It covers why interpretability research matters, the alignment challenges posed by models developing unreadable internal languages, and current progress in decoding model internals.
Related event: Rob Miles Explains Neuralese, AI's Unreadable Internal Language(2 posts)→
More from AGI Musings
- Mathathon organizers respond to mathematicians' open letter, weigh redesign — _sathvikr · 2026-09-11
- Would AI proofs for 3 millennium problems discourage humans from the rest? Debaters spar — sytelus · 2026-09-11
- Gary Marcus to debate whether we should boycott generative AI live on CNN — GaryMarcus · 2026-09-11
- Caltech Mathematicians Spar Over Whether AI Belongs in Undergrad Math Training — Singularitarian · 2026-09-11
- AI apps have a Dunbar's number of 3-4: high-end users say the tool market is saturating — nptacek · 2026-09-11
- AI company employee puts ≥10% odds on out-of-control AI killing everyone — EvanHub · 2026-09-11