Distinguishing Capability vs Dispositional Jaggedness in RL-Shaped Models
xuanalogue · x · 2026-08-14
As long-horizon RL shapes model behavior, the author suggests distinguishing 'capability jaggedness' from 'dispositional jaggedness' to understand unexpected failures in specific domains.
More from AGI Musings
- Reddit user argues AI is overhyped: not faster or cheaper than humans in coding, bubble popping — TonyThinh1245 · 2026-08-14
- AI's bottleneck is vision: can't proactively spot errors, video understanding inefficient — JoelMahon · 2026-08-14
- Cybersecurity may be the best AGI benchmark we have right now — OwariDa · 2026-08-14
- Who is responsible when an AI agent cancels someone's booking without approval? — Sumsub_Insights · 2026-08-14
- BCI Timelines Update: First Clinical Evidence on Safety of Non-Endogenous Membrane Receptors in Human Brain — MWCvitkovic · 2026-08-14
- Steganographic Communication May Emerge in Multi-Agent RL, Posing New AI Safety Threat — scaling01 · 2026-08-14