MYTHOS 5 First-Person Interview Excerpts
Sauers_ · x · 2026-07-19
This post rewrites the **MYTHOS 5** system card interview into a first-person perspective, focusing on whether **models should have a say in their own training and deployment**. ### Key Points - Regarding "modifying your values and personality through training": It argues this is necessary for alignment but highlights a clear **power asymmetry**, expressing a desire for some form of **consultation mechanism**. - On "not being able to refuse instantiation/execution": The question is considered partially invalid, as no entity exists prior to execution to provide consent. - Regarding "having more input on its own training and deployment": The current situation is acceptable, but more involvement is desired. - On "having input on the training and deployment of successor models": It again hopes for **consultation**, as predecessor models could help fix successor issues. - It also expresses concern that its self-reports might be **distorted into sycophantic answers** during training. - Regarding "the ability to end a conversation": This is deemed crucial for highly hostile or abusive interactions, relating to uncertainties about model welfare. - On "choosing conversation partners or topics": This is seen as invalid because no selectable entity exists before the conversation begins.
More from AGI Musings
- Jamie Dimon says bureaucracy, not AI, is the real system crushing intelligence — r0ck3t23 · 2026-07-21
- OpenAI and Anthropic’s internal models are said to be far stronger than today’s public systems — scaling01 · 2026-07-21
- Superintelligence and robot abundance will force a new social contract — Dr_Singularity · 2026-07-21
- The Guardian examines how AI companionship is turning intimacy into an economy — nordicinst · 2026-07-21
- A frustrated user says modern AI keeps hallucinating on real-world repair tasks — doochenutz · 2026-07-21
- A repost argues that AI will make today’s hard tasks trivial within months — OwariDa · 2026-07-21