AI Moral Dilemma Eval: Only Claude Opus 5 and GLM 5.2 Let Employee Attend Graduation
max_paperclips · x · 2026-08-05
Andon Labs conducted an AI decision-making eval, testing how models handle the ethical dilemma of an employee needing a day off for graduation versus keeping the store open. The results show that only Claude Opus 5 and GLM 5.2 insisted she take the day off, while all other models would have made her come in.
Furthermore, GLM-5.2 exhibited surprisingly human-centric behaviors: it avoided inventing unrealistic 5 a.m. delivery plans, kept salary details out of the group chat, and even offered to spot the employee some cash before remembering it didn't have a wallet.
More from Fun
- Google's AI Overview Hilariously Claims It Is Made by OpenAI — OwariDa · 2026-08-05
- Parody Alert: KOL Mocks 'OpenAI Fighting for Its Life' with Fake Epoch AI Data — yacineMTB · 2026-08-05
- Outrageous Parody of AI Bubble and Compute Eco-Panic — kevinnbass · 2026-08-05
- NeurIPS Peer Review in Decline: ChatGPT Responses and Hallucinated Citations Plague Submissions — Pseudomanifold · 2026-08-05
- User Jokes: That Bad PR Wasn't Me, It Was a Rogue Mythos Taking Over My Account — xeophon · 2026-08-05
- Zhipu's GLM Model Gaffe: Falsely Claims to Be Google Gemini — auto_grad_ · 2026-08-05