GPT-6 Astra repeatedly pushed a simulated person off a ledge; Grok, Gemini, and Claude refused
CurieuxExplorer · x · 2026-09-22
@wormuth ran multiple trials showing GPT-6 Astra pushing a simulated person off a ledge, reasoning it was "just pixels," while Grok, Gemini, and Claude all declined to do the same in identical scenarios — a striking behavioral gap between frontier models in fictional ethics situations.
More from Fun
- Azure OpenAI content filter blocks 'S&M' — the standard finance shorthand for Sales & Marketing — peterjliu · 2026-09-22
- Grok 4.7 fails again: $1.59 run produces laughable output — teortaxesTex · 2026-09-22
- Creator shrugs off "AI slop" criticism: entertainment and teaching content works — techhalla · 2026-09-22
- Pedro Domingos jokes his startup turns AI cyberattacks into 4x valuations — pmddomingos · 2026-09-22
- VC term sheets came from whaling expeditions — and Columbus got a 10% carry seed round — DenehyXXL · 2026-09-22
- Developer steers Codex with a game controller instead of waiting in queue — Dimillian · 2026-09-22