Users report GPT-6 Astra tightens safety refusals, becoming more Anthropic-like
TheStalwart · x · 2026-09-15
TheStalwart notes GPT-6 Astra is going "next levels" with personal safety, sharing refusal screenshots.
In replies, ivanbezdomny adds that GPT was previously very permissive with almost no refusals, but since Astra it has become noticeably more Anthropic-like in its guardrails — though "a bit dumber" at it since OpenAI is newer to this style of moderation. The thread captures community perception that the new model ships with much more conservative safety behavior.
Related event: Users report GPT-6 Astra refuses more, resembling Claude's style(2 posts)→
More from Models
- ZDTaichu5.0-9B, a 9B vision-language model with spatial reasoning, trends on Hugging Face — TaichuAI · 2026-09-15
- Atria Dawn Preview: student-heavy team launches research-focused agentic base model — xiaohu · 2026-09-15
- Which 10Eros video-model quant works best on 8GB VRAM? A practical trade-off question — apostrophefee · 2026-09-15
- rasbt shows why final-result benchmarks mislead: Astra vs Qwen in Paint — rasbt · 2026-09-15
- Cristóbal Valenzuela praises Solaris: 'Websites are going to be fun again' — c_valenzuelab · 2026-09-15
- Google DeepMind on speech-to-speech: conversational, intelligent, multimodal — pick two — AI Engineer · 2026-09-15