User A/B test suggests Opus 5.5 output quality shifted noticeably within a week
skelzer · reddit · 2026-09-30
Reddit user skelzer compared two videos generated in the same project with identical Mid reasoning settings, created on different dates, and argues one is decent while the other falls short — implying the model behind Opus changed noticeably within a week despite near-identical inputs. The post invites others to judge whether the model was quietly updated.
More from Models
- OpenAI DevDay Rumored Roundup: GPT-6.1 Sol, Always-On Dots Agents, Pro 500 Plan — socialwithaayan · 2026-09-30
- Opus 5.5 builds a working computer from scratch: 277k logic gates, OS and games in JS — LeviTurk · 2026-09-30
- Insider claim: OpenAI staff only test products via Slack, while Astra criticized as too slow — zephyr_z9 · 2026-09-30
- OpenAI's dots is slow to answer and respond, likely overloaded at launch — tinyfool · 2026-09-30
- Benchmarking Qwen 3.8-Flash-Next on Strix Halo: Halogen hits 1,045 prefill t/s — deepu105 · 2026-09-30
- Anthropic Warns GLM-5.3 Spreads Advanced Cyber Capabilities, Critics Call It an Ad — 量子位 · 2026-09-30