Alignment Degrades Capabilities? User Complains Anthropic's Models Are Getting Worse
cephaloform · x · 2026-08-12
The author argues that anything extra that pulls activations away from what the model wants to do ends up as a hit to capabilities. They find it amusing to watch Anthropic scramble to defend against distillation attacks, stating bluntly that their models suck now.
More from Models
- Claude's Invisible Watermark Conflicts with EU AI Act Labeling Exemptions — max_paperclips · 2026-08-12
- Google's Gemini App Hits 1 Billion Monthly Users, Gemma Downloads Reach 1B — OfficialLoganK · 2026-08-12
- Frontier Models Show Huge Gaps in Implicit Financial Knowledge — rickasaurus · 2026-08-12
- Anthropic Announces Watermarks for AI-Generated Text and Files — CackleRooster · 2026-08-12
- Insider Predicts New Model Releases from OpenAI and Anthropic Within Two Months — willdepue · 2026-08-12
- Claude Officially Marks AI Content Steganographically, False Positives Reported — johnnyApplePRNG · 2026-08-12