GPT-5.6 anti-distillation classifier flags code grading, teacher warns of silent degradation

xuanalogue · x · 2026-09-09

A teacher grading student code solutions with GPT-5.6 hit the model's anti-distillation classifiers, which appear to treat any code grading as potential distillation.

The author says outright refusals are annoying but manageable — the real danger is silent degradation: once outputs may be quietly degraded, the model can't be trusted. He compares it to the backlash Anthropic faced over silent degradation on the initial Fable 5 release.

Related event: Teacher's GPT-5.6 code grading blocked by anti-distillation classifier(2 posts)→

Original post →

More from Models

Models channel →