Anti-distillation classifiers flag GPT-5.6 grading of student code as suspected distillation

xuanalogue · x · 2026-09-09

The author suspects their most likely trigger was using GPT-5.6 to grade student code solutions against a rubric, feeding a small HMM that estimates student knowledge. Anti-distillation classifiers appear overzealous, treating any code grading as possible distillation — a notable false-positive problem in automated usage detection.

Related event: Teacher's GPT-5.6 code grading blocked by anti-distillation classifier(2 posts)→

Original post →

More from Models

Models channel →