Pangram Can Be Defeated by a Fine-Tuned LLM
theshawwn · x · 2026-07-19
The author cautions: Pangram can indeed be defeated by a fine-tuned LLM. Even though the model was fine-tuned for other purposes, this result is still highly noteworthy.
The attached image shows Pangram's detection interface: a document is labeled as Human Written, with the right panel stating 100% of this text is Human Written. The author presents this as an interesting case study demonstrating that such detection systems are not invulnerable to adversarial attacks.
More from Models
- Grok 4.5 is now free inside Cursor, the popular AI coding IDE — mark_k · 2026-07-21
- GPT often converges on the same near-miss ideas in math problems — yacineMTB · 2026-07-21
- Eno Reyes says model distillation is basically unstoppable — LangChain · 2026-07-21
- Sakana says multiple diffusion models plus MCTS beat test-time scaling on coding and math — SakanaAILabs · 2026-07-21
- OpenAI hackathon project stalls as Codex struggles on voice, while Claude spots the issue — ColleenMBrady · 2026-07-21
- Kimi K3 lands exactly on China’s 2-year AI capability trend line — peterwildeford · 2026-07-21