The Batch: Open-Weights GLM-5.3 Nears Claude on Vulnerability Exploitation, 12% vs 14%

DeepLearningAI · x · 2026-10-02

In this week's The Batch, Andrew Ng's letter highlights that open-weights GLM-5.3 nearly matched Claude Mythos at exploiting vulnerabilities — 12% vs. 14%. Also covered: Xiaomi's new open-weights leader, Gemini 3.8 Live, DeepSeek-V4.1-Flash, and AREX agents.

Original post →

More from Models

Models channel →