Next-Gen GLM Seeks Hard Prompts for Evaluation

The GLM team is collecting challenging prompts across reasoning, coding, and Chinese to build a targeted test set for evaluating their next-generation model.

2026-07-19 ~ 2026-07-19 · 2 related posts