Calling for Eval Prompts for Next-Gen GLM

petrusenko_max · x · 2026-07-19

This post is collecting "prompts that current models struggle with," spanning areas like reasoning, coding, SVG, and Chinese, to evaluate the next-generation GLM. While light on details, it signals their effort to build targeted test sets and evaluation samples for upcoming models.

Related event: Next-Gen GLM Seeks Hard Prompts for Evaluation(2 posts)→

Original post →

More from Models

Models channel →