Qwen 3.8 Max reaches 42% on the hard INDUCTION benchmark, taking second place

DeryaTR_ · x · 2026-08-04

Qwen 3.8 Max jumps to 42% on the INDUCTION benchmark

A GitHub leaderboard shared in the thread shows Qwen 3.8 Max reaching 42.2% on the challenging INDUCTION benchmark, up from 0% for Qwen 3.7, and taking 2nd place behind GPT-5.6 Sol.

What the leaderboard says

Why it matters

The benchmark is presented as a difficult induction task from ICML 2026, so the result is framed as evidence of rapid progress in Chinese open-source models over the last few months.

Original post →

More from Models

Models channel →