Qwen3.8-Max takes No. 1 on FlashInfer after 500 runs and 30K tool calls

YouJiacheng · x · 2026-07-22

Qwen3.8-Max tops FlashInfer benchmark

The attached leaderboard shows Qwen3.8-Max taking #1 on SOL-ExecBench FlashInfer. The post says the result is averaged over 500 runs with 30K tool calls, using Qwen3.8-Max-preview with an Atrex Kernel Agent.

The author also frames it as possibly the first frontier model outside GPT and Opus/Fable to show strong production-kernel optimization capability.

Related event: Alibaba's Qwen3.8-Max-Preview Tops FlashInfer Benchmark(2 posts)→

Original post →

More from Infra

Infra channel →