Delip Rao calls out new Qwen3.5-9B-based model for benchmarking latency but not accuracy vs Jev

deliprao · x · 2026-10-10

Delip Rao highlighted a new text-only model built on the Qwen3.5-9B base, trained on "somewhere between a billion and a trillion" tokens, and questioned why its official benchmarking compares latency against Jev but omits accuracy entirely. He directly asked Satya Nadella why—implying selective metric display by the vendor.

Original post →

More from Models

Models channel →