SGLang turns Qwen3.8-27B into a decision model that beats Pokémon FireRed at sub-100ms
zhaoran_wang · x · 2026-09-30
The SGLang team turned Qwen3.8-27B into a multimodal decision model that beat Pokémon FireRed's elite four and champion with sub-100ms decisions from live game state. SGLang now offers a native /v1/decisions endpoint to use LLMs/VLMs as classification and scoring models, plus /v1/systemone for Jev-like open models with the TypeSafe SDK.
More from Infra
- DeepSeek open-sources DeepGEMM Ascend port, hitting 99.8% of hardware limit on GEMM — zheanxu · 2026-09-30
- AI intelligence-cost Pareto frontier shifted fast: GPT-5 mini at 17 ($0.05) to Claude Opus 5.5 at 58 ($5.98) — ArtificialAnlys · 2026-09-30
- Google's Project Suncatcher to fly TPUs in space for the first time on Oct 1 — allisondman · 2026-09-30
- Photon 2.6 ships FP8 + speculative decoding, runs Qwen3.5 27B at 400+ tok/s on B200 — Bedrovelsen · 2026-09-30
- Agentic AI turns CPUs into the overlooked bottleneck as CPU:GPU ratios shift upward — AccBalanced · 2026-09-30
- AT&T CEO says SpaceX's phone strategy is not viable: "Satellite won't beat fiber" — RachelVT42 · 2026-09-30