NVIDIA Provides Day-0 Support for Alibaba's Qwen3.8-Flash-Next with NeMo
Alibaba_Qwen · x · 2026-08-27
NVIDIA announced Day-0 support for Alibaba's Qwen3.8-Flash-Next model. Developers can now use NVIDIA NeMo AutoModel and NeMo RL to fine-tune the model for domain-specific use cases. Recipes are also available to run the model with sgl, vllm, and TokenSpeed. The model is a hybrid-attention MoE architecture with 125B parameters.
More from Infra
- TokenSpeed adds Day-0 support for Qwen 3.8 Flash Next architecture — Alibaba_Qwen · 2026-08-27
- Z.ai serving 100T tokens/day on Chinese hardware implies training capability — SumitGup · 2026-08-27
- DLSS 4.5 Ray Reconstruction released with 2nd-gen joint denoiser — ctnzr · 2026-08-27
- Startups may measure runway in tokens by 2027 — MillionInt · 2026-08-27
- NVIDIA Covers Full AI Stack via Licenses and Investments — himanshustwts · 2026-08-27
- Concerns Arise Over HuggingFace's Hardware Neutrality After NVIDIA Acquisition — QuixiAI · 2026-08-27