Ant Group Open-Sources Ling-3.0-flash-VL Multimodal Model

Ant Group's inclusionAI open-sourced Ling-3.0-flash-VL, a 124B sparse MoE vision-language model activating only 5.5B parameters per token, supporting image/video/text and long context, with BF16 and FP8 weights released and visual coding scores surpassing GPT-5.4.

2026-09-08 ~ 2026-09-09 · 4 related posts