dots3-note Preview: 280B Open Multimodal Model for Long-Horizon Agents

dots3-note preview is an open-weight 280B-parameter (16B active) multimodal MoE model for long-horizon agents, introducing TEMPO, a reinforcement learning method using self-critique and time-scaled value estimation for long-horizon training.

2026-08-17 ~ 2026-08-17 · 3 related posts