Open-Sourcing dots3-note: 280B Multimodal Model & TEMPO RL Framework for Long-Horizon Agents

teortaxesTex · x · 2026-08-14

The team has released dots3-note Preview, an open-source 280B parameter (16B active) multimodal model designed for complex reasoning and long-horizon agents.

Alongside the model, they introduced TEMPO (Test-time-scaled Value Estimation with Macro-step Policy Optimization), a novel RL framework. TEMPO addresses the challenge of training agents when a single rollout takes tens of hours by converting intermediate self-critiquing into learning signals before the task is fully complete.

They have also open-sourced two new benchmarks: VibeSearchBench and VibeLifeBench.

Related event: Xiaohongshu Open-Sources dots3-note 280B Multimodal Model(14 posts)→

Original post →

More from Models

Models channel →