dots3-note: 280B MoE agent learns memory via RL

Aiden_Tech_Ai · x · 2026-08-17

dots3-note preview is released as a step towards long-horizon agency. It features TEMPO, a new RL approach for agent training via self-critiquing.

Key Specs:

It learns what to remember through RL and logs its own bugs. Open weights are available on Hugging Face, along with two new benchmarks: VibeSearchBench and VibeLifeBench.

Original post →

More from Models

Models channel →