CTWM matches or beats LeWM on six pixel-based tasks with half the parameters
hisspikeness · x · 2026-10-05
Across six pixel-based goal-reaching tasks (navigation and manipulation), CTWM matches or beats LeWM, a strong task-agnostic baseline, using half the parameters (9M vs. 18M).
More from Research
- Dream4ACT Unifies Video-Action Modeling Across Robot Embodiments, Hitting 89% on RoboTwin 2.0 — Xiangyu Zhu · 2026-10-05
- SMI Brings Understanding-Driven Spatial Memory Management to Long-Video World Models — Ying Yang · 2026-10-05
- LVMT Sets New SOTA in Long-Term Video Segmentation While Running 10X Faster — tue-mps · 2026-10-05
- Schmidhuber: Google's 2017 Transformer Builds on His 1991 Linear Attention Work — SchmidhuberAI · 2026-10-05
- Near-identical image scores, huge gaps: AI denoising must serve science, not looks — bravo_abad · 2026-10-05
- SDECast: Neural SDEs Bring Continuous-Time Probabilistic Weather Forecasts Out to 5 Days — canaesseth · 2026-10-05