New Models to Know: MoE-ViE, τ0-VLA, 4DAnyone, and More
TheTuringPost · x · 2026-08-26
Notable new models this week include:
- MoE-ViE: Meta's sparse mixture-of-experts vision encoder for efficient image and video understanding.
- τ0-VLA: A robot foundation model that "thinks" through possible outcomes before acting.
- 4DAnyone: Turns a single casual video into an animatable 4D human (Ant Group & Zhejiang University).
- WithEveryone: Tencent's image model for generating group scenes while preserving identity.
- PixRestore: A compact diffusion transformer for pixel-space image restoration (OPPO & PolyU).
Related event: This Week's New Models: Meta Vision Encoder and Robot Foundation Models(2 posts)→
More from Models
- Nvidia may have funded 100T tokens for free GLM 5.3 Flash release — bindureddy · 2026-08-26
- Flaw in anti-finetuning: Cost > Quality once models are saturated — rhythmrg · 2026-08-26
- sanoTTS: 1.4M-Param Model Runs Real-Time on $3 Chip — kastnerkyle · 2026-08-26
- User Finds Sol Max More Reliable Than Sol Ultra for Complex Tasks — imjustnewatai · 2026-08-26
- GPT Auto-Titles Conversation in Chinese, Baffling User — rodrigoinfloripa · 2026-08-26
- QUASAR Releases Fully Quantized NVFP4 Qwen3.8-27B, Near-BF16 Performance — arty_photography · 2026-08-26