EVOKE 14B World Model: 3-Step CFG-Free Interactive Generation
Crazy-Repeat-2006 · reddit · 2026-08-16
EVOKE is a 14B autoregressive world model that generates 384×640@24fps video in 3 steps without CFG, maintaining coherence over 30s rollouts. It decouples world state from generation via an external camera-indexed world state bank, enabling endless scenes and mid-flight re-prompting. Weights on Hugging Face, code and demos on GitHub.
More from Multimodal
- LTX 2.5 Full-Resolution Workflows: No Downscaling, Better Detail but Slower — nickinnov · 2026-08-16
- Fixing poor face details in Minimax H3 wide shots — ItsLukeHill · 2026-08-16
- MiniMax H3: Issue with poor face details on wide shots — ItsLukeHill · 2026-08-16
- Black Hole Effect: Krea 2 and MiniMax H3 Workflow — -becausereasons- · 2026-08-16
- Flova × Seedance 2.5 released with stunning 1080p AI video results — SimplyAnnisa · 2026-08-16
- AI video production costs plummet; ByteDance and MiniMax valued like Hollywood studios — Xianbao_QIAN · 2026-08-16