After Two Years, We Finally Have a 'Local Sora'
MustBeSomethingThere · reddit · 2026-08-10
The poster expressed that two years after OpenAI first previewed Sora, open-source video models have finally achieved an impressive 'Local Sora' level.
The post includes a highly cinematic video demo and shares the detailed prompt structure:
- Visuals: A static close-up shot framing a glass train window looking out at Tokyo suburbs, with a clear reflection of a young woman looking at her phone inside the cabin.
- Sound design: Features a continuous low train rumble, rhythmic track click-clacks, and the muffled swoosh of wind.
This demonstrates that current open-source models can handle complex physical reflections while simultaneously generating matching environmental audio.
More from Multimodal
- BBC-Style AI Short Film: The Bizarre World of an Alien Planet — ctrl-shift-face · 2026-08-10
- MiniMax-H3-Image-VAE: Experimental Single-Image Model Released — physalisx · 2026-08-10
- AI Animated Short: The Adventures of Super Gus and Power Whiskers — Judindor · 2026-08-10
- Targeting Single Character Face/Body Swap in Wan 2.2 Multi-Subject Videos — SilentThree · 2026-08-10
- AI Video Generation Test: POV Riding a Maglev on Venus — fofrAI · 2026-08-10
- Minimax H3 Hack: Compositing 50+ Reference Elements in One Video — No_Damage_8420 · 2026-08-10