Ref2V demo: H3 model handles complex prompts in just 15 seconds
Jeffu · reddit · 2026-08-26
The author shared a video clip generated using the Ref2V method (based on the H3 model), demonstrating the model's ability to handle complex prompts within just 15 seconds.
More from Multimodal
- OraRL: Efficient and Scalable RL for Video MLLMs — Yunheng Li · 2026-08-26
- Seeking Fast HD MiniMax Video Generation Without Quality Loss — OkMeat6773 · 2026-08-26
- Emotional animation of girl touching sky whale generated by Google Gemini — michaelrabone · 2026-08-26
- Wan 3.0 vs. MiniMax H3 Video Comparison: Smoother Audio and Transitions — SimplyAnnisa · 2026-08-26
- AI generated dance video shows cool moves — No-Bookkeeper-char · 2026-08-26
- Face Anything: 4D Face Reconstruction from Any Image Sequence (ECCV 2026) — rsasaki0109 · 2026-08-26