Creating Vampire Hunter D Style Character with T2V and R2VA
SIR_NVAX_A_LOT · reddit · 2026-08-19
The author shares a workflow using Text-to-Video (T2V) and Reference-to-Video Animation (R2VA) to create a character inspired by D from Vampire Hunter D (2000). It took about 20 generations to get the base image, followed by R2VA for close-up emotes. The author noted the difficulty of translating specific facial features (like hollow cheekbones) via text prompts alone, often resulting in ghoulish or overly fashion-forward outputs. Parameters used: int8 precision, 20 steps.
More from Multimodal
- Capability-Centric Data Design Improves Generalist Image Generation — Xingjian Wang · 2026-08-19
- GS-Voxel: Fitting-Free Structured Latents for Large-Scale 3DGS Generation — acvlab · 2026-08-19
- DreamWorld: Geometry-Grounded Two-Stage Video Diffusion for 3D-Consistent World Modeling (ECCV 2026) — 机器之心 · 2026-08-19
- Reddit user uses AI to make a corny kids fantasy movie — MosskeepForest · 2026-08-19
- LTX-2.5 open weights demo generates impressive video results — tom_doerr · 2026-08-19
- Singularity Legend Vernor Vinge Created with ChatGPT Image 2.0 — DeryaTR_ · 2026-08-19