单镜头 30-60 秒长剪辑难产:社区痛陈三种接续方案全部翻车
SorryINeedHelp1 · reddit · 2026-09-08
A Redditor trying to generate a 30-60 second static single-speaker dialogue clip locally reports every continuation method degrades: latent continuation smudges colors and ruins skin texture past 30s; first-last-frame chaining shifts skin tone darker and dims backgrounds; ref2vid segmentation with a bridge clip shows the same color drift.
Even with high step counts and no turbo/attention shortcuts, 15 seconds is the OOM limit on decent settings. The poster asks whether a single uncut long take is simply not achievable with current local video models.
「多模态」频道最新
- 无提示词直出大片:Grok 视频风格实测 4 连发 — azed_ai · 2026-09-08
- 多模型混搭产出全 AI 视频:Qwen 3.8、Flux 2 Klein、LTX 2.5 各司其职 — CQDSN · 2026-09-08
- Astra 依照片与旧手册 2 分钟建模停产马桶配件 — rms80 · 2026-09-08
- 从一张图造角色数据集:img2img 复现同一人物屡屡失败求工作流 — ThePerfectStormy · 2026-09-08
- MiniMax 短片《Cabin Pressure》:视频模型叙事能力的新展示 — Sleepy_Bandit · 2026-09-08
- 用 500 美元 DJI 无人机遥测数据在 Blender 重建房屋 3D 场景 — doodlestein · 2026-09-08