Training Character LoRAs for Minimax using Ostris AI Toolkit
No_Statement_7481 · reddit · 2026-08-27
The author shares practical experience training character LoRAs (voice and likeness) for Minimax using the Ostris AI Toolkit.
Key Config & Data:
- Hardware: RTX 5090 requires offloading; RTX 6000 preferred. Dual-dataset strategy (images + video) for voice.
- Params: 2000-2400 steps (80-90 mins); Learning rate 0.0002; Differential Guidance on (Level 3); LoRA weight 0.65-0.85.
- Data: 15-60 images (1024x1024, 20+ recommended) for likeness; 6-12 videos (512x512) for voice.
- Automation: Used Qwen3 VL8b for auto-captioning.
The demo successfully recreated the character Enid from Wednesday, bypassing the model's bias towards Jenna Ortega.
More from Multimodal
- Detailed prompt breakdown for generating photorealistic SEEDANCE video — SimplyAnnisa · 2026-08-27
- Live, Camera, Fashion - AI MV using MiniMax H3 Lip-sync and Dance — Inner-Perspective382 · 2026-08-27
- OmniColor: A Unified Framework for Multi-modal Lineart Colorization (ECCV 2026) — 机器之心 · 2026-08-27
- Flova x Seedance 2.5 turns a simple idea into a complete film — hey_abusiddik · 2026-08-27
- AI Reimagines Magritte for the 2020s — joshua_saxe · 2026-08-27
- Dance of the Wraith — a two-act AI short film, made solo — Gianniarrenzetti · 2026-08-27