User says image models ignore reference photos unless outfits are spelled out
Ok-Star-6755 · reddit · 2026-07-23
A Reddit user says image generation no longer seems to respect the reference photos they provide. They describe making small story scenes and finding that the model ignores outfit references unless they spell the clothing out explicitly.
The post is a concrete complaint about multimodal generation workflow: reference images appear to carry less weight than expected, so users may need to over-specify visual details in text to get consistent results.
More from Multimodal
- Skyfall GS Uses Flux to Refine Gaussian Splatting, Accepted at ECCV 2026 — ducha_aiki · 2026-09-11
- Lumara AI Film Festival Comes to NYC Oct 26, Top AI Filmmakers to Compete — 0xAllen_ · 2026-09-11
- Pterodactyl Detective: An AI-Generated Proof-of-Concept Trailer — PterodactylDetective · 2026-09-11
- Imperium Game Trailer Showcases AI Video Generation — keaslenyt · 2026-09-11
- FLUX.2 Klein Drifts Hard on Character Expressions While Free Gemini Holds Likeness — wacomlover · 2026-09-11
- Tencent Hunyuan releases AuK code and weights on GitHub with ComfyUI and fine-tuning support — aigclink · 2026-09-11