How Consistent is ChatGPT's Cross-Scene Portrait Generation

TruthTellerTom · reddit · 2026-07-09

The post compares ChatGPT's image generation capabilities with ComfyUI workflows. The author notes that ChatGPT can seamlessly place a person into new scenes while preserving facial features and overall identity, yielding more natural results than typical face-swapping pipelines.

They ask whether open-source solutions have caught up, mentioning tools like Flux, SDXL, InstantID, PuLID, IP-Adapter, and LoRA.

Original post →

More from Multimodal

Multimodal channel →