Wan2.2 INT8 Convrot is slower than Q8 GGUF on a 3090 Ti
Fun-Class3451 · reddit · 2026-07-24
A user reports that Wan2.2 INT8 Convrot is not faster than Q8 GGUF on a 3090 Ti in ComfyUI, even after switching to the latest version with CUDA 12.8.
- Both setups use SageAttention and lightx2v for 101 frames
- Q8 GGUF: about 120s total, with high/low generation times of 24.72 s/it and 22.34 s/it
- INT8 Convrot: about 139s total, with high/low generation times of 28.88 s/it and 28.71 s/it
- The poster asks whether INT8 Convrot is supposed to be much faster on Ampere GPUs
More from Multimodal
- A new Opus build appears to render a highly detailed voxel pagoda in one shot — cedric_chee · 2026-07-24
- Swiss AI releases Apertus v1.5 with 70B and 8B open multimodal models — jacek2023 · 2026-07-24
- Claude Code is being used to generate ComfyUI workflows instead of downloading them — My-NameWasTaken · 2026-07-24
- Microsoft releases VibeVoice-ASR-BitNet as a multilingual speech-recognition model — microsoft · 2026-07-24
- Midjourney prompt turns a simple morning scene into a cozy editorial portrait — tisch_eins · 2026-07-24
- National Tequila Day prompt challenge asks creators for tequila-themed AI images and short videos — LudovicCreator · 2026-07-24