Krea 2 Local Generation Slower Than Expected
imataruto · reddit · 2026-07-19
The author tested the local generation speed of Krea 2 on an RTX 4090 and found it underperformed.
- Using krea2rawint8convrot + qwen3vl4bfp8scaled at 1MP resolution with 51 steps, the speed was about 4.19s/it, totaling 3 minutes 32 seconds per generation.
- With turbo enabled and 8 steps, it was about 2.2s/ts, total 16 seconds, which the author still finds slow.
- Peak VRAM usage was about 20GB, which doesn’t seem to be a bottleneck on a 24GB 4090.
- Trying Dynamic VRAM toggle and updating ComfyUI didn’t improve speed significantly, so the author suspects a configuration issue and asks if this is normal and what speeds others achieve with similar hardware.
More from Infra
- SF Compute founder: buying compute is 'an absolutely awful experience' right now — IgorCarron · 2026-09-11
- SmolVM open-sources persistent computer infrastructure for agents that outlive chat sessions — aniketmaurya · 2026-09-11
- PyTorch Day Korea 2026 launches first offline conf, CFP closes Sept 13 — PyTorch · 2026-09-11
- Local LLM server dilemma: 4x CMP-170HX (price up 53% in 20 days) vs Mac Studio M5 Ultra — rumboll · 2026-09-11
- llama.cpp lands Flash Attention tuning for RDNA4, big prefill gains on AMD — pmttyji · 2026-09-11
- Your p99 latency benchmark may be lying: a deep dive into coordinated omission — Franc0Fernand0 · 2026-09-11