Running Krea 2 on 8GB VRAM Causes RAM Leaks; GGUF Becomes the Only Stable Workaround

Full-Belt3640 · reddit · 2026-08-13

A developer using an 8GB AMD RDNA2 GPU and 32GB RAM on Linux shared their struggles with running the Krea 2 model. When using fp8 or int8 quantizations, RAM usage spikes to 99% after the first generation, freezing the system entirely, even if cache clearing is attempted.

Currently, GGUF quants around 7-8GB in size are the only way to run the model reliably, though very few Krea 2 models offer GGUF versions. The author notes that while ComfyUI's dynamic memory management was supposed to make GGUFs obsolete, unsupported hardware like RDNA2 cards still face significant hurdles.

Original post →

More from Infra

Infra channel →