Running Krea2 GGUF text-to-image in under 15 seconds on a 12GB laptop GPU
Symphony_of_Heat · reddit · 2026-09-24
The author shares a local text-to-image workflow using Krea2turbo-Q4-CR.gguf with Qwen3VL-4B-Instruct and dolphin3.0-qwen2.5:0.5b as the LLM. On an RTX 4080 Laptop GPU (12GB VRAM), 1024x1024 generation takes 14s with a new prompt and 11s with the same prompt. Workflow shared via pastebin.
More from Infra
- Modal Labs in funding talks at ~$15B valuation, Baseten at $26B — nmasc_ · 2026-09-24
- Fireworks launches Ember-1: Kimi K3-based model cuts reasoning tokens by ~40% — sophiamyang · 2026-09-24
- NVIDIA's BioNeMo Runtime Boosts Protein Structure Prediction Throughput 2.9x — AllThingsApx · 2026-09-24
- Liquid AI's DSpark speculative decoding makes LFM2.5-VL-3B up to 3.13x faster — helloiamleonie · 2026-09-24
- Stripe Publishes Data on European Entrepreneurship and Compute — jeff_weinstein · 2026-09-24
- MiniMax H3 speed-ups measured: every known trick on one RTX 5090 — fruesome · 2026-09-24