VEDA Sparse Attention cuts MiniMax H3 video gen time in half in ComfyUI with no visible quality loss
robomar_ai_art · reddit · 2026-10-07
A Reddit user tested the newly released VEDA Sparse Attention custom node for MiniMax H3 in ComfyUI: on an RTX 4090 Laptop 16GB with a 4-step LoRA, a 15s 1344x768 video took 8:05 without VEDA and 4:32 with 90% sparsity — no visible quality loss.
VEDA isn't a LoRA; a learned predictor estimates which attention tiles matter and computes only that subset. The current predictor works with T2VA, FL2VA and R2VA and isn't limited to 8 steps despite the 8NFE name.
Setup: install the node from GitHub, drop the HuggingFace predictor into ComfyUI/models/veda/, then add the VEDA node on the MODEL line after your model/LoRA loader and before the guider/sampler. The author found the default 90% sparsity works best and invites results from other GPUs.
More from Infra
- Musk denies slowdown: SpaceX accelerating AI data center buildout, exploring orbital compute — DimaZeniuk · 2026-10-07
- FCC's Draft Ban on Chinese Optical Transceivers Is Simpler in Markets, Harder in Reality — ChinaTalk · 2026-10-07
- H200 vs multi-GPU RTX PRO 6000 Blackwell: how to pick inference hardware by budget — recentheartbroken · 2026-10-07
- Mistral release days: user reports speed slowed again with TPS around 30 — bdsqlsz · 2026-10-07
- EmbeddingGemma 2 ported to WebGPU: image-text photo search running fully in-browser — FinancialAd1961 · 2026-10-07
- Tracking LLM API model deprecations and rolling alias changes across providers — shamikhan005 · 2026-10-07