VEDA Sparse Attention cuts MiniMax H3 video gen time in half in ComfyUI with no visible quality loss

robomar_ai_art · reddit · 2026-10-07

A Reddit user tested the newly released VEDA Sparse Attention custom node for MiniMax H3 in ComfyUI: on an RTX 4090 Laptop 16GB with a 4-step LoRA, a 15s 1344x768 video took 8:05 without VEDA and 4:32 with 90% sparsity — no visible quality loss.

VEDA isn't a LoRA; a learned predictor estimates which attention tiles matter and computes only that subset. The current predictor works with T2VA, FL2VA and R2VA and isn't limited to 8 steps despite the 8NFE name.

Setup: install the node from GitHub, drop the HuggingFace predictor into ComfyUI/models/veda/, then add the VEDA node on the MODEL line after your model/LoRA loader and before the guider/sampler. The author found the default 90% sparsity works best and invites results from other GPUs.

Original post →

More from Infra

Infra channel →