ComfyUI-QwenVL 2.3.0: Auto Video Downscaling to Kill OOM, Zero-Config Model Downloader
Narrow-Particular202 · reddit · 2026-08-27
ComfyUI-QwenVL v2.3.0 targets the dreaded OOM when feeding multi-frame HD video to Qwen-VL in ComfyUI.
- Smart video auto-scaling & token budget: nodes compute a safe per-frame pixel budget from the model's context window and frame count; 1080p/4K clips are bicubic-downscaled to fit while small videos stay untouched. Manual frame-size override (384–768) available; works on both GGUF and Transformers backends.
- Zero-config HF downloader: auto-scans downloaded files, auto-discovers matching mmproj.gguf projector files, and writes custommodels. registration automatically.
- Modular cross-plugin engine: exposes runqwenvlvision/text APIs and a standalone CLI so other plugins (e.g. MiniMax-H3-Promptor) can reuse your local Qwen-VL without reloading weights.
Repo: github.com/1038lab/ComfyUI-QwenVL
More from coding & agent
- Bronx high schoolers host AI dev competition to support local businesses — ziv_ravid · 2026-08-27
- Automating parsing of Indian subcontinental language books with Claude and Codex — aryaman2020 · 2026-08-27
- User claims $75 stake given to Grok bot grew to $6,140 in 48 hours — RachelVT42 · 2026-08-27
- Which coding tasks justify the highest-capability model in your workflow? — whereismy1 · 2026-08-27
- WebMCP enables seamless website integration for agents with just embedded JS — HankYeomans · 2026-08-27
- Which agent steps deserve the expensive model in long-running workflows? — Top-Construction938 · 2026-08-27