X-MinimaxH3: Local acceleration for MiniMax H3 on consumer GPUs
This_Temporary_8537 · reddit · 2026-08-30
An indie dev open-sourced X-MinimaxH3, accelerating MiniMax H3 on consumer NVIDIA GPUs (tested on RTX 4090):
- Benchmarks: Native 720p→1440p second sampling takes 112s/223s/334s for 5s/10s/15s clips.
- Technique: Reuses latent states for second sampling instead of frame-by-frame upscaling; uses a quality-aware scheduler to allocate compute budgets dynamically.
- Profiles: Supports 8GB (W4A8), 16GB (INT8), and 24GB (INT8) via ComfyUI, REST API, or Web UI.
- Integration: Connects to ComfyUI via HTTP to avoid double model loading.
Related event: Open-Source Project Runs MiniMax H3 on a Single RTX 4090(3 posts)→
More from Infra
- Switching to GLM Flash: Zero quality drop at lower cost — TheZachMueller · 2026-08-30
- Billionaire Mandel Bets Big on Nebius Amid 454% Revenue Growth — DavidLinthicum · 2026-08-30
- Uber's Software Factory: 70% of PRs Attributed to AI Agents — JohnAlexander · 2026-08-30
- 50% price cut drives 14x volume, showing intelligence's price elasticity — skorusARK · 2026-08-30
- The Cache TTL Trap: Why We Need Paid Retention Windows — EmetInteractive · 2026-08-30
- MiniMax H3 native 1440p on a single RTX 4090 via auto-scheduled sparse attention, open-sourced — This_Temporary_8537 · 2026-08-30