Self-hosted MiniMax H3 renders for $0.03, compared to fal's faster H3 Max
tobowers · x · 2026-08-27
A developer compared self-hosted MiniMax H3 with fal.ai's H3 Max for video generation. Using a rented Nebius H200 GPU with NVIDIA Sol-Attn optimizations, the self-hosted H3 in 'draft' mode renders extremely cheaply ($0.03, 13s), while H3 Max is faster and higher quality but much more costly. The post details performance under identical prompts, noting that cheap draft iterations open new possibilities.
More from Infra
- llama.cpp PR adds dspark support for Nanbeige4.2-3B — jacek2023 · 2026-08-27
- Oak CEO: Every AI agent needs distinct identity and auditability — TechNadu · 2026-08-27
- Structured generation with 0% throughput overhead in production — remilouf · 2026-08-27
- M5 Pro vs M6 for local AI: memory bandwidth debate resurfaces — jonejy · 2026-08-27
- Volcengine Unveils Embodied AI Data Solution with Seedance Video Generation — 火山引擎 · 2026-08-27
- Dual RTX 3090s run ComfyUI video gen 2.7x faster via new open-source Windows NCCL backend — unjusti · 2026-08-27