Recommended inference engine/model for LTX/Wan/MiniMax on 2-node GB10 Spark cluster
ElSrJuez · reddit · 2026-08-25
A user with a 2-node Nvidia GB10 Spark cluster is seeking recommendations for a local inference engine and suitable models to run LTX, Wan, or MiniMax. The goal is to provide an OpenAI-compatible API for generating short image-to-video (i2v) character and scene animations.
More from Infra
- Smaller models could reshape deployment economics with high efficiency — eyishazyer · 2026-08-25
- Hugging Face libraries trade raw speed for broad compatibility and feature coverage — bclavie · 2026-08-25
- Chimera Boosts Multi-Vector Retrieval Throughput by 16x via GPU-CPU Co-Processing — _reachsumit · 2026-08-25
- Fix 7900 XTX Linux Crashes via amdgpu.runpm=0 — Snoo_81913 · 2026-08-25
- Leaked Apple Siri Server Features 32 Apple Silicon Chips, No Local SSDs — asymco · 2026-08-25
- Blackwell performance questioned in high interactivity scenarios — zephyr_z9 · 2026-08-25