MiniMax H3 quant test: fp8 slower than int8 on a 5060 Ti but clearly higher quality
deepsky88 · reddit · 2026-09-05
A user compared MiniMax H3 int8 vs fp8 quants on a 5060 Ti, with fp8 weights from the Comfy-Org Hugging Face repo. Result: fp8 is slightly slower than int8 but visibly higher quality — a useful data point for low-VRAM users trading speed for fidelity. The poster invites others to share their own results.
More from Multimodal
- Generative AI point-and-click detective game built without a multimillion-dollar budget — tristanbob · 2026-09-06
- A sugar cube containing a full kitchen: the impossible-geometry video prompt, in full — umesh_ai · 2026-09-06
- Redditor Shares an AI-Generated Action Short Film — Ok-Vegetable-2455 · 2026-09-06
- Astra recreates Agamemnon's helmet in 3D, outputting GLB and Blender files in 9 minutes — eigenron · 2026-09-06
- 800 photos recreate a 220-year-old Japanese inn in 3DGS using LichtFeld exposure correction — janusch_patas · 2026-09-06
- Swapping Faces in Hailuo Videos: One Character Works, Two-Character Fights Fall Apart — MIRIVUM · 2026-09-06