Cosmos3 (64B) INT4 Quants Bring Local Image and Video Gen to Mac and CUDA
Formal-Swordfish-228 · reddit · 2026-09-09
A developer released INT4 quantized builds of NVIDIA's 64B-parameter Cosmos3 supporting CUDA and Apple Silicon MLX, enabling local text-to-image and image-to-video.
- Code: gtrg55/cosmos3-quant-mlx-cuda on GitHub
- Weights: JuliaML/Cosmos3-Super-Text2Image-4Step-INT4-G64-BF16 on HF, 4-step generation
- Benchmarked at 5 minutes per clip on an M4 Max 128GB
- Includes a quality comparison against Grok
More from Infra
- llama.cpp launches llama.app: one-line local LLM install, zero telemetry — ngxson · 2026-09-09
- Dev sarcastically 'thanks' OpenAI for boosting local, privacy-first LLM inference — ngxson · 2026-09-09
- Fluidstack hits an $18 billion valuation building data centers for Google and Anthropic — MxMnr · 2026-09-09
- The AI hardware paradox: datacenter demand pricing out its own users — Demon-llord · 2026-09-09
- Industry's embodied CV data dwarfs academia's, chart shows log-scale gap — ducha_aiki · 2026-09-09
- Smart LLM routing cuts costs 69% on 120 tasks while keeping 99.2% success rate — shensi · 2026-09-09