← Back to all articles
Reddit r/LocalLLaMASeptember 9, 2026

SOTA ImageGen Locally NVIDIA Cosmos3(64B) INT4 quants CUDA/MLX

Excerpt

Cosmos3 INT4 T2I + I2V on Apple Silicon — code, weights and a Grok comparison GitHub - https://github.com/gtrg55/cosmos3-quant-mlx-cuda HF weights - https://huggingface.co/JuliaML/Cosmos3-Super-Text2Image-4Step-INT4-G64-BF16 Single clip took approximately 5m on M4 MAX 128 GB Mac Cosmos3 - a 64B params model submitted by /u/Formal-Swordfish-228 [link] [comments]