Reddit r/LocalLLaMAAugust 26, 2026
we made Qwen 3.8 27b MLX vision quants and compared them against other popular community publishers (lm-studio, lukaskremla, mlx-community and etc)
Excerpt
we made vision mlx quants of qwen3.8 27b (9 builds from 8bit at 29.5 GB down to 3.23bpw DWQ at 11.8 GB) and compared them against other community vision mlx quants from hf (we only compared vision builds) the layout comes out of a clipping search we wrote for mlx and on top of that we wanted to try distillation, so the low-bit files are dwq builds, trained against the bf16 model as the teacher. there are still a couple of percent lying around in almost every one of them, so we will keep playing