← Back to all articles
Reddit r/LocalLLaMASeptember 26, 2026

Run Qwen3.8+Flash-Next and tiny models on Apple Silicon up to 3x faster

Excerpt

Maybe you'll like it? I hope I get to use my self-promotion credit a tiny little bit here after being in the community so long haha. I was the top of MLX.fast for a while and remain the winner on chips below M5. If you have capacity to contribute further enhancements I'd love that https://github.com/struffl/ishizuki submitted by /u/VagabondTruffle [link] [comments]