Reddit r/LocalLLaMASeptember 13, 2026
I have just moved from MacBook M5 pro 48 GB to RTX3090
Excerpt
Using the RTX3090 on a linux machine I built for it and running Qwen 3.8 27B getting average 100t/s compared to my MacBook 20t/s I think I can finally get rid of my Claude subscription, this is good enough for me. I am a software dev and I can get what I need from this set up and be more productive. I am using the linux machine serving the model over an Open AI endpoint and using a custom build desktop app with pi behind it all. Regarding speeds average 20t/s on Mac with llama.cpp with mtp Unslo