Reddit r/LocalLLaMASeptember 26, 2026
For the longest time I’ve felt this sub should have a pinned section where a detailed post about each model should get featured.
Excerpt
For instance whenever a model comes out, what’s the best engine to run it, the best harness and absolute minimum you need to get same or near same re results that the benchmark of that model claims. And whenever a quant from Unsloth guys comes out the guide can either be updated or new guide could be added for that quant. For example Gemini keeps telling me an 8x v100 server is no good to self host deepseek v4.1 but it’s super difficult to find the right answer to my question from an hallucinati