Reddit r/LocalLLaMAAugust 21, 2026
Qwen 3.8 vs 3.6 27b low reasoning loops way less now
Excerpt
Have seen some people say Qwen 3.8 still overthinks even when reasoning is set to low. Which on my case has been way better compared to 3.6, eveb on a 3 bit quant. I think it's worth mentioning that the default is actually xhigh, so first make sure to specify it if not already. Also, Qwen 3.8 has an additional parameter preserve_thinking . It allows to keep/discard the reasoning after every turn. So make sure its activated, otherwise the model may end up reasoning through the same stuff again. M