Reddit r/LocalLLaMAAugust 27, 2026
DeepSeek V4 0731 -> Qwen 3.8 Flash -> GLM 5.3 Flash (and back again!)
Excerpt
Spent yesterday getting Qwen3.8 Flash and GLM 5.3 Flash up and running on my cluster of 4 x DGX Sparks with a view to replacing DeepSeek 0731... but.. really not that impressed with GLM 5.3 - overly verbose and takes for ever (was getting around 22 tok/s on dual spark setup). Have now got myself setup as DeepSeek V4 0731 running on 2 of the sparks and Qwen 3.8 Flash running on the other two. DS is my plan and build and Qwen is explore / scout / subagent work. Seems to be running as a pretty good