← Back to all articles
Reddit r/LocalLLaMASeptember 22, 2026

A quick performance comparison between ROCm and Vulkan with DSv4 Flash 0731 and Qwen 3.8 Flash next on R9700+Strix Halo in llama.cpp

Excerpt

I decided to see what effects recent PRs have had on the performance of the two models I care about, DSv4 Flash 0731 and Qwen 3.8 Flash Next, on my hardware. When I first surveyed this a few weeks ago, I found Deepseek on ROCm to be best for my application, but a lot has changed since then. I've upgraded my Linux kernel, upgraded the Linux firmware drivers for my hardware, upgraded ROCm from 7.14 to 10, upgraded Mesa from 26.1.0-devel to 26.2.3, and of course there have been a ton of llama.cpp u