← Back to all articles
Reddit r/LocalLLaMASeptember 1, 2026

MTP released for Qwen3.8-Flash-Next-GGUF

Excerpt

Can't wait to test! This should significantly boost TPS! Now we just need more llama cpp optimizations to be merged in! Edit: For anyone who wants to test this: https://github.com/unslothai/llama.cpp/pull/144/changes More info: https://huggingface.co/unsloth/Qwen3.8-Flash-Next-GGUF/blob/main/MTP/README.md submitted by /u/vini542reddit [link] [comments]