← Back to all articles
Reddit r/LocalLLaMAAugust 21, 2026

Llama.cpp DSpark PC Tree Fork (up to 3%-29.5% faster!)

Excerpt

Hello gang, I made an implementation of DSpark PC Tree (Parent conditioned drafting tree). This is an implementation of this research paper: https://arxiv.org/abs/2608.02123 Unaffiliated, just found it and implemented it. And I have to preface: This is just a first shot, I have no feedback from anyone yet! These are some stats im getting with Qwen 3.0: GPU: SM120 (RTX5090) llama-bench combined Configuration tok/s vs plain vs DSpark n3 Acceptance ━━━━━━━━━━━━━━━ ━━━━━━━━ ━━━━━━━━━━ ━━━━━━━━━━━━━━