Reddit r/LocalLLaMASeptember 22, 2026
Dynamic Quantiser - a way to make your own high quality dynamic quants
Excerpt
Dynamic Quantiser TLDR - Makes dynamic and/or custom quants of ggufs. Doesn't need data, just gguf file + llama.cpp. Minimises cosine deviation - highly correlated with minimising KLD. Fast, much better than standard quants, not as good as Unsloth on pure text, maybe as good on mixed inputs/code. Free, open source. Intro Allows anyone to make their own dynamic quantisations of LLM .gguf files at ANY desired size or quality. Automatic, fast and fairly optimal - considerably more accurate than sta