← Back to all articles
Reddit r/LocalLLaMASeptember 10, 2026

Running Vision Qwen 3.8 27B on a 16GB Card, the config (45tks).

Excerpt

I am just sharing my config for Qwen 3.8 27b that fits on a 5060TI, what is cool about this is that you can even get vision! and a 85K context (I have 1.5gb of headroom for more context or a better quant) Model: IQ3_XXS-mtp from https://huggingface.co/ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF Using beellama https://github.com/Anbeeld/beellama.cpp Config used: [*] model = ..\llm-models\Qwen3.8-27B-GSQ-RCO-IQ3_XXS-mtp.gguf mmproj = ..\llm-models\mmproj-Qwen3.8-27B-BF16.gguf image-min-tokens = 256 gpu-l