← Back to all articles
Reddit r/LocalLLaMASeptember 11, 2026

Terminal Bench v4 scores

Excerpt

Some people says terminal bench reflects model intelligence better than the intelligent index. From the look of it, the ranking does seem to reflect how people feel about the open and closed models. For the open models, GLM-5.3 is in a league of its own. GLM-5.3-Flash is leading the current gen of top flash models. Kimi-K3 did pretty bad in this benchmark for its size. Qwen3.8-27B is the only small model that can do something on this bench. Model Score GLM-5.3 41.9% GLM-5.3-Flash 32.8% DSV4.1-Fl