← Back to all articles
arXiv cs.LGOctober 1, 2026

Towards Better Exploration in Sequential Test-Time Scaling

Excerpt

arXiv:2609.39632v1 Announce Type: new Abstract: Test-time scaling improves language model reasoning by spending additional compute at inference. However, both classes of existing methods often fail to continue improving over long timescales. Parallel methods repeatedly sample independent answers from the model, scaling poorly on problems the model is unlikely to solve in a single attempt. In contrast, sequential methods build on previous answers to access new ideas, yet so far have not been show