arXiv · 2511.14772
Test-time Scaling of LLMs: A Survey from A Subproblem Structure Perspective
Abstract
With this paper, we survey techniques for improving the predictive accuracy of pretrained large language models by allocating additional compute at inference time. In categorizing test-time scaling methods, we place special emphasis on how a problem is decomposed into subproblems and on the topological organization of these subproblems whether sequential, parallel, or tree-structured. This perspective allows us to unify diverse approaches such as Chain-of-Thought, Branch-Solve-Merge, and Tree-of-Thought under a common lens. We further synthesize existing analyses of these techniques, highlighting their respective strengths and weaknesses, and conclude by outlining promising directions for future research
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Zhuoyi Yang, Xu Guo, Tong Zhang, Huijuan Xu, Boyang Li. 2025-11-01. Test-time Scaling of LLMs: A Survey from A Subproblem Structure Perspective. https://arxiv.org/abs/2511.14772
Cite the original work for its findings. Save a collection to share your selection of sources.