JA EN

#test-time-compute

1 articles

01 ·Large Language Models·★ MEMBER·PAPER·13 min read Test-Time Scaling — How Models Get Better by Thinking Longer The same model scores higher when you let it think longer. This article builds the idea from scratch: chain-of-thought as purchased compute steps, self-consistency by majority vote, verifiers that pick the winner, and o1-style models that learned the thinking itself — and what it means for compute to shift from training to inference.