Scaling LLM Test-Time Compute Optimally Can be More Effective than Scaling Parameters for Reasoning.
Charlie Victor Snell, Jaehoon Lee, Kelvin Xu, Aviral Kumar
Browse the full ICLR paper archive.
Charlie Victor Snell, Jaehoon Lee, Kelvin Xu, Aviral Kumar
Browse the full ICLR paper archive.