This paper investigates how the way large language models generate multiple candidate responses affects their performance and energy consumption. Practitioners might care because optimizing test-time scaling can lead to significant improvements in model accuracy and efficiency.
Firehose
Filtered to Papers, tagged “test-time scaling” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives