← all papers · overview

Think Deep, Think Fast: Investigating Efficiency Of Verifier-free Inference-time-scaling Methods

Abstract

There is intense interest in investigating how inference time compute (ITC) (e.g. repeated sampling, refinements, etc) can improve large language model (LLM) capabilities. At the same time, recent breakthroughs in reasoning models, such as Deepseek-R1, unlock the opportunity for reinforcement learning to improve LLM reasoning skills. An in-depth understanding of how ITC interacts with reasoning ac

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).