← all papers · overview

Distilling Reasoning Ability From Large Language Models With Adaptive Thinking

Abstract

Chain of thought finetuning (cot-finetuning) aims to endow small language models (SLM) with reasoning ability to improve their performance towards specific tasks by allowing them to imitate the reasoning procedure of large language models (LLM) beyond simply predicting the answers. Most existing cot-finetuning methods adopt a pre-thinking mechanism, allowing the SLM to generate a rationale before

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).