← all papers · overview

TPD: Enhancing Student Language Model Reasoning Via Principle Discovery And Guidance

Abstract

Large Language Models (LLMs) have recently showcased remarkable reasoning abilities. However, larger models often surpass their smaller counterparts in reasoning tasks, posing the challenge of effectively transferring these capabilities from larger models. Existing approaches heavily rely on extensive fine-tuning data or continuous interactions with a superior teacher LLM during inference. We intr

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).