← all papers · overview

Twt: Thinking Without Tokens By Habitual Reasoning Distillation With Multi-teachers' Guidance

Abstract

Large Language Models (LLMs) have made significant strides in problem-solving by incorporating reasoning processes. However, this enhanced reasoning capability results in an increased number of output tokens during inference, leading to higher computational costs. To address this challenge, we propose TwT (Thinking without Tokens), a method that reduces inference-time costs through habitual reason

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).