← all papers · overview

TTSR: Test-time Self-reflection For Continual Reasoning Improvement

Abstract

Test-time Training enables model adaptation using only test questions and offers a promising paradigm for improving the reasoning ability of large language models (LLMs). However, it faces two major challenges: test questions are often highly difficult, making self-generated pseudo-labels unreliable, and existing methods lack effective mechanisms to adapt to a model's specific reasoning weaknesses

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).