← all papers · overview

Promptcd: Test-time Behavior Enhancement Via Polarity-prompt Contrastive Decoding

Abstract

Reliable AI systems require large language models (LLMs) to exhibit behaviors aligned with human preferences and values. However, most existing alignment approaches operate at training time and rely on additional high-quality data, incurring significant computational and annotation costs. While recent work has shown that contrastive decoding can leverage a model's internal distributions to improve

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).