← all papers · overview

Utilize The Flow Before Stepping Into The Same River Twice: Certainty Represented Knowledge Flow For Refusal-aware Instruction Tuning

Abstract

Refusal-Aware Instruction Tuning (RAIT) enables Large Language Models (LLMs) to refuse to answer unknown questions. By modifying responses of unknown questions in the training data to refusal responses such as "I don't know", RAIT enhances the reliability of LLMs and reduces their hallucination. Generally, RAIT modifies training samples based on the correctness of the initial LLM's response. Howev

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).