← all papers · overview

Source-aware Training Enables Knowledge Attribution In Language Models

Abstract

Large language models (LLMs) learn a vast amount of knowledge during pretraining, but they are often oblivious to the source(s) of such knowledge. We investigate the problem of intrinsic source citation, where LLMs are required to cite the pretraining source supporting a generated response. Intrinsic source citation can enhance LLM transparency, interpretability, and verifiability. To give LLMs su

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).