← all papers · overview

Pace: Parsimonious Concept Engineering For Large Language Models

Abstract

Large Language Models (LLMs) are being used for a wide variety of tasks. While they are capable of generating human-like responses, they can also produce undesirable output including potentially harmful information, racist or sexist language, and hallucinations. Alignment methods are designed to reduce such undesirable outputs via techniques such as fine-tuning, prompt engineering, and representat

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).