← all papers · overview

Ensuring Safe And High-quality Outputs: A Guideline Library Approach For Language Models

Abstract

Large Language Models (LLMs) exhibit impressive capabilities but also present risks such as biased content generation and privacy issues. One of the current alignment techniques includes principle-driven integration, but it faces challenges arising from the imprecision of manually crafted rules and inadequate risk perception in models without safety training. To address these, we introduce Guide-A

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).