← all papers · overview

Human-instruction-free LLM Self-alignment With Limited Samples

Abstract

Aligning large language models (LLMs) with human values is a vital task for LLM practitioners. Current alignment techniques have several limitations: (1) requiring a large amount of annotated data; (2) demanding heavy human involvement; (3) lacking a systematic mechanism to continuously improve. In this work, we study aligning LLMs to a new domain with limited samples (e.g. < 100). We propose an a

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).