← all papers · overview

Value FULCRA: Mapping Large Language Models To The Multidimensional Spectrum Of Basic Human Values

Abstract

The rapid advancement of Large Language Models (LLMs) has attracted much attention to value alignment for their responsible development. However, how to define values in this context remains a largely unexplored question. Existing work mainly follows the Helpful, Honest, Harmless principle and specifies values as risk criteria formulated in the AI community, e.g., fairness and privacy protection,

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).