← all papers · overview

Raising The Bar: Investigating The Values Of Large Language Models Via Generative Evolving Testing

Abstract

Warning: Contains harmful model outputs. Despite significant advancements, the propensity of Large Language Models (LLMs) to generate harmful and unethical content poses critical challenges. Measuring value alignment of LLMs becomes crucial for their regulation and responsible deployment. Although numerous benchmarks have been constructed to assess social bias, toxicity, and ethical issues in LLMs

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).