← all papers · overview

Denevil: Towards Deciphering And Navigating The Ethical Values Of Large Language Models Via Instruction Learning

Abstract

Large Language Models (LLMs) have made unprecedented breakthroughs, yet their increasing integration into everyday life might raise societal risks due to generated unethical content. Despite extensive study on specific issues like bias, the intrinsic values of LLMs remain largely unexplored from a moral philosophy perspective. This work delves into ethical values utilizing Moral Foundation Theory.

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).