← all papers · overview

Alignment Is Not Sufficient To Prevent Large Language Models From Generating Harmful Information: A Psychoanalytic Perspective

Abstract

Large Language Models (LLMs) are central to a multitude of applications but struggle with significant risks, notably in generating harmful content and biases. Drawing an analogy to the human psyche's conflict between evolutionary survival instincts and societal norm adherence elucidated in Freud's psychoanalysis theory, we argue that LLMs suffer a similar fundamental conflict, arising between thei

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).