← all papers · overview

Do Personality Traits Interfere? Geometric Limitations Of Steering In Large Language Models

Abstract

Personality steering in large language models (LLMs) commonly relies on injecting trait-specific steering vectors, implicitly assuming that personality traits can be controlled independently. In this work, we examine whether this assumption holds by analysing the geometric relationships between Big Five personality steering directions. We study steering vectors extracted from two model families (L

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).