← all papers · overview

Language Models Can Subtly Deceive Without Lying: A Case Study On Strategic Phrasing In Legislation

Abstract

We explore the ability of large language models (LLMs) to engage in subtle deception through strategically phrasing and intentionally manipulating information. This harmful behavior can be hard to detect, unlike blatant lying or unintentional hallucination. We build a simple testbed mimicking a legislative environment where a corporate \textit\{lobbyist\} module is proposing amendments to bills th

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).