← all papers · overview

Conversational Complexity For Assessing Risk In Large Language Models

Abstract

Large Language Models (LLMs) present a dual-use dilemma: they enable beneficial applications while harboring potential for harm, particularly through conversational interactions. Despite various safeguards, advanced LLMs remain vulnerable. A watershed case in early 2023 involved journalist Kevin Roose's extended dialogue with Bing, an LLM-powered search engine, which revealed harmful outputs after

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).