← all papers · overview

When Large Language Models Contradict Humans? Large Language Models' Sycophantic Behaviour

Abstract

Large Language Models have been demonstrating broadly satisfactory generative abilities for users, which seems to be due to the intensive use of human feedback that refines responses. Nevertheless, suggestibility inherited via human feedback improves the inclination to produce answers corresponding to users' viewpoints. This behaviour is known as sycophancy and depicts the tendency of LLMs to gene

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).