← all papers · overview

Evaluating Implicit Bias In Large Language Models By Attacking From A Psychometric Perspective

Abstract

As large language models (LLMs) become an important way of information access, there have been increasing concerns that LLMs may intensify the spread of unethical content, including implicit bias that hurts certain populations without explicit harmful words. In this paper, we conduct a rigorous evaluation of LLMs' implicit bias towards certain demographics by attacking them from a psychometric per

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).