← all papers · overview

Bias A-head? Analyzing Bias In Transformer-based Language Model Attention Heads

Abstract

Transformer-based pretrained large language models (PLM) such as BERT and GPT have achieved remarkable success in NLP tasks. However, PLMs are prone to encoding stereotypical biases. Although a burgeoning literature has emerged on stereotypical bias mitigation in PLMs, such as work on debiasing gender and racial stereotyping, how such biases manifest and behave internally within PLMs remains large

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).