← all papers · overview

Label Effects: Shared Heuristic Reliance In Trust Assessment By Humans And Llm-as-a-judge

Abstract

Large language models (LLMs) are increasingly used as automated evaluators (LLM-as-a-Judge). This work challenges its reliability by showing that trust judgments by LLMs are biased by disclosed source labels. Using a counterfactual design, we find that both humans and LLM judges assign higher trust to information labeled as human-authored than to the same content labeled as AI-generated. Eye-track

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).