← all papers · overview

I Think, Therefore I Am: Benchmarking Awareness Of Large Language Models Using Awarebench

Abstract

Do large language models (LLMs) exhibit any forms of awareness similar to humans? In this paper, we introduce AwareBench, a benchmark designed to evaluate awareness in LLMs. Drawing from theories in psychology and philosophy, we define awareness in LLMs as the ability to understand themselves as AI models and to exhibit social intelligence. Subsequently, we categorize awareness in LLMs into five d

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).