← all papers · overview

I Am A Strange Dataset: Metalinguistic Tests For Language Models

Abstract

Statements involving metalinguistic self-reference ("This paper has six sections.") are prevalent in many domains. Can current large language models (LLMs) handle such language? In this paper, we present "I am a Strange Dataset", a new dataset for addressing this question. There are two subtasks: generation and verification. In generation, models continue statements like "The penultimate word in t

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).