← all papers · overview

Artifacts Or Abduction: How Do Llms Answer Multiple-choice Questions Without The Question?

Abstract

Multiple-choice question answering (MCQA) is often used to evaluate large language models (LLMs). To see if MCQA assesses LLMs as intended, we probe if LLMs can perform MCQA with choices-only prompts, where models must select the correct answer only from the choices. In three MCQA datasets and four LLMs, this prompt bests a majority baseline in 11/12 cases, with up to 0.33 accuracy gain. To help e

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).