← all papers · overview

Reading Subtext: Evaluating Large Language Models On Short Story Summarization With Writers

Abstract

We evaluate recent Large Language Models (LLMs) on the challenging task of summarizing short stories, which can be lengthy, and include nuanced subtext or scrambled timelines. Importantly, we work directly with authors to ensure that the stories have not been shared online (and therefore are unseen by the models), and to obtain informed evaluations of summary quality using judgments from the autho

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).