← all papers · overview

To Burst Or Not To Burst: Generating And Quantifying Improbable Text

Abstract

While large language models (LLMs) are extremely capable at text generation, their outputs are still distinguishable from human-authored text. We explore this separation across many metrics over text, many sampling techniques, many types of text data, and across two popular LLMs, LLaMA and Vicuna. Along the way, we introduce a new metric, recoverability, to highlight differences between human and

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).