← all papers · overview

Postmark: A Robust Blackbox Watermark For Large Language Models

Abstract

The most effective techniques to detect LLM-generated text rely on inserting a detectable signature -- or watermark -- during the model's decoding process. Most existing watermarking methods require access to the underlying LLM's logits, which LLM API providers are loath to share due to fears of model distillation. As such, these watermarks must be implemented independently by each LLM provider. I

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).