← all papers · overview

Watch Your Language: Investigating Content Moderation With Large Language Models

Abstract

Large language models (LLMs) have exploded in popularity due to their ability to perform a wide array of natural language tasks. Text-based content moderation is one LLM use case that has received recent enthusiasm, however, there is little research investigating how LLMs perform in content moderation settings. In this work, we evaluate a suite of commodity LLMs on two common content moderation ta

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).