← all papers · overview

Harmaug: Effective Data Augmentation For Knowledge Distillation Of Safety Guard Models

Abstract

Safety guard models that detect malicious queries aimed at large language models (LLMs) are essential for ensuring the secure and responsible deployment of LLMs in real-world applications. However, deploying existing safety guard models with billions of parameters alongside LLMs on mobile devices is impractical due to substantial memory requirements and latency. To reduce this cost, we distill a l

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).