← all papers · overview

Performance-guided LLM Knowledge Distillation For Efficient Text Classification At Scale

Abstract

Large Language Models (LLMs) face significant challenges at inference time due to their high computational demands. To address this, we present Performance-Guided Knowledge Distillation (PGKD), a cost-effective and high-throughput solution for production text classification applications. PGKD utilizes teacher-student Knowledge Distillation to distill the knowledge of LLMs into smaller, task-specif

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).