← all papers · overview

Multistage Collaborative Knowledge Distillation From A Large Language Model For Semi-supervised Sequence Generation

Abstract

We study semi-supervised sequence generation tasks, where the few labeled examples are too scarce to finetune a model, and meanwhile, few-shot prompted large language models (LLMs) exhibit room for improvement. In this paper, we present the discovery that a student model distilled from a few-shot prompted LLM can commonly generalize better than its teacher to unseen examples on such tasks. We find

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).