← all papers · overview

Distilling LLM Reasoning Into Graph Of Concept Predictors

Abstract

Deploying Large Language Models (LLMs) for discriminative workloads is often limited by inference latency, compute, and API costs at scale. Active distillation reduces these costs by querying an LLM oracle to train compact discriminative students, but most pipelines distill only final labels, discarding intermediate reasoning signals and offering limited diagnostics of what reasoning is missing an

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).