← all papers · overview

Towards Understanding Multi-task Learning (generalization) Of Llms Via Detecting And Exploring Task-specific Neurons

Abstract

While large language models (LLMs) have demonstrated superior multi-task capabilities, understanding the learning mechanisms behind this is still a challenging problem. In this paper, we attempt to understand such mechanisms from the perspective of neurons. Specifically, we detect task-sensitive neurons in LLMs via gradient attribution on task-specific data. Through extensive deactivation and fine

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).