← all papers · overview

Exploring Activation Patterns Of Parameters In Language Models

Abstract

Most work treats large language models as black boxes without in-depth understanding of their internal working mechanism. In order to explain the internal representations of LLMs, we propose a gradient-based metric to assess the activation level of model parameters. Based on this metric, we obtain three preliminary findings. (1) When the inputs are in the same domain, parameters in the shallow lay

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).