← all papers · overview

Deepdecipher: Accessing And Investigating Neuron Activation In Large Language Models

Abstract

As large language models (LLMs) become more capable, there is an urgent need for interpretable and transparent tools. Current methods are difficult to implement, and accessible tools to analyze model internals are lacking. To bridge this gap, we present DeepDecipher - an API and interface for probing neurons in transformer models' MLP layers. DeepDecipher makes the outputs of advanced interpretabi

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).