← all papers · overview

Llms Explain't: A Post-mortem On Semantic Interpretability In Transformer Models

Abstract

Large Language Models (LLMs) are becoming increasingly popular in pervasive computing due to their versatility and strong performance. However, despite their ubiquitous use, the exact mechanisms underlying their outstanding performance remain unclear. Different methods for LLM explainability exist, and many are, as a method, not fully understood themselves. We started with the question of how ling

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).