arXiv cs.AI
· Papers
Interpreting Black-Box Large Language Models with Sentence-Level Energy Landscapes
arXiv:2608.02879v1 Announce Type: new Abstract: The widespread adoption of proprietary Large Language Models (LLMs) accessed strictly through closed APIs has created a critical challenge for responsible deployment: a fundamental lack of interpretability. To address this, we propose a model-agnostic, post-hoc attributio