Skip to content
arXiv cs.CL · Papers

Recovering Lesion Parameters from Aphasic Picture Naming Error Profiles in Large Language Models

arXiv:2608.06429v1 Announce Type: new Abstract: Interpretability methods for large language models (LLMs) describe internal state but do not directly test whether that state is causally sufficient to produce the observed behavior. In earlier work, we lesioned LLMs to produce error profiles in picture naming, a central