arXiv cs.CL
· Papers
Recovering Lesion Parameters from Aphasic Picture Naming Error Profiles in Large Language Models
arXiv:2608.06429v1 Announce Type: new Abstract: Interpretability methods for large language models (LLMs) describe internal state but do not directly test whether that state is causally sufficient to produce the observed behavior. In earlier work, we lesioned LLMs to produce error profiles in picture naming, a central