Skip to content
arXiv cs.AI · Papers

Logit-Contribution Scoring Identifies Non-Literal Retrieval Heads

arXiv:2607.01002v1 Announce Type: cross Abstract: In long-context use, large language models frequently synthesize answers from the meaning of a relevant context span rather than literally copy-pasting them. Identifying which attention heads perform this synthesis matters for interpreting long-context model behavior. Y