Skip to content
HF Daily Papers · Papers

Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning

Fine-tuning LLMs to inject new knowledge faces a critical challenge: LLMs can quickly memorize new facts, yet fail to use them for downstream reasoning tasks. We formalize this failure as the textbf{Knowing--Using Gap}, characterized by an accuracy gap and a temporal lag between memorization and ge