HF Daily Papers
· Papers
Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning
Fine-tuning LLMs to inject new knowledge faces a critical challenge: LLMs can quickly memorize new facts, yet fail to use them for downstream reasoning tasks. We formalize this failure as the textbf{Knowing--Using Gap}, characterized by an accuracy gap and a temporal lag between memorization and ge