arXiv cs.AI
· Papers
RippleBench: Capturing Ripple Effects Using Existing Knowledge Repositories
arXiv:2512.04144v3 Announce Type: replace Abstract: Targeted interventions on language models, such as unlearning or model editing, aim to modify specific information, but their effects often propagate to related, unintended areas (e.g., removing virology content may degrade performance on allergies); these side-effect