Skip to content
arXiv cs.AI · Papers

RippleBench: Capturing Ripple Effects Using Existing Knowledge Repositories

arXiv:2512.04144v3 Announce Type: replace Abstract: Targeted interventions on language models, such as unlearning or model editing, aim to modify specific information, but their effects often propagate to related, unintended areas (e.g., removing virology content may degrade performance on allergies); these side-effect