relearning-attacks

Tag

Cards List
#relearning-attacks

Robust LLM Unlearning Against Relearning Attacks: The Minor Components in Representations Matter

arXiv cs.CL · 2026-05-13 Cached

This paper introduces Minor Component Unlearning (MCU), a novel approach to LLM unlearning that targets minor components in representations to resist relearning attacks. It addresses the vulnerability of existing methods by focusing on robust directions within the model's spectral structure.

0 favorites 0 likes
← Back to home

Submit Feedback