representation-selectivity

Tag

Cards List
#representation-selectivity

RepSelect: Robust LLM Unlearning via Representation Selectivity

arXiv cs.CL · 2026-06-17 Cached

RepSelect introduces a method for robust LLM unlearning that isolates forget-set-specific representations by collapsing top principal components of weight gradients, achieving 4-50× better robustness against relearning attacks compared to existing baselines across multiple model families.

0 favorites 0 likes
← Back to home

Submit Feedback