Tag
This paper proposes a spillover-aware method for multi-value activation steering to achieve pluralistic alignment in LLMs, improving control over multiple value dimensions without fine-tuning or reward models.