llm-personas

Tag

Cards List
#llm-personas

"Many Are My Names": The Anatomy of the Assistant and Its Personas via Sparse Autoencoders

arXiv cs.CL ↗ · 2026-08-11 Cached

This paper uses sparse autoencoders to decompose how language models represent the default Assistant, roleplay personas, and story characters, finding that personas retain an Assistant core while differentiating across layers, and story characters lack that core.

0 favorites 0 likes
← Back to home

Submit Feedback