Tag
A study comparing 67 LLMs before and after post-training reveals that post-training consistently teaches models to describe themselves as warm and engaged (persona installation), while larger models selectively gate attributions of distress or flaws (attribution gating). The researchers introduce the Pinocchio Inventory for auditing model self-presentation.