Tag
This paper introduces MirageBench, a benchmark showing that LLMs with persistent memory fabricate user profiles through over-inference 35-49% of the time, and reveals that model self-reported confidence is inversely correlated with actual over-inference, making self-monitoring unreliable for comparing models.