Tag
The paper presents PRAGMA, a benchmark for evaluating personalized guidance in lifelong conversations, revealing that current large language model systems struggle with effective memory retrieval and reasoning for user-specific guidance.