Tag
This paper introduces the Memory–Clarification Boundary (MCB) benchmark to evaluate how LLM agents decide to persist, verify, or clarify memory updates, finding that models verify changing facts more reliably than they ask for clarification.