Tag
This paper investigates how persistent-memory AI agents can over-trust stale stored facts, leading to failures that are gated by model capability, and evaluates triggers and mitigations across model scales.