标签
本文表明,对于编程代理的记忆系统,检查特定声明在代码仓库变更后是否仍然有效,比评估差异中的行为保持提供更高精确度,此结论通过多个LLMs和真实数据的实验得到验证。
This paper empirically studies how VLM agents with persistent spatial memory fail when memory becomes stale, using a dynamic FrozenLake testbed. It finds that trusting stale memory can more than double death rates, and that read-time auditing helps but does not fully close the gap.
本文实证研究了视觉语言模型智能体中的空间记忆过时性问题,发现模型常常忽略矛盾的视觉证据,并且信任过时记忆会增加安全风险。作者提出了审计机制,但表明在记忆-观察冲突下的视觉接地仍是一个重大开放挑战。