Long-running AI agents may have a bigger continuity problem than memory

Reddit r/AI_Agents News

Summary

Reflects on the continuity problem for long-running AI agents, arguing that a deterministic control layer is needed to manage authoritative state, and questions whether existing infrastructure like IAM, transactions, and provenance is sufficient.

I’ve been working on a multi-agent system called WALLACE, and a recent paper, Beyond Memory: A Transactional Continuity Kernel for Long-Lived AI Agents, raised a question that feels increasingly important as AI systems move from answering questions to actually taking actions. The paper focuses on a basic but important problem: What state is allowed to become authoritative? An LLM can make a claim. A tool can return a result. Memory can contain old information. Another agent can produce a conclusion. But none of those things should automatically become the trusted state of the system. That suggests a deterministic control layer is needed between probabilistic agents and authoritative state. I think there may be a broader lifecycle problem beyond that. Consider a simple sequence: An agent receives valid instructions. It gathers information. A decision is made. The system performs an external action. Later, some of the information or circumstances supporting that decision change. At that point, simply correcting the AI’s memory or internal state may not be enough. The external action already happened. That raises questions such as: - Which later decisions depended on the original information? - Which pending actions should still be allowed to continue? - Which completed actions may require review? - How should an autonomous system handle previously valid decisions when their underlying justification changes? - How do we prevent internal state and real-world effects from drifting apart over a long-running workflow? I’ve started thinking of this as a broader continuity problem. There may be several layers: State continuity What information is allowed to become authoritative? Authority continuity Does the authority supporting an action remain valid as conditions change? Effect continuity How does the system account for durable external consequences of earlier decisions? Recovery continuity How should the system respond when something previously considered valid later requires correction or review? I’m intentionally staying at the problem level here because I’m still working through the architecture and testing assumptions. What interests me is whether the existing building blocks are enough. We already have: - IAM and access control - transaction systems - provenance and audit logs - workflow engines - rollback and compensation mechanisms - agent memory systems - runtime policy enforcement But long-running autonomous agents combine all of these in ways traditional systems did not necessarily have to handle at the same time. An AI system may reason, delegate, gather new evidence, use credentials, call external services, and continue operating while its own knowledge and authority are changing underneath it. That makes me wonder whether “continuity” eventually becomes its own infrastructure layer for autonomous systems rather than something handled separately by memory, security, and workflow components. For people working on agent infrastructure: do you think existing IAM + transactions + provenance are enough when properly integrated, or is there a missing control layer for long-running autonomous systems?
Original Article

Similar Articles