We synthesized 27 papers on AI agent safety into a citation-backed mindmap
Summary
This article presents a synthesized mindmap of 27 papers on AI agent safety, providing a citation-backed overview of the field.
Similar Articles
@Sa4d_k1: Among the distinguished papers recently published in the field of AI Agents. The research involved 60 researchers from …
This paper surveys agent memory in AI agents, addressing the problem of context explosion and proposing a framework for self-evolving agents through various memory types and management strategies.
@UnTalNixon_exe: The Definitive Map of AI Agents: 35 Agentic Architectures with Comparative Benchmarks Building an AI agent isn’t just a…
This article presents a repository that systematically gathers and benchmarks 35 AI agent architectures, helping developers choose effective control structures for production systems.
[R] AI Agent Security: The Complete Guide to Threats, Defenses, and the Future of Autonomous AI Safety [R]
A comprehensive guide to AI agent security covering major incidents from April–June 2026, defensive architectures, and government regulatory responses, synthesizing 18 articles from The Agent Report.
Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security
This survey provides a comprehensive examination of trustworthy agentic AI, focusing on safety, robustness, privacy, and system security. It clarifies key concepts, identifies risks along the agent workflow, summarizes mitigation strategies, and consolidates evaluation metrics and benchmarks, aiming to serve as a practical reference for deploying agentic AI in high-stakes environments.
we keep talking about making agents smarter but not about making them safe around data
The article argues that AI agent safety focuses too much on instruction-following and not enough on data access governance, highlighting the Agentic Data Protocol as an early effort to put policy in infrastructure.