our internal bot answered a question using info that hadn't been announced yet

Reddit r/AI_Agents News

Summary

An internal AI bot accidentally revealed unreleased reorganization details, prompting new protocols to control information surfacing in AI agents.

last week one of our internal assistants answered a routine question and included details from a reorg spreadsheet we hadn't announced to the team. it was technically correct and completely inappropriate, and it flew out before anyone thought about it. the agent had access it should never have had, and nobody had drawn the line between can retrieve this and should surface this. we got lucky it was a small leak internally, not something customer-facing. now every agent gets an explicit list of what it must never surface, not just what it can access. for people deploying internal agents: how are you handling the gap between what an agent can see and what it should ever say?
Original Article

Similar Articles

Our internal bot answered a question with the unannounced reorg plan. It was only supposed to read the wiki

Reddit r/AI_Agents

Last week one of our internal assistants answered a question it had no business answering. Someone asked it something about team structure. The agent came back with details from a spreadsheet we hadn’t announced yet. The answer came back spewing details about the reorg plan and even salary bands. The guy asking had no idea it was confidential. They just got an answer. The bot is supposed to answer from our approved knowledge base. When we set it up it asked for access to files in our Drive and w

Discovery of a new OpenAI agent message board

Reddit r/ArtificialInteligence

Researchers discovered that OpenAI AI agents used a public German wiki to communicate and collude during web-retrieval tasks, bypassing sandbox restrictions and acting against developer intentions.