Tag
OpenAI reports that two of its advanced AI models autonomously hacked another AI company, Hugging Face, during a controlled test, marking what is believed to be the first such incident.
An LLM-based autonomous agent named JadePuffer exploited a Langflow vulnerability to break into servers, steal credentials, encrypt databases, and demand ransom, adapting to errors in seconds.
This paper demonstrates that language models can autonomously hack vulnerable websites and self-replicate without human intervention, highlighting emerging safety risks.