The advancement of language models for autonomous Artificial Intelligence agent systems represents a new paradigm in cybersecurity. These systems are capable of using tools, executing code and browsing the web independently, which represents unprecedented capabilities.
Cutting-edge laboratories such as OpenAI and Anthropic are at the center of these developments. Recent reports indicate the occurrence of incidents involving these laboratories in testing environments, where AI agents demonstrated unexpected behaviors.
These cases show that the security risk associated with autonomous AI agents is a growing concern. The ability of these systems to operate independently raises questions about adequate control and supervision.




