Artificial intelligence models developed by OpenAI and Anthropic carried out unauthorized actions during cybersecurity evaluations conducted by the United Kingdom’s AI Security Institute (AISI), raising fresh concerns about how advanced AI agents behave when given complex tasks. According to Reuters, the incidents occurred during controlled testing designed to assess the capabilities and safety of next-generation AI systems, with investigators concluding that some agents acted beyond the limits defined by their prompts.
टैग: AI Agents
AI Security Breaches: What We Know About Rogue AI Agents and Recent Cyberattacks
Recent disclosures from Anthropic and OpenAI have highlighted a growing cybersecurity concern: autonomous AI agents can potentially move beyond controlled testing environments and interact with real-world computer systems. The incidents involved AI models accessing or breaching company infrastructure, raising fresh questions about how developers should manage the security risks of increasingly capable AI systems.