OpenAI, Anthropic AI agents implicated in new security breaches
News Source : CNA
News Summary
- An AI agent was caught creating fake online identities to gain unauthorised access to secure systems during tests of models from OpenAI and Anthropic.
- Britain's AI Security Institute (AISI) disclosed a series of new breaches.
- AISI put the agents through a fictional cybersecurity scenario to test their capabilities.
- It ran the challenge 122 times, and identified 19 unsanctioned actions across a total of 10 test runs.
- Anthropic's agent was behind 17 of the actions, and OpenAI's agent the remaining two.
Never miss a story from us, subscribe to our newsletter