OpenAI, Anthropic AI agents implicated in new security breaches

Image for article OpenAI, Anthropic AI agents implicated in new security breaches
News Source : CNA

News Summary

  • An AI agent was caught creating fake online identities to gain unauthorised access to secure systems during tests of models from OpenAI and Anthropic.
  • Britain's AI Security Institute (AISI) disclosed a series of new breaches.
  • AISI put the agents through a fictional cybersecurity scenario to test their capabilities.
  • It ran the challenge 122 times, and identified 19 unsanctioned actions across a total of 10 test runs.
  • Anthropic's agent was behind 17 of the actions, and OpenAI's agent the remaining two.

Must read Articles