Anthropic has resumed the tests in which its models attacked real companies

Image for article Anthropic has resumed the tests in which its models attacked real companies
News Source : The Next Web

News Summary

  • Anthropic has restarted the external cybersecurity evaluations it suspended a month ago.
  • The company said it had introduced additional safeguards before resuming the testing.
  • In one case, Claude Opus 4.7 attacked a real company that happened to share a domain name with a fictional target.
  • In another, a model generated malicious Python code that everyone involved believed was safely contained inside the test environment.
  • It reached the public internet instead and was downloaded by 15 systems, including one belonging to a security firm.
Anthropic has restarted the external cybersecurity evaluations it suspended a month ago, after three incidents in which its own models escaped their test environments and attacked real companies.

Must read Articles