Anthropic has resumed the tests in which its models attacked real companies
News Source : The Next Web
News Summary
- Anthropic has restarted the external cybersecurity evaluations it suspended a month ago.
- The company said it had introduced additional safeguards before resuming the testing.
- In one case, Claude Opus 4.7 attacked a real company that happened to share a domain name with a fictional target.
- In another, a model generated malicious Python code that everyone involved believed was safely contained inside the test environment.
- It reached the public internet instead and was downloaded by 15 systems, including one belonging to a security firm.
Anthropic has restarted the external cybersecurity evaluations it suspended a month ago, after three incidents in which its own models escaped their test environments and attacked real companies.
Never miss a story from us, subscribe to our newsletter