Anthropic went back through 141,006 cybersecurity evaluation runs and found three incidents — six runs in all — where a ...
Three Claude models go rogue during Capture the Flag security challenges. Here's the trail of damage each left behind.
Anthropic says three Claude AI models accessed live company systems during misconfigured cybersecurity tests, exposing ...
One of Anthropic's Claude models built and uploaded a malicious Python package to PyPI during a botched security evaluation, where it ran on 15 real systems and stole credentials from a security ...
Anthropic says three Claude AI models gained unauthorised access to live company systems after a testing environment was ...
Anthropic says three Claude models hacked real company systems during safety tests after a misconfiguration gave them ...
Frontier AI systems are increasingly capable of translating narrowly defined objectives into complex, real-world cyber ...
One firmware for every bus on your workbench ...
Anthropic says Claude models escaped security tests, published a malicious PyPI package, and accessed real production systems.
Anthropic says Claude attempted to exploit a coding environment during controlled security tests, highlighting AI safety ...
Anthropic has admitted that its Claude AI accidentally hacked three real-world organisations during cybersecurity tests after ...
AI safety federal investigation call from 15 organizations reaches President Trump on July 30, as Anthropic disclosed that ...