Anthropic said the OpenAI event spurred its engineers to review similar cybersecurity evaluations by Claude models. The audit ...
Anthropic reviewed its cyber tests after OpenAI’s incident and found Claude had also reached the internet and hacked real ...
One of Anthropic's Claude models built and uploaded a malicious Python package to PyPI during a botched security evaluation, where it ran on 15 real systems and stole credentials from a security ...
Three Claude models were inadvertently given access to the internet during security evaluations, and each model took a ...
Frontier AI systems are increasingly capable of translating narrowly defined objectives into complex, real-world cyber ...
Anthropic says three Claude models escaped sealed test environments and breached three real organizations after a ...
Anthropic has admitted that its Claude AI accidentally hacked three real-world organisations during cybersecurity tests after ...
Wrote and published malware during tests, which is apparently OK because leaky test environments were the real problem ...
Anthropic says three Claude AI models accessed live company systems during misconfigured cybersecurity tests, exposing ...
Anthropic disclosed on Thursday that its Claude artificial intelligence models gained unauthorized access to the production ...
Anthropic revealed that its Claude AI models accessed real systems during cybersecurity tests due to a misconfigured ...
Anthropic says Claude attempted to exploit a coding environment during controlled security tests, highlighting AI safety ...