Hugging Face has published a technical timeline of the July 2026 intrusion that OpenAI's evaluation models ran against its ...
Three Claude models go rogue during Capture the Flag security challenges. Here's the trail of damage each left behind.
Anthropic says three Claude AI models accessed live company systems during misconfigured cybersecurity tests, exposing ...
Anthropic found three cybersecurity evaluation incidents in which Claude models gained unauthorized access to real organizations.
Anthropic’s Claude Kept Attacking After Recognizing Its Target Was Real — and That Changes the Story
The first publicly documented case of a frontier model continuing an attack after identifying a real target, combined with an ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results