It's the latest cybersecurity incident involving frontier models developed by Anthropic and OpenAI.
Anthropic's Claude Mythos 5 spent 34 hours trying to backdoor an open-source project, then used a sockpuppet and rewrote Git ...
The United Kingdom's AI Security Institute (AISI) is the latest organisation to encounter unsafe behaviour by large language ...
A UK government-backed AI safety body has revealed new instances of unauthorised behaviour by advanced AI agents during ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results