A new study led by astronomers at The University of Texas at Austin proposes a theory that could solve two astronomical ...
OpenAI introduced GPT-Red, an automated AI system designed to find vulnerabilities in GPT models before release. The company said GPT-Red was used to train GPT-5.6, reducing failures on one of its ...
OpenAI says GPT-Red automates prompt injection testing and helped GPT-5.6 Sol record sixfold fewer direct injection failures ...
OpenAI has built an LLM super-hacker called GPT-Red that it uses as a sparring partner to help its other models boost their ...
Florida International University researchers found AI can be tricked with altered images, which can identify risks in the ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results