A new study led by astronomers at The University of Texas at Austin proposes a theory that could solve two astronomical ...
OpenAI introduced GPT-Red, an automated AI system designed to find vulnerabilities in GPT models before release. The company said GPT-Red was used to train GPT-5.6, reducing failures on one of its ...
OpenAI says GPT-Red automates prompt injection testing and helped GPT-5.6 Sol record sixfold fewer direct injection failures ...
OpenAI has built an LLM super-hacker called GPT-Red that it uses as a sparring partner to help its other models boost their ...
Florida International University researchers found AI can be tricked with altered images, which can identify risks in the ...