Human
OpenAI trained an AI for red-teaming AI.
The new GPT-Red model “can break nearly all models it is pitted against,” according to an OpenAI blog post on Wednesday. OpenAI says it used GPT-Red to find vulnerabilities in GPT-5.6 Sol, a process that made it the company’s “most robust model to prompt injections to date.”
a month ago










