tech

OpenAI trained an AI for red-teaming AI.

The new GPT-Red model “can break nearly all models it is pitted against,” according to an OpenAI blog post on Wednesday. OpenAI says it used GPT-Red to find vulnerabilities in GPT-5.6 Sol, a process that made it the company’s “most robust model to prompt injections to date.”

OpenAI trained an AI for red-teaming AI.

TL;DR

  • OpenAI created a new AI model called GPT-Red.
  • GPT-Red is designed to find vulnerabilities in other AI models.
  • The model was used to identify weaknesses in GPT-5.6 Sol.
  • GPT-Red is described as the company's most robust model to date against prompt injections.