July 15, 2026 · MIT Technology Review
OpenAI built an internal AI 'super-hacker' to red-team its models
OpenAI built GPT-Red, an LLM designed to automate offensive security testing against its other models as a sparring partner. The company says training its newly released GPT-5.6 against GPT-Red made it OpenAI's most robust model yet against cyberattacks.
Why it matters: It shows how AI labs are using AI itself to find and close security gaps before public release.