parallelquant
July 15, 2026 · MIT Technology Review

OpenAI built an internal AI 'super-hacker' to red-team its models

OpenAI built GPT-Red, an LLM designed to automate offensive security testing against its other models as a sparring partner. The company says training its newly released GPT-5.6 against GPT-Red made it OpenAI's most robust model yet against cyberattacks.

Why it matters: It shows how AI labs are using AI itself to find and close security gaps before public release.

Related updates