AI Matters
Why it matters
OpenAI builds automated AI hacker to test its own defenses
MIT Technology Review AI Meet GPT-Red: an LLM super-hacker OpenAI built to make its models safer
OpenAI developed GPT-Red, an AI system that uses self-play to attack other models and find security flaws, which helped secure the newly released GPT-5.6.
Why it matters
Automating security testing helps developers find and patch model vulnerabilities much faster than human teams can.
Latest brief · Jul 22, 2026