AI Matters

Why it matters

OpenAI builds automated AI hacker to test its own defenses

MIT Technology Review AI Meet GPT-Red: an LLM super-hacker OpenAI built to make its models safer

OpenAI developed GPT-Red, an AI system that uses self-play to attack other models and find security flaws, which helped secure the newly released GPT-5.6.

Why it matters

Automating security testing helps developers find and patch model vulnerabilities much faster than human teams can.