July 16, 2026
- OpenAI unveiled GPT-Red, an internal automated red-teamer that finds prompt-injection vulnerabilities at scale via self-play — reportedly succeeding in ~84% of test scenarios versus ~13% for human teams, with results fed into hardening GPT-5.6. One practitioner cautioned that AI testing AI "should not become the only judge of its own" defenses. OpenAI, MIT Tech Review · AI & Model Security
in Relay Chains, Bind-Link Blindspots, and a Wave of Live Zero-Days