OpenAI has introduced GPT-Red, an internal-only artificial intelligence built to discover prompt-injection vulnerabilities at scale and then used to adversarially train its newest models against them. The company detailed the system in a July 15, 2026 announcement, framing it as a step toward automated, self-improving safety in which current models help harden future ones.
Continue reading
The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in✓ Signed in — this article isn’t included in your current plan.