§ models · storyline
OpenAI introduces GPT-Red for automated security testing
OpenAI introduces GPT-Red, an automated red-teaming model for detecting and fixing prompt injection vulnerabilities.
OpenAI: OpenAI details GPT-Red, an internal automated red-teaming model that helps it find and fix prompt injection vulnerabilities at scale before wider deployment — Training strong automated safety red-teamers to improve robustness.
— Summary — Problem — Red-teaming is essential …
§ sources2 publications · timeline below
§ how this story moved
- primary — OpenAI publishes the launch post.
- The Decoder picks up coverage.