Skip to content
Kernelia
All news
SafetyOpenAI News

GPT-Red: Unlocking Self-Improvement for Robustness

OpenAI introduced GPT-Red, an automated red‑teaming system that uses self‑play to enhance AI safety, alignment, and resistance to prompt injection attacks. The system aims to improve model robustness through continuous internal testing.

Summary written by Kernelia from the original article by OpenAI News. The story and its rights belong to its author.