Introducing GPT‑Red, an automated system that tests our AI models for weaknesses—helping us make them safer and more se…
By OpenAI · AI
Introducing GPT‑Red, an automated system that tests our AI models for weaknesses—helping us make them safer and more secure before they’re widely released. As AI models become more capable, our safety measures need to keep pace. Red-teaming—the practice of testing systems by deliberately trying to find their weaknesses—is essential, but today’s methods are difficult to scale. GPT‑Red helps us test more efficiently and strengthen our safeguards.