---
id: 20260716-T0-04
title: "OpenAI发布GPT-Red：让AI通过自博弈自动测试自身安全漏洞"
title_en: "OpenAI Launches GPT-Red: AI Self-Play for Automated Red Teaming"
url: https://ai.daily.yangsir.net/daily/20260716-T0-04
issue_date: 2026-07-16
publish_date: 2026-07-15T10:00:00.000Z
category: research
source_name: "OpenAI News"
source_url: https://openai.com/index/unlocking-self-improvement-gpt-red
---

# OpenAI发布GPT-Red：让AI通过自博弈自动测试自身安全漏洞

OpenAI发布GPT-Red，一个自动红队测试系统。该系统利用自博弈（self-play）机制，让AI模型自动生成对抗性输入并测试自身安全边界，从而持续提升面对提示注入等攻击时的鲁棒性。该方法旨在减少人工红队测试的依赖，实现安全评估的规模化自动化。

## English Version

**OpenAI Launches GPT-Red: AI Self-Play for Automated Red Teaming**

OpenAI released GPT-Red, an automated red teaming system that uses self-play to improve AI safety. The system allows models to generate adversarial inputs and probe their own security boundaries, enhancing robustness against prompt injection and other attacks. It aims to reduce reliance on manual red teaming for scalable safety evaluation.

---

**来源**：[OpenAI News](https://openai.com/index/unlocking-self-improvement-gpt-red)

**详情页**：https://ai.daily.yangsir.net/daily/20260716-T0-04

---

*智语观潮 · Daily — https://ai.daily.yangsir.net/llms.txt*