Developing story Security3 updates over 1 day

OpenAI GPT-Red safety system

OpenAI introduced GPT-Red, an automated red‑team system that uses self‑play to test and improve model robustness against prompt injection and alignment failures.

One continuously updated timeline instead of dozens of separate articles. New developments are appended as the story evolves.

  1. AnnouncementJul 15, 2026, 10:00 AM

    OpenAI announces GPT-Red, an automated red‑team system for improving model robustness

    OpenAI introduced GPT-Red, an automated red‑team system that uses self‑play to test and improve model robustness against prompt injection and alignment failures.

    Read the full story →
  2. 7 hours later
    AnnouncementJul 15, 2026, 05:09 PM

    OpenAI releases GPT-Red to enhance model security

    OpenAI has developed GPT-Red, an LLM designed to test the defenses of its other models. The latest GPT-5.6 model was trained against GPT-Red, making it the most robust release yet.

    Read the full story →
  3. 1 day later
    AnnouncementJul 16, 2026, 06:48 PM

    OpenAI reveals details of its internal automated red-teaming model, GPT-Red

    OpenAI's GPT-Red model outperformed human red-teamers in a prompt injection test, with a success rate of 84% compared to 13%. The model was trained using self-play reinforcement learning against a population of defender LLMs.

    Read the full story →