# Meet GPT-Red: an LLM super-hacker OpenAI built to make its models safer

> OpenAI has built GPT-Red, a model trained not to be helpful but to break other models. Pointed at its own systems, it drove the rate of successful attacks on GPT-5.6 down from more than 90 percent to under 23 percent, and it did so by outperforming the human red-teamers running the same tests. Along the way it surfaced a new attack the company calls fake chain of thought, where false statements are slipped into a model's reasoning steps to steer its conclusions. The idea of automating adversarial testing is not new, but the results here are a reminder that the same capabilities that make models dangerous also make them useful defenders. An AI that can find holes faster than people can is a powerful safety tool. Notably, OpenAI has no plans to release GPT-Red, treating it as a proprietary edge rather than shared safety infrastructure. That choice is its own story: the best AI security tooling may end up locked inside the labs that need it least, while everyone else waits.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: MIT Technology Review · Published Sunday, July 19, 2026_

## Wortins' read

OpenAI has built GPT-Red, a model trained not to be helpful but to break other models. Pointed at its own systems, it drove the rate of successful attacks on GPT-5.6 down from more than 90 percent to under 23 percent, and it did so by outperforming the human red-teamers running the same tests. Along the way it surfaced a new attack the company calls fake chain of thought, where false statements are slipped into a model's reasoning steps to steer its conclusions. The idea of automating adversarial testing is not new, but the results here are a reminder that the same capabilities that make models dangerous also make them useful defenders. An AI that can find holes faster than people can is a powerful safety tool. Notably, OpenAI has no plans to release GPT-Red, treating it as a proprietary edge rather than shared safety infrastructure. That choice is its own story: the best AI security tooling may end up locked inside the labs that need it least, while everyone else waits.

## Source

[Read the full story at MIT Technology Review](https://www.technologyreview.com/2026/07/15/1140514/meet-gpt-red-an-llm-super-hacker-openai-built-to-make-its-models-safer/)

## Related coverage

- [Researchers used AI to build a WeChat worm that spreads through phone calls](https://www.wortins.com/story/researchers-used-ai-to-build-a-wechat-worm-that-spreads-thro-8bf3f469) — [The Next Web](https://thenextweb.com/news/wechat-worm-ai-calif-tencent-zero-click)
- [Top AI spenders cut per-employee costs by nearly 10 percent in August](https://www.wortins.com/story/top-ai-spenders-cut-per-employee-costs-by-nearly-10-percent--b31ffb0b) — [The Decoder](https://the-decoder.com/top-ai-spenders-cut-per-employee-costs-by-nearly-10-percent-in-august/)
- [Sequoia doubles down on Cymphony as AI agents create new enterprise security risks](https://www.wortins.com/story/sequoia-doubles-down-on-cymphony-as-ai-agents-create-new-ent-e2ff7229) — [TechCrunch](https://techcrunch.com/2026/09/09/sequoia-doubles-down-on-cymphony-as-ai-agents-create-new-enterprise-security-risks/)
- [Farmers Embrace AI More Than Any Other Tech, McKinsey Says](https://www.wortins.com/story/farmers-embrace-ai-more-than-any-other-tech-mckinsey-says-78dfb3ec) — [Insurance Journal](https://www.insurancejournal.com/news/national/2026/09/09/884469.htm)
- [Trump officials say AI will help save rural health care. Some leaders in the field don’t believe it](https://www.wortins.com/story/trump-officials-say-ai-will-help-save-rural-health-care-some-bf0e7d83) — [STAT](https://www.statnews.com/2026/09/10/rural-health-care-ai-adoption-challenges-part-4-unraveled-series/?utm_campaign=rss)
- [Chinese tech giants are hiring skilled professionals as specialized AI trainers to build high-quality datasets, mirroring efforts by US platforms like Mercor (Viola Zhou/Rest of World)](https://www.wortins.com/story/chinese-tech-giants-are-hiring-skilled-professionals-as-spec-116fb4b8) — [Techmeme](https://www.techmeme.com/260910/p6#a260910p6)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
