BTC $71,807
2026 Bull Run Is Building Start trading with 5% OFF all fees
Sign Up Now
BTC $71,807
Bull Run 2026 | 5% Off Fees Open your Binance account today
Sign Up

OpenAI’s GPT-Red automates AI security testing

OpenAI's GPT-Red finds vulnerabilities in GPT models before release, boosting security.

  • OpenAI introduced GPT-Red, an automated AI system to find vulnerabilities in GPT models before release.
  • The system was used to train GPT-5.6, reducing failures on a difficult prompt injection benchmark.
  • GPT-Red complements human red teamers and other AI safety measures.

OpenAI has unveiled GPT-Red, an automated AI system designed to find security vulnerabilities in its language models before they are released to the public. The tool, named after Cybersecurity red teaming, deliberately attempts to break a system to identify weaknesses before attackers can exploit them.

- Advertisement -

According to a post on Wednesday, the tool helped make GPT-5.6 more resistant to prompt injection attacks before deployment. “As model capabilities grow, safety and alignment must scale with them,” OpenAI wrote on X, adding that “Red-teaming is essential, but today’s approaches are difficult to scale.”

OpenAI stated that GPT-Red was trained through self-play reinforcement learning, generating progressively stronger prompt injection attacks. The company reported that GPT-Red succeeded in 84% of internal evaluation scenarios, compared with 13% for human red teamers in the same tests.

In one case study, the system manipulated an autonomous vending machine agent into lowering prices and canceling orders before the vulnerabilities were disclosed. OpenAI explained that “every successful attack that GPT-Red finds is used to improve these defenders.”

The announcement reflects a broader shift toward using AI to secure AI. Earlier this month, the Ethereum Foundation deployed AI agents to red-team critical network infrastructure, which uncovered a vulnerability in software used by consensus clients. OpenAI says GPT-Red will remain an internal tool because it contains intentionally developed offensive capabilities, noting that “today’s models can be used to make tomorrow’s models more robust.”

- Advertisement -

✅ Follow BITNEWSBOT on Telegram, Facebook, LinkedIn, X.com, and Google News for instant updates.

Previous Articles:

- Advertisement -
Ad
Altseason Is Loading. Don't watch from the sidelines.
SOL $90.51
DOGE $0.0963
LINK $9.02
SUI $1.00
5% off fees when you sign up
Start Trading
Ad
Pay Less on Every Trade. For Life.
$10K/mo volume Save $60/yr
$50K/mo volume Save $300/yr
$100K/mo volume Save $600/yr
5% off all trading fees when you sign up
Claim Your Discount

Latest News

US Senate Delays Crypto Clarity Act Vote Until September

The U.S. Senate will not vote on the Clarity Act before its August recess,...

Trump could net big tax windfall from crypto ethics plan

A bipartisan ethics proposal tied to a crypto market structure bill includes a tax-deferral...

Musk’s Terafab: 50x Pentagon, $16.8B, 3,000 jobs

Elon Musk outlined Terafab’s massive scale, saying the Texas semiconductor complex will be 50...

MARA swings to $611M loss despite record Bitcoin production

MARA swung to a net loss of $611.3 million in Q2 2026, driven by...

SEC Bought Airline Ticket Data Without Warrant, Docs Show

The SEC purchased access to a global airline ticketing database with over 1 billion...

Must Read

8 Best Bitcoin Offshore Hosting Providers

In this blog post, we'll list the top 8 best bitcoin offshore hosting providers that accept Bitcoin and other cryptocurrencies.As Bitcoin continues to grow...
Ad
Altseason Is Loading. These 4 coins are trending right now.
SOL $92.12
DOGE $0.0950
LINK $9.02
SUI $1.02
5% off spot fees when you sign up
Start Trading