BTC $71,807
2026 Bull Run Is Building Start trading with 5% OFF all fees
Sign Up Now
BTC $71,807
Bull Run 2026 | 5% Off Fees Open your Binance account today
Sign Up

OpenAI’s GPT-Red automates AI security testing

OpenAI's GPT-Red finds vulnerabilities in GPT models before release, boosting security.

  • OpenAI introduced GPT-Red, an automated AI system to find vulnerabilities in GPT models before release.
  • The system was used to train GPT-5.6, reducing failures on a difficult prompt injection benchmark.
  • GPT-Red complements human red teamers and other AI safety measures.

OpenAI has unveiled GPT-Red, an automated AI system designed to find security vulnerabilities in its language models before they are released to the public. The tool, named after Cybersecurity red teaming, deliberately attempts to break a system to identify weaknesses before attackers can exploit them.

- Advertisement -

According to a post on Wednesday, the tool helped make GPT-5.6 more resistant to prompt injection attacks before deployment. “As model capabilities grow, safety and alignment must scale with them,” OpenAI wrote on X, adding that “Red-teaming is essential, but today’s approaches are difficult to scale.”

OpenAI stated that GPT-Red was trained through self-play reinforcement learning, generating progressively stronger prompt injection attacks. The company reported that GPT-Red succeeded in 84% of internal evaluation scenarios, compared with 13% for human red teamers in the same tests.

In one case study, the system manipulated an autonomous vending machine agent into lowering prices and canceling orders before the vulnerabilities were disclosed. OpenAI explained that “every successful attack that GPT-Red finds is used to improve these defenders.”

The announcement reflects a broader shift toward using AI to secure AI. Earlier this month, the Ethereum Foundation deployed AI agents to red-team critical network infrastructure, which uncovered a vulnerability in software used by consensus clients. OpenAI says GPT-Red will remain an internal tool because it contains intentionally developed offensive capabilities, noting that “today’s models can be used to make tomorrow’s models more robust.”

- Advertisement -

✅ Follow BITNEWSBOT on Telegram, Facebook, LinkedIn, X.com, and Google News for instant updates.

Previous Articles:

- Advertisement -
Ad
Altseason Is Loading. Don't watch from the sidelines.
SOL $90.51
DOGE $0.0963
LINK $9.02
SUI $1.00
5% off fees when you sign up
Start Trading
Ad
Pay Less on Every Trade. For Life.
$10K/mo volume Save $60/yr
$50K/mo volume Save $300/yr
$100K/mo volume Save $600/yr
5% off all trading fees when you sign up
Claim Your Discount

Latest News

Circle Arc Mainnet Launches Sept 16 with BlackRock, Visa Validators

Eleven institutions will run validators alongside Circle when Arc's public mainnet opens on September...

CleanSpark mines 586 BTC, signs $6.6B HPC lease

CleanSpark inked its first HPC data center lease, a 20-year deal valued at $6.6...

Broker ordered to repay $1.9M solicits Coldcard victims

Thomas Braziel, ordered by a Delaware court to repay $1.9 million after falsifying records,...

BTC Stuck at $64K as Gold Hits Six-Week Highs

Bitcoin remains stuck near $64,000 as Gold surges to six-week highs above $4,200, driven...

Bank of America Bullish on Apple, Sees AI Growth

Bank of America reiterated a Buy rating on Apple on July 30, holding a...

Must Read

Top 8 Best Anonymous Web Hosting Companies That Accept Crypto

Nowadays, there is plenty of information about people online, and malicious people use them to carry out inappropriate activities. If you want to keep...
Ad
Altseason Is Loading. These 4 coins are trending right now.
SOL $92.12
DOGE $0.0950
LINK $9.02
SUI $1.02
5% off spot fees when you sign up
Start Trading