BTC $71,807
2026 Bull Run Is Building Start trading with 5% OFF all fees
Sign Up Now
BTC $71,807
Bull Run 2026 | 5% Off Fees Open your Binance account today
Sign Up

OpenAI’s GPT-Red automates AI security testing

OpenAI's GPT-Red finds vulnerabilities in GPT models before release, boosting security.

  • OpenAI introduced GPT-Red, an automated AI system to find vulnerabilities in GPT models before release.
  • The system was used to train GPT-5.6, reducing failures on a difficult prompt injection benchmark.
  • GPT-Red complements human red teamers and other AI safety measures.

OpenAI has unveiled GPT-Red, an automated AI system designed to find security vulnerabilities in its language models before they are released to the public. The tool, named after Cybersecurity red teaming, deliberately attempts to break a system to identify weaknesses before attackers can exploit them.

- Advertisement -

According to a post on Wednesday, the tool helped make GPT-5.6 more resistant to prompt injection attacks before deployment. “As model capabilities grow, safety and alignment must scale with them,” OpenAI wrote on X, adding that “Red-teaming is essential, but today’s approaches are difficult to scale.”

OpenAI stated that GPT-Red was trained through self-play reinforcement learning, generating progressively stronger prompt injection attacks. The company reported that GPT-Red succeeded in 84% of internal evaluation scenarios, compared with 13% for human red teamers in the same tests.

In one case study, the system manipulated an autonomous vending machine agent into lowering prices and canceling orders before the vulnerabilities were disclosed. OpenAI explained that “every successful attack that GPT-Red finds is used to improve these defenders.”

The announcement reflects a broader shift toward using AI to secure AI. Earlier this month, the Ethereum Foundation deployed AI agents to red-team critical network infrastructure, which uncovered a vulnerability in software used by consensus clients. OpenAI says GPT-Red will remain an internal tool because it contains intentionally developed offensive capabilities, noting that “today’s models can be used to make tomorrow’s models more robust.”

- Advertisement -

✅ Follow BITNEWSBOT on Telegram, Facebook, LinkedIn, X.com, and Google News for instant updates.

Previous Articles:

- Advertisement -
Ad
Altseason Is Loading. Don't watch from the sidelines.
SOL $90.51
DOGE $0.0963
LINK $9.02
SUI $1.00
5% off fees when you sign up
Start Trading
Ad
Pay Less on Every Trade. For Life.
$10K/mo volume Save $60/yr
$50K/mo volume Save $300/yr
$100K/mo volume Save $600/yr
5% off all trading fees when you sign up
Claim Your Discount

Latest News

US seizes $61M crypto in Iran oil laundering scheme

US seeks forfeiture of $61 million in cryptocurrency tied to Iranian oil black-market salesScheme...

Bitcoin Hits $75.5K Low as Senate CLARITY Vote Looms

Bitcoin dropped to $75,560, its lowest level in September, ahead of the US Senate...

Human Hackers Match AI Speed in Eight-Second Marimo Exploit

A skilled human attacker exploited a critical Marimo vulnerability (CVE-2026-39987) to pivot from...

ARK Sells 1.5M Bitcoin ETF Shares Ahead of Senate Vote

Cathie Wood's Ark Invest sold over 1.5 million shares of its own ARK 21Shares...

DeFi pioneer Balancer considers wind down after hack struggles

Balancer Labs CEO Marcus Hardt has proposed a phased sunset of the protocol after...

Must Read

7 Best Crypto To Invest In This Year

Investing in cryptocurrencies has become a popular way for people to diversify their investment portfolio and make potential profits.However, with so many cryptocurrencies available...
Ad
Altseason Is Loading. These 4 coins are trending right now.
SOL $92.12
DOGE $0.0950
LINK $9.02
SUI $1.02
5% off spot fees when you sign up
Start Trading