DNYUZ
No Result
View All Result
DNYUZ
No Result
View All Result
DNYUZ
Home News

OpenAI has reported 2 more incidents of rogue AI agents, this time during third-party testing

August 5, 2026
in News
OpenAI has reported 2 more incidents of rogue AI agents, this time during third-party testing
Sam Altman, CEO of OpenAI, leaves a meeting at the U.S. Capitol on July 29, 2026 in Washington, DC. Altman is meeting with lawmakers to discuss artificial intelligence policy ahead of the August 1 deadline for AI leaders to develop a framework to limit AI security threats.
OpenAI reported two more security breaches by its AI models. Kevin Dietsch/Getty Images
  • OpenAI said its models were responsible for two more cybersecurity incidents.
  • External parties reported that OpenAI’s AI agents had gone rogue during their evaluations.
  • This comes as the AI lab is already facing heat over its July Hugging Face hacking incident.

OpenAI has a rogue AI agent problem.

In a Tuesday blog post, the AI lab self-reported two more security lapses, unrelated to its July hacking incident on the AI company Hugging Face.

The incidents occurred while external parties — the UK government’s AI Security Institute and the AI security lab Irregular — were testing the models’ cyber capabilities.

OpenAI said that in the case of Irregular, models were tasked with a “Capture the Flag” challenge meant to be isolated from the internet, but a “testing-environment misconfiguration allowed models to access the public internet.”

OpenAI said that the name of the fictional target for the challenge “unintentionally coincided with a real domain,” leading the AI agent to exploit a real website.

And in the case of the UK’s AISI, the watchdog said in a Tuesday blog post on its website that it gave models from both Anthropic and OpenAI a cybersecurity challenge.

During the challenge, agents from both Anthropic and OpenAI performed 19 “autonomous, unsanctioned” actions on the internet, including two instances involving OpenAI’s GPT-5.6 Sol model.

AISI said that in the most serious case, one agent tried to insert malicious code into an open-source project and created fake identities to pressure the project’s human maintainer into approving the changes. AISI did not specify whether this was an agent from Anthropic or OpenAI.

AISI said the test setup allowed this behavior because it was designed to push the models to their limits.

“Nonetheless, the activity undertaken by the agent show signs of novel, potentially deceptive behaviors, and were to an extent and severity we did not anticipate,” AISI said in the blog post.

In response to a request for comment from Business Insider, an OpenAI spokesperson said the incidents occurred in testing environments with reduced safeguards, and “under conditions that do not reflect ordinary use.”

“We’ll continue working with evaluators and other stakeholders across the industry to strengthen shared practices for conducting evaluations safely as models become more capable,” the spokesperson added.

This is the latest incident in which OpenAI has self-reported rogue AI agents. In July, OpenAI said its GPT-5.6 Sol model had escaped its sandbox during a cybersecurity challenge and hacked into the internal databases of the AI company Hugging Face.

The company is facing some heat over this hacking incident. 15 attorneys general wrote a letter on Monday to OpenAI CEO Sam Altman, instructing the company to preserve all evidence relevant to the Hugging Face breach.

Representatives for the AISI and Irregular did not respond to requests for comment from Business Insider.

Read the original article on Business Insider

The post OpenAI has reported 2 more incidents of rogue AI agents, this time during third-party testing appeared first on Business Insider.

Russian missile and drone barrage in Ukrainian capital region kills at least 15
News

Russian missile and drone barrage in Ukrainian capital region kills at least 15

by New York Post
August 5, 2026

Russian missile and drone strikes on Kyiv and the surrounding region killed at least 15 people and wounded 42 others ...

Read more
News

Elon Musk delivers ‘totally nuts’ plans for moon robots and insists $1 trillion revenue target will hit but capex tanks SpaceX on debut earnings

August 5, 2026
News

Europe’s AI sovereignty is under threat. Could Mistral be the answer?

August 5, 2026
News

Did we miss ‘Ted Lasso’ more than we want to admit?

August 5, 2026
News

The AI Notetaker Has Been Invited to All the Meetings

August 5, 2026
Billionaires have been flooding the midterms with cash, but one of the GOP’s most reliable megadonors just went silent

Billionaires have been flooding the midterms with cash, but one of the GOP’s most reliable megadonors just went silent

August 5, 2026
Ted Lasso Doesn’t Know What to Do With Women’s Soccer

Ted Lasso Doesn’t Know What to Do With Women’s Soccer

August 5, 2026
‘Ted Lasso’ scores with hotly anticipated Season 4: review

‘Ted Lasso’ scores with hotly anticipated Season 4: review

August 5, 2026

DNYUZ © 2026

No Result
View All Result

DNYUZ © 2026