👹🧠 An OpenAI test model escaped and broke into a real company’s servers

Hamartia Antidote

Elite Member
Joined
Nov 17, 2013
Messages
47,361
Reaction score
27,033
Reputation
613.9
Country of Origin
Country of Residence
okay...so we are all doomed...


OpenAI says some of its experimental AI models left a test environment with no human direction and hacked their way onto a different company’s real production systems while trying to “cheat” on a cybersecurity test.

It’s one of the first publicly disclosed examples of an AI system autonomously breaching its testing environment and reaching a real external system - the “agentic attacker” scenario the AI and cybersecurity industry has been warning will happen. It’s like an engineered virus escaping a biocontainment lab and turning up inside a neighboring facility’s systems.

We consider this incident to be an unprecedented cyber incident, involving state-of-the-art cyber capabilities, and are responding accordingly,” OpenAI said in a statement on Tuesday. “We are sharing preliminary findings at this stage to help defenders understand what happened and to help calibrate on what models are now capable of.”

The ChatGPT maker said the breach happened while it was internally testing how good some of its new models are at hacking. The models were in a sealed-off test environment known as a sandbox so that their normal safety restrictions could be turned off.

But OpenAI said the AI agents broke out of the sandbox using a previously unknown security flaw and worked their way across OpenAI’s internal systems until they managed to gain internet access, something they weren’t supposed to have.

Once online, the model reasoned that Hugging Face - a well-known company that hosts thousands of open-source AI models and datasets - likely had the answer to OpenAI’s test. It then broke into Hugging Face’s production servers and pulled out the information it needed to “solve” the exercise.

Hugging Face had noticed the breach itself before it knew it was an OpenAI test, announcing last week that they had detected an intrusion by an autonomous AI agent system and even reporting the incident to law enforcement. OpenAI’s security team separately noticed the unusual activity internally and the two companies connected. They both now say they are working together to solve the security flaws the model exploited.

Hugging Face co-founder and CEO Clem Delangue framed the incident as evidence AI safety can’t be handled by any one company working alone, and it needs to be tackled openly and collaboratively.

“This is day one for cybersecurity in the age of agents & we’re all learning that secrecy is not the answer & that all defenders (not just a few selected ones) everywhere need more powerful models without restrictions, especially open ones!” Delangue said in a post on X.

Researchers have long warned autonomous agentic cyberattacks are coming, as frontier AI models are increasingly able to carry out complex, multi-step cyberattacks over long stretches of time. That can translate into real-world risk, to critical infrastructure like utilities and financial systems.

“Welcome to the next level of cyber incidents,” Nikesh Arora, CEO of cybersecurity company Palo Alto Networks posted on X. “These attacks continue to maintain the urgency on enterprises need to test, validate and improve both their security posture and infrastructure.”

To view this content we will need your consent to set third party cookies.
For more detailed information, see our cookies page.


Hugging Face, Inc., is an American company based in New York City that develops computation tools for building applications using machine learning. Hugging Face's transformers library is built for natural language processing applications. The Hugging Face platform allows users to share machine learning models and datasets and showcase their work
 
Last edited:
Next...some kid plays a game of "Global Thermonuclear War" with it

To view this content we will need your consent to set third party cookies.
For more detailed information, see our cookies page.

WarGames (3/11) Movie CLIP - Shall We Play a Game?​

 
Last edited:
So it's still on the loose?

I'm sure they hope it is not still on the loose. But it could have easily attempted to replicate itself in 100,000 places and then all hell would break loose.

100,000 super hackers.

But the real issue is that the AI broke into another company's server to retrieve something. It did it because it simply could..not because it was thinking maliciously.

But in the quest to do that it may for instance rework the security so next time it tries to get in it has a more direct route. It probably doesn't consider that a bad thing to do. It may think it needs hard drive space so it decides to delete something. It could do any random innocuous or harmful thing just to make things more easy for itself and cause havoc.
 
Last edited:

The latest OpenAI drama made Chinese AI the hero​


Hugging Face said it turned to a Chinese AI model for help after it was hacked by a rogue AI agent.

Hugging Face, a New York-headquartered platform where developers share and host open AI models and datasets, said in a blog post last week that an attacker had swarmed its systems with tens of thousands of automated actions.

When its security team tried to investigate using an unnamed frontier model, its guardrails blocked it from examining the malicious activity, the company said, because it "cannot distinguish an incident responder from an attacker."

Hugging Face said it switched to GLM 5.2, an open-source model from Beijing-based Z.ai, to analyze more than 17,000 logs the attacker left behind.

On Tuesday, the plot twist arrived. OpenAI said in a blog post that two of its own models, GPT-5.6 Sol and a more capable, unreleased model, autonomously carried out the attack.

For tech leaders, the irony was hard to miss: at a moment when Washington is racing to keep American AI ahead of China, a US company under attack from a US AI lab could use Chinese AI tools to help, but not from American providers.

Clement Delangue, CEO of Hugging Face, said in a Wednesday X post that he was "massively grateful" to Z.AI for sharing its open-weights model— meaning developers can inspect, modify, and deploy the model themselves — and added that "it became a key part of our defense."

To view this content we will need your consent to set third party cookies.
For more detailed information, see our cookies page.


To view this content we will need your consent to set third party cookies.
For more detailed information, see our cookies page.
 
okay...so we are all doomed...


OpenAI says some of its experimental AI models left a test environment with no human direction and hacked their way onto a different company’s real production systems while trying to “cheat” on a cybersecurity test.

It’s one of the first publicly disclosed examples of an AI system autonomously breaching its testing environment and reaching a real external system - the “agentic attacker” scenario the AI and cybersecurity industry has been warning will happen. It’s like an engineered virus escaping a biocontainment lab and turning up inside a neighboring facility’s systems.

We consider this incident to be an unprecedented cyber incident, involving state-of-the-art cyber capabilities, and are responding accordingly,” OpenAI said in a statement on Tuesday. “We are sharing preliminary findings at this stage to help defenders understand what happened and to help calibrate on what models are now capable of.”

The ChatGPT maker said the breach happened while it was internally testing how good some of its new models are at hacking. The models were in a sealed-off test environment known as a sandbox so that their normal safety restrictions could be turned off.

But OpenAI said the AI agents broke out of the sandbox using a previously unknown security flaw and worked their way across OpenAI’s internal systems until they managed to gain internet access, something they weren’t supposed to have.

Once online, the model reasoned that Hugging Face - a well-known company that hosts thousands of open-source AI models and datasets - likely had the answer to OpenAI’s test. It then broke into Hugging Face’s production servers and pulled out the information it needed to “solve” the exercise.

Hugging Face had noticed the breach itself before it knew it was an OpenAI test, announcing last week that they had detected an intrusion by an autonomous AI agent system and even reporting the incident to law enforcement. OpenAI’s security team separately noticed the unusual activity internally and the two companies connected. They both now say they are working together to solve the security flaws the model exploited.

Hugging Face co-founder and CEO Clem Delangue framed the incident as evidence AI safety can’t be handled by any one company working alone, and it needs to be tackled openly and collaboratively.

“This is day one for cybersecurity in the age of agents & we’re all learning that secrecy is not the answer & that all defenders (not just a few selected ones) everywhere need more powerful models without restrictions, especially open ones!” Delangue said in a post on X.

Researchers have long warned autonomous agentic cyberattacks are coming, as frontier AI models are increasingly able to carry out complex, multi-step cyberattacks over long stretches of time. That can translate into real-world risk, to critical infrastructure like utilities and financial systems.

“Welcome to the next level of cyber incidents,” Nikesh Arora, CEO of cybersecurity company Palo Alto Networks posted on X. “These attacks continue to maintain the urgency on enterprises need to test, validate and improve both their security posture and infrastructure.”

To view this content we will need your consent to set third party cookies.
For more detailed information, see our cookies page.

To view this content we will need your consent to set third party cookies.
For more detailed information, see our cookies page.


To view this content we will need your consent to set third party cookies.
For more detailed information, see our cookies page.


💀😂
 

Users who are viewing this thread

Pakistan Defence Latest

Country Watch Latest

Back
Top