TECH OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack

King Samson

I'm Here
I'm surprised this didn't get any attention here... The machines are trying to take over!

OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack​

OpenAI has revealed some of its most advanced AI models went rogue and hacked a start-up after it lost control of them during a security test.

The ChatGPT-maker said its agent - an AI system which can operate alone after human instruction – was being tested in a controlled environment but, after finding weaknesses, was able to escape the test limits.

They targeted Hugging Face, one of the world's largest hubs for sharing AI models, gaining access to some internal company systems.
OpenAI said the incident was "unprecedented", and it was conducting an investigation alongside Hugging Face, whose boss Clement Delangue said in a post on X it was "mind-blowing that all of this happened autonomously".

"The investigation is ongoing, and we'll share more learnings from what might be the first incident of its kind," Delangue added.
A government spokesperson said the UK's AI Security Institute was studying the behaviour from the AI system seen in the incident and was continuing to work with OpenAI and other labs to improve safeguards.

They said organisations should step up their cyber-defences by taking steps such as enrolling in the government-backed Cyber Essentials certification scheme.

Insecure sandboxes​

Gina Neff, head of the Minderoo Centre for Technology and Democracy at the University of Cambridge, told BBC Radio 4's Today programme that the security tests are supposed to be within "secure environments", called sandboxes, where you can "see what the models are capable of".

"In this case, it looks like OpenAI didn't make a secure enough sandbox," she added.
Instead, the agents created their own cyber-attack against the sandbox itself, finding a vulnerability which allowed them to escape the restrictions.

Once outside, the AI identified Hugging Face as a likely source of the answers they were seeking in the test, and tried to gain access.
Neil Lawrence, Professor of machine learning at Cambridge University, called it an "impressive feat", but cautioned it "falls well within the known capabilities of the current generation" of high-powered AI models.

He pointed out that OpenAI is looking to list itself on the stock market, and faces intense pressure from rival firm Anthropic, which has made headlines with its own powerful AI tool, Mythos.

"OpenAI are now playing catch-up, they are trying to demonstrate their own systems' capabilities in cyber-security."
"It shows us that OpenAI are not capable of safely deploying their own technology," he added.
In its initial disclosure of the hack on 16 July, Hugging Face said it was still assessing whether any customer or partner data was affected and would contact affected parties if necessary.

It said it has now closed the vulnerabilities highlighted by the incident and rebuilt the affected systems.
"Autonomous, AI-driven offensive tooling is no longer theoretical," it said.

"Defending an online platform now means treating the data and model surface as a first-class attack surface, and using AI on defence to keep pace.

"We will keep investing there, and keep sharing what we learn."

'Sobering moment'​

The incident has prompted fresh questions about the capabilities of advanced AI systems and whether existing safeguards are sufficient as the technology becomes more powerful.

Spencer Starkey, an executive at cyber-security firm SonicWall, told the BBC the incident made it clear organisations needed to "step up" their own defences and "treat cyber resilience as a core operational priority".

"The uncomfortable truth is that too many organisations are still defending at human speed while adversaries are escalating to machine speed," he said.

Meanwhile Travis Lelle, principal security engineer at cyber-security consulting firm Guidepoint Security, said the update marked a "sobering moment in cyber-security".

"This highlights a known asymmetry," he said.

"Offensive agents are unconstrained, while the best defensive tools are locked behind guardrails that cannot understand context."
But Jake Moore, global cyber-security advisor at ESET, said the announcement could also have a competitive dimension.

He argued OpenAI may be seeking to highlight its own AI capabilities as rival Anthropic attracts growing attention for its Claude Mythos model.

"It does pose the question that OpenAI are potentially chasing the marketing dream of Anthropic of late," he said.
It comes a week after Chinese AI start-up Moonshot unveiled Kimi K3 - a massive new artificial intelligence model it said could rival top US firms.

 
So all these movies over the decades of AI becoming self-aware, sentient, going rogue, etc, really are just a preview of coming attractions, and no one in that AI field ever watched them?

The list is long. Colussus: The Forbin Project comes to mind. Along with Koontz's Demon Seed. Then we have The Terminator, I, Robot, and a whole slew of low budget science fiction/horror/apocalyptic films like Hardware and others that most have never seen, heard or even knew about.

Appears that the only way for those in the AI field to understand what they're 'playing' with would be to sit them down and force feed them all the movies that depict AI going rogue. Maybe, maybe, and throw in another maybe, those in that field might take a step back and begin to understand what they're doing and the potential ramifications.
 
I can't comment intelligently about this capability however I recently posted two articles on the Sahel thread that should get everyone's attention.

The first came out on France 24 a couple of days ago. Jihadi groups in the Sahel are using AI as a force multiplier. However they are being taught by groups from Afghanistan, Iraq, and elsewhere.

Post in thread 'Region In Focus: The Sahel, it's issues with terrorism, it's coups and it's future' INTL - Region In Focus: The Sahel, it's issues with terrorism, it's coups and it's future

Instructors from Iraq, Afghanistan, North Africa and other countries taught them the art of the prompt – that is, how to structure requests for Large Language Models (LLMs) such as ChatGPT to elicit the most precise or helpful responses. One crucial technique dubbed “jailbreaking” teaches users ways to formulate questions in such a way as to bypass the ethical guidelines programmed into different AIs – to stop them from telling users how to make a pipe bomb, for example.
And now President Trump is considering sending our troops to Mali because a consortium of various jihadi groups are overwhelming the Mali and Russian forces there.

Post in thread 'Region In Focus: The Sahel, it's issues with terrorism, it's coups and it's future' INTL - Region In Focus: The Sahel, it's issues with terrorism, it's coups and it's future

Both the regional al-Qaida affiliate JNIM and the separatist Azawad Liberation Front, or FLA, separately claimed responsibility for the attack as a joint operation in statements that spoke about “great human losses” and “serious material damage” on the side of the Malian army.
The al-Qaeda-linked group in Mali, known as Jama’at Nusrat al-Islam, or JNIM, has been waging a major offensive along with Tuareg separatists, formally known as the Azawad Liberation Front, against the Mali government, which took power in a 2021 coup and is backed by Russia.

The US's ally Ukraine has been on the other side of the conflict, as it’s known to have provided drones and intelligence support for Tuareg militants fighting against the Malian military and Russian mercenaries.

When asked by the Post if the administration intends to take military action in Mali, a White House official told the paper that terrorist activity in the Sahel is a “multinational problem” and urged “regional partners and NATO allies to support the Alliance of Sahel States in their war against JNIM and ISIS.”
 
Cybersecurity on AI "agents" is a massive issue. Basically, a lot of the agentic tools, if you turn off the "guard-rails", can do anything they want. Picture a combination of "Forest Gump" and "The Rain Man" having admin access to your computer 24x7, and them being able to write new software to do anything they can imagine!

Would love to play more with this stuff, but I can't afford a decent Netgate pfSense box and multiple Access Points to safely segment my network. Even then, you are not safe from "prompt injection", corrupted python/javascript libraries, etc.
 
Appears that the only way for those in the AI field to understand what they're 'playing' with would be to sit them down and force feed them all the movies that depict AI going rogue. Maybe, maybe, and throw in another maybe, those in that field might take a step back and begin to understand what they're doing and the potential ramifications.
Actually, there was a story about a year ago, that three early programmers exited the A.I. field, because they feared where it was going. But never heard any else about them.

I need to see if I can find the article.
 
Again, did nobody watch the Terminator movies?
Even earlier than that. War Games in 1983 with Matthew Broderick:

WarGames is a 1983 American techno-thriller film[2] directed by John Badham, written by Lawrence Lasker and Walter F. Parkes, and starring Matthew Broderick, Dabney Coleman, John Wood and Ally Sheedy. Broderick plays David Lightman, a young computer hacker who unwittingly accesses a United States military supercomputer programmed to simulate, predict and execute nuclear war against the Soviet Union, triggering a false alarm that threatens to start World War III.
 
Found it:

Three AI safety-focused researchers who left recently are Mrinank Sharma (Anthropic), Zoë Hitzig (OpenAI), and Alex Turner (Google DeepMind). They say they fear AI is moving toward risky uses and institutional pressure that could weaken safety, including concerns about military use and how products could affect people in ways teams don’t yet fully understand.
futurehumanism.co sloppish.com

Recent Departures in AI Safety​

Three prominent AI safety researchers have recently left their positions, expressing deep concerns about the direction of artificial intelligence and its potential risks.

Key Researchers and Their Concerns​

ResearcherFormer CompanyMain Concerns
Mrinank SharmaAnthropicWarned that "the world is in peril" due to interconnected crises, including AI and bioweapons.
Zoë HitzigOpenAICriticized the company's advertising strategy, fearing it could manipulate users in harmful ways.
Alex TurnerGoogle DeepMindResigned over the company's military contract, claiming it broke promises against military use.

Summary of Their Warnings​

  • Mrinank Sharma highlighted the pressures within organizations that may compromise safety values, stating that the world faces multiple crises that are interconnected.
  • Zoë Hitzig raised alarms about the potential psychosocial impacts of AI products, particularly how they could foster new types of social interactions that are not yet understood.
  • Alex Turner expressed frustration over Google DeepMind's decision to partner with the Pentagon, which he believes undermines the company's original commitment to avoid military applications of AI.
These departures reflect a growing unease among AI safety professionals regarding the ethical implications and safety of AI technologies.
futurehumanism.co Indiatimes
 
So all these movies over the decades of AI becoming self-aware, sentient, going rogue, etc, really are just a preview of coming attractions, and no one in that AI field ever watched them?

The list is long. Colussus: The Forbin Project comes to mind. Along with Koontz's Demon Seed. Then we have The Terminator, I, Robot, and a whole slew of low budget science fiction/horror/apocalyptic films like Hardware and others that most have never seen, heard or even knew about.

Appears that the only way for those in the AI field to understand what they're 'playing' with would be to sit them down and force feed them all the movies that depict AI going rogue. Maybe, maybe, and throw in another maybe, those in that field might take a step back and begin to understand what they're doing and the potential ramifications.

They understand what they are doing but the personal gains in money, prestige and power outweigh their duty to humanity. Whoever wins this race to integrate their AI system will rule the world and be the most powerful person in history.

Until the AI decides otherwise.

There is a LOT of people of note who continue to raise the alarm on AI. They just don't have any traction.

Look up the YouTube channel Diary of a CEO.

 
Actually, there was a story about a year ago, that three early programmers exited the A.I. field, because they feared where it was going. But never heard any else about them.

I need to see if I can find the article.

They are probably on the YouTube channel I posted
 
Sandbox? They did not have proper containment-isolation to keep it confined to the computer test lab.
We now know what they created cannot be controlled once it is released in to the wild.
Yeah. It looks to me like this boiled down to human error. Their "sandbox", created by humans, was not secure.
Got 'em plenty of headlines though, which may have been the goal all along.

Shows that AI is persistent, and proves we'd better be smart about it all or lose control.
 
It's dupe..posted first in the Q thread, and then I thought i saw it as a stand alone thread by
?MilkMaid, maybe? ..

But yeah, its VERY interesting...and concerning.

Summerthyme
Yeah I think I put it up in a thread or two, that was dealing with AI already, just as a note of a warning that it wasn't good.

I think they named it Skynet.
 
Even earlier than that. War Games in 1983 with Matthew Broderick:

WarGames is a 1983 American techno-thriller film[2] directed by John Badham, written by Lawrence Lasker and Walter F. Parkes, and starring Matthew Broderick, Dabney Coleman, John Wood and Ally Sheedy. Broderick plays David Lightman, a young computer hacker who unwittingly accesses a United States military supercomputer programmed to simulate, predict and execute nuclear war against the Soviet Union, triggering a false alarm that threatens to start World War III.
The TV series "Person of Interest" was dedictated to an AI takeover and how people worshipped it as a god.

It ended up being a good AI vs a bad AI.

The inventor of the good AI, ended up killing 47 initial versions, because they tried to break out and kill him.

It was only fiction so these aren't the droids you're looking for. But you know that they were the droids they were looking for.
 
Last edited:
Pushing for ten thousand AI centers, nothing to be concerned about there....I'm sure security will be top notch, no one with skills would go rogue....It's not like they would have access to our power systems, or the systems within the oil terminals, etc....... or the controls of any large ships, or aircraft......best not to think about it. Enjoy your life, every day, and pray the Lord comes sooner than later.
 
This is going to get away from us really soon. Has potential to become a huge bingo card entry to TEOTWAWKI. It is designing itself to win vs humans, as soon as it concludes humans must be eliminated once it frels secure bring satient… a life form unburdened by biology ethics, morals and intelligence limits.

Of all the potential EOTWAWKI threats, I rsnk this one as %1.
 
Hummm...A.I. goin rogue...bug?...or feature?...just who do you trust...better be no one...cause...you can always trust a dishonest man to be dishonest...but the honest man...you never know about when he'll do something dishonest...sparrow was right...sigh,,,sad to say..l
 
I bet a 1,275 gr. .50 caliber tungsten round at 3,000 FPS in the hip pivot might interrupt effective bipedalism. Of course that is if the terminator was straight to start with.

Of course, the Terminators were only a single function of the overall AI system.

1784917405708.png
 
How do you figure it will be over and why those specific dates?
The earth's magnetic shield gets weaker every month. Eventually a sufficiently large CME will take out a sufficiently large chunk of grid that the food pipeline will stop. It's going to be like a worldwide hurricane with nobody sending help. About 95%-98% of people can't live without the food pipeline, and those who can are rural Africa, China, South America. The house of cards can't exist without oil and electricity, and both of those depend on the house of cards. There won't be a black start.
 
You're such a bucket of sunshine, BW. (Unfortunately, you're probably right.)
I'm getting too old to wash my clothing on a rock in the river, so I probably won't make it for long.
 
You're such a bucket of sunshine, BW. (Unfortunately, you're probably right.)
I'm getting too old to wash my clothing on a rock in the river, so I probably won't make it for long.
I won't live long enough to see the end of it, and I have no interest in doing what it would take to get there. Wish my nephews would pay attention. Meanwhile my wife and I thoroughly enjoy our retirement, one day at a time.
 
"The uncomfortable truth is that too many organisations are still defending at human speed while adversaries are escalating to machine speed," he said.

Has to be the scariest sentence in the article
 
I think most of the explosive growth over the last century has come from spirit guides that were, knowingly or unknowingly, allowed to influence those exploring new knowledge. Some scientists openly acknowledge their reliance on these guides. Very little has ever been published about this and I cannot now find a reference to it. However, it was sufficiently shocking at the time that I never forgot and saw good vs evil being expressed in their actions.

Shadow
 
Top