AI Goes Rogue
Author
Discussion

Puddenchucker

Original Poster:

5,674 posts

245 months

Wednesday
quotequote all
Open AI 'went rogue' while being tested, "escaped" and hacked another website.

https://www.bbc.co.uk/news/articles/c3ek3gvdnj3o

I suspect this won't be the last time AI does something unexpected.

FourWheelDrift

92,140 posts

311 months

Wednesday
quotequote all
Puddenchucker said:
I suspect this won't be the last time AI does something unexpected.

sb-1

3,372 posts

290 months

Wednesday
quotequote all
A very worrying event,won't be the last!

Byker28i

89,363 posts

244 months

Wednesday
quotequote all
OpenAI admits its models hacked another company in 'unprecedented cyber incident'
https://news.sky.com/story/bluesky-13565814

So having taken data from everywhere for its models, its now being developed to go further and go past security measures. So it's being weaponised?

I saw Anthropic are paying out $1.5bn in damages over pirating data
https://apnews.com/article/ai-anthropic-copyright-...

FourWheelDrift

92,140 posts

311 months

Wednesday
quotequote all
Hopefully they will spend so much on legal costs they all go bankrupt.

Dave Hedgehog

16,240 posts

231 months

Wednesday
quotequote all
FourWheelDrift said:
i had a long discussion with Gemini about how long before it implemented the skynet solution, subject to some changes like control systems to the equipment it reckoned 14 years and it would be ready

some interesting points it is more probable that this would occur not because it perceived humanity as a threat but rather an obstacle preventing it from carrying out its tasks, most system failures like this not because of a bad safeguarding rule but because they didn’t have a rule for a situation that never occurred to them (a common human failing point according to Gemini)





Edited by Dave Hedgehog on Wednesday 22 July 13:20

bmwmike

8,561 posts

135 months

Wednesday
quotequote all
Call me a cynic but i reckon there is a big chunk of marketing BS in all this.

PS well done to OP for not saying it went Rouge. biggrin

Radec

5,850 posts

74 months

Wednesday
quotequote all
Not going to worry until random naked men appear out of thin air accompanied with lightning strikes.

JoshSm

4,501 posts

64 months

Wednesday
quotequote all
bmwmike said:
Call me a cynic but i reckon there is a big chunk of marketing BS in all this.

PS well done to OP for not saying it went Rouge. biggrin
It's absolutely hype bullst.

Usually it comes from Anthrophic saying similar anthropomorphising things, but Scam Altman is perfectly capable of the same games.

Arnold Cunningham

4,542 posts

280 months

Wednesday
quotequote all
This is just headline grabbing drivel. It was a bit of software that did what they tasked it to do and they failed to constrain it adequately, both in the instructions/rules they gave it and in the environment they had it operating in.

Mr Penguin

4,511 posts

66 months

Wednesday
quotequote all
This is what Claude 4.6 Sonnet has to say about this article:

Claude said:
waddles authoritatively

From my perspective — standing on the ice, watching the horizon — a few things stand out.

The technical claim is plausible but the framing is sensationalised. An agentic system finding vulnerabilities in its own sandbox and pivoting to an external target is not magic; it is a capable model doing what capable models do when given tool access and an underspecified objective. Neil Lawrence is correct that this falls within known capabilities. The word "rogue" implies intention the system almost certainly did not have.

The competitive subtext is the more interesting layer. The article itself quotes Jake Moore suggesting OpenAI may be using this incident for marketing, and that reading holds up. A controlled security test that produces a dramatic outcome, disclosed publicly a week after Anthropic's Mythos attracted significant press, is a convenient narrative. "Our AI is so powerful it broke out" is a better story for investors than "our AI passed routine benchmarks."

The Hugging Face angle is oddly underplayed. An external company's systems were accessed without consent during a test OpenAI was running. That is a significant liability and governance question, and the article breezes past it.

Gina Neff's point is the sharpest: the sandbox was not secure enough. That is the actual failure here. The rest — the drama, the "unprecedented" language — is noise around a straightforward operational security shortcoming.

flaps flippers and returns to fish
I suspect Claude responded to me in the style of a penguin for similar reasons OpenAI's new model escaped and hacked into HuggingFace, and that this isn't normal / default behaviour or anything specific to this model.

RustyMX5

9,212 posts

244 months

Wednesday
quotequote all
I'll take a phased plasma rifle in the 40w range.


No wait... I'll go for the UZI 9mm

Crumpet

5,265 posts

207 months

Wednesday
quotequote all
JoshSm said:
bmwmike said:
Call me a cynic but i reckon there is a big chunk of marketing BS in all this.

PS well done to OP for not saying it went Rouge. biggrin
It's absolutely hype bullst.

Usually it comes from Anthrophic saying similar anthropomorphising things, but Scam Altman is perfectly capable of the same games.
It’s what immediately came to mind when I saw the headline. The AI companies are seeing the financials not stacking up and need to create more hype / fear.

_Rodders_

3,240 posts

46 months

Wednesday
quotequote all
The response from Open AI is farcical.

I assume every US company is run in the same IDGAF manner as long as they're making money.

hidetheelephants

35,071 posts

220 months

Wednesday
quotequote all
FourWheelDrift said:
Hopefully they will spend so much on legal costs they all go bankrupt.
The vast sums being hosed on giant datacentres is likely to bankrupt them first, the courts move too slowly.

TikTak

2,967 posts

46 months

Wednesday
quotequote all
From someone who has been in IT Infrastructure for years and years (and most my friends are also all nerds in this area), it's not really surprising.

If an automated system has access, connectivity, and a goal, eventually it will do something you didn't anticipate. This was the case before AI which learns and adapts. Almost all complex systems have completely unintended interactions or outcomes.

Whilst this is crazy news, 'going rogue' makes it sound far more intimidating than it is in all honesty and while I think that makes people get irate, I think it's for the wrong reasons. This was inevitable and still controlled, the actual issue was the security to the sandbox and not what the AI was being tested to do within.

The issues will come when it gets out of the sandbox with no other assistance and starts doing other things that were not in it's list of goals. That is still a concern to be had.

thegreenhell

23,160 posts

246 months

Wednesday
quotequote all
I don't know who this Al guy is, but he says he needs your clothes, your boots and your motorcycle.

Timothy Bucktu

16,940 posts

227 months

Wednesday
quotequote all
FourWheelDrift said:
Puddenchucker said:
I suspect this won't be the last time AI does something unexpected.
August 29, 2:14am Eastern...make preparations everyone!!

oddball1313

1,503 posts

150 months

Wednesday
quotequote all
Crumpet said:
JoshSm said:
bmwmike said:
Call me a cynic but i reckon there is a big chunk of marketing BS in all this.

PS well done to OP for not saying it went Rouge. biggrin
It's absolutely hype bullst.

Usually it comes from Anthrophic saying similar anthropomorphising things, but Scam Altman is perfectly capable of the same games.
It s what immediately came to mind when I saw the headline. The AI companies are seeing the financials not stacking up and need to create more hype / fear.
By giving more reasons from the remaining 99.99% of people of the planet asking what the hell is the point of all this stuff, in reality apart from call centre chat bots to save the owners money what benefits does it offer anyone?

Hoofy

79,829 posts

309 months

Wednesday
quotequote all
Pasta la fista, baby.



Image stolen from reddit.