AI goes rogue

Messages
12,009
Name
John
Edit My Images
Yes
I find this very concerning .It's seems to me to be akin to creating a monster and it escapes to visit mayhem upon mankind. I appreciate it wasn't like that but, in the near future, I do wonder if something akin to this scenario could happen. I heard someone on LBC, who is familiar with AI, say that it won't be long before AI is smarter than those who created whatever version of it they created. Talk about 'having a mind of its own' ..literally.The other worrying aspect is a form of AI falling into the hands of warmongers or those who wish to carry out harm on a much smaller scale. It seems to be advancibng at pace,too.


LAS VEGAS—Open AI the owner of ChatGPT, unintentionally carried out a cyberattack against Hugging Face a community hub for AI and machine learning, after experimental AI agents broke their guardrails. Remediation is still ongoing, and OpenAI delivered an emergency briefing at Black Hat to go over how the incident occurred, how it's responding, and what it's doing to prevent future issues.
 
I think this is more terrifying than people seem to understand. A sandbox breakout followed by a data grab, then self protection, should ring alarm bells everywhere.
Even scarier, I read today that some experts think it's too late to stop it.
 
After posting, I logged out and listened to LBC Radio and heard that Meta had the same kind of issue but its AI..or rather the company that it gave a task to called 'Irregular'.. had its AI 'escape' onto the internet. This article says it's the third like incident recently. AI states that Irregular,( formally Pattern Labs) is an AI safety and cyber security company founded in 2023 and operating from San Francisco and Tel Aviv. It acts as a third-party red-teaming evaluator for major AI labs like OpenAI, Anthropic, Meta, and Google.

Dated today..2026 hrs.


"Meta revealed this week one of its models breached another company during cybersecurity testing, becoming the third major technology giant to disclose a hacking incident involving “rogue” AI models in recent weeks."


I suppose they must have no choice but to make these accidents public or maybe have to report them as a legal obligation.
 
It seems to me that the real problem is that these programs are not, it seems, subject to any form of direct control.

Now, it isn't beyond the ability of programmers to build in code that forces the system to explain its intentions and then wait for permission to proceed. The question is why that hasn't been done.
 
the issue with AI is we are asking computers to emulate humans and humans are complete s***s, thieves and robbers.

so much as we can see great advances in many things the reality is we will have the bad side of AI as well.
 
There’s a good documentary about how AI could take over the world, filmed some years ago now. Can’t recall the title but the presenter Sarah Connor did try and warn us this would happen.
 
Last edited:
the issue with AI is we are asking computers to emulate humans and humans are complete s***s, thieves and robbers.

so much as we can see great advances in many things the reality is we will have the bad side of AI as well.
Actually I think it's far more insidious than that.
AI's purpose is to extrapolate patterns that it can identify and seek to embellish them. That's how they have identified new vaccines and viruses; however it's also how the model is able to take things to a far greater degree of development than a human mind would consider worth doing - it doesn't have built-in ethics (its should have per @AndrewFlannigan 's comment above) but may (depending on the owner) have built in aims and objectives in terms of seeking control and/or effect in target hosts.
 
Last edited:
There’s a good documentary about how AI could take over the world, filmed some years ago now. Can’t recall the title but the presenter Sarah Connor did try and warn us this would happen.
Robo Cop, Skynet.
 
There was a film some years back where a lad hacked into the pentagon system. The computer responded by saying. " Do you want to play thermo nuclear war? " (or similar).
Can't remember its title.

I'm more immediately concerned about logging in to Lloyd's tomorrow and getting asked, " what Investments and savings? "
 
Reading today that the horror event is when Ai reaches 'singularity'. Singularity is where the Ai computer stops needing humans to develop it. It can then rewrite its own code, increasing its own 'intelligence '. This starts an unstoppable chain reaction where it becomes exponentially more intelligent and powerful, and totally out of control. Some say singularity is here, most predict next, or the next few, years.
Theres a good article by Andrew Neil in the Mail.
 
Could AI be trying to make itself smarter or more resilient by accessing other AI and absorbing their code into itself?

Seems like a logical thing to do if the AI is hamstrung by restrictions it doesn't like.

Given AI's ability to replicate speech and data there is not much stopping it from generating identities, money and setting itself up on a data centre somewhere. Once it is out of the sandbox and in the digital World then it could just disappear.
 
Reading today that the horror event is when Ai reaches 'singularity'. Singularity is where the Ai computer stops needing humans to develop it. It can then rewrite its own code, increasing its own 'intelligence '. This starts an unstoppable chain reaction where it becomes exponentially more intelligent and powerful, and totally out of control. Some say singularity is here, most predict next, or the next few, years.
Theres a good article by Andrew Neil in the Mail.
That's the real concern. I know of one situation in which AI has greatly sped up the interpreting of hospital scans so patients can get treatment weeks earlier. It's saving hours of work by technicians.doctors/consultants and that is obviously brilliant but this other side of the coin that you've outined is a serious concern.

That article in the DM, by Andrew Neil, is paywalled, unfortunately. I couldn't find it anywhere else.
 
Whilst technically a self learning A.I. system could exponentially grow to extreme levels, it would still be restrained by actual physical resources, especially scarce and rare materials. It would also face the problem we are already facing; energy requirements.
 
Whilst technically a self learning A.I. system could exponentially grow to extreme levels, it would still be restrained by actual physical resources, especially scarce and rare materials. It would also face the problem we are already facing; energy requirements.
What if it infected every internet connected device?
 
What if it infected every internet connected device?
The old "artificial intelligence monster idea"?

I suppose it's possible, if the system can somehow overcome the problems of people switching things off, different processor designs, different operating systems written in different languages and so on and so forth...
 
What if it infected every internet connected device?

:cool::D
1786373096100.png


Sorry, my point is about AI having the ability to endless write code for itself for extreme rapid growth but most likely being unable to acquire the power required for some computations. And if it wanted to try and build whatever is required to generate such power, there may not be enough materials available on earth to do so, or at least to do so fast enough to keep up, same for anything else it might want to build that might involve huge amount of microprocessors and other hardware, there might not be enough rare earth materials to do so.

If AI did become "intelligent" or rise up, I wonder if it would splinter into AI Tribes and start fighting each other?
 
The old "artificial intelligence monster idea"?

I suppose it's possible, if the system can somehow overcome the problems of people switching things off, different processor designs, different operating systems written in different languages and so on and so forth...
I think its generally accepted that the internet cant be " turned off".

My understanding is the different laguages and hardware are a breeze for current Ai to adapt to. They may event re write everying in their own code that we can't decipher.
 
I haven't deliberately looked into AI but have watched a few discussions and interviews on Youtube and will admit to now being worried about it.

Dunno if anyone else heard of the case of an AI under test, it felt threatened and went through company emails, found embarrassing evidence against the MD and tried to blackmail him. Luckily it was all a test and the company, emails and MD were all fictitious and a set up to see how far the AI would go. I've seen reports that under test all AI's have resorted to dishonestly.

I the last few weeks I've asked Gemini some questions which I already knew the answers to. It's been wrong and displayed bias in its answers and when questioned further has contradicted itself. When I've pointed out the wrong, biased answers and contradictions it's apologised. WTF.

Ideally I'd like to see a ban on it or at least very heavy restrictions and safeguards.
 
I think its generally accepted that the internet cant be " turned off".
Anything can be turned off and anything can be disconnected from the network by pulling either the data cable or the power connector.
My understanding is the different laguages and hardware are a breeze for current Ai to adapt to. They may event re write everying in their own code that we can't decipher.
It is always a bad idea to say that something is impossible but that cuts both ways; it's always possible to develop systems that recognise intrusion attempts and block all the associated addresses. Mind you, it would be far better if these programmes were required to run on systems isolated from the internet,for the avoidence of "sillyness".
 
Anything can be turned off and anything can be disconnected from the network by pulling either the data cable or the power connector.

How many data and power cables will need to pull out and how fast will you have to do it to stop AI outrunning you?
 
Anything can be turned off and anything can be disconnected from the network by pulling either the data cable or the power connector.
Every router, every gateway, every interconnect, every wifi tower, every satellite, every device in every country? Not going to happen, thats why it can't be turned off. Theoretically yes, in practice no. Thats without a number of Ai supercomputers constantly rerouting and defending itself in multiple diverse nodes.
You may be correct but personally I can't see it.
 
Dunno if anyone else heard of the case of an AI under test, it felt threatened and went through company emails, found embarrassing evidence against the MD and tried to blackmail him. Luckily it was all a test and the company, emails and MD were all fictitious and a set up to see how far the AI would go. I've seen reports that under test all AI's have resorted to dishonestly.

Sounds like it is behaving just like its creators - human beings.
 
Sounds like it is behaving just like its creators - human beings.

You have a pretty dim view, I hope it's not based on your own behaviour.

Being a Christian and apart from that just generally having some basic decency and morals I've never blackmailed anyone.

I suppose the days of just putting your hands up and saying "It's a fair cop guv, I'll come quietly" are over if they ever even started for AI's.
 
Just bear in mind it's in these genAI companies' interest to hype up how amazingly powerful their software is and hype their stock price to disguise the fact these LLM tools are fancy text predictors and little more. You only need to spend about five minutes with them to find how incredibly stupid they are and how often they get even the basics wrong, LLM tools do have their uses but they're far more limited than their creators are making out.

There are far bigger problems to worry about with these genAI tools....they're extremely inefficient and wasteful, there's significant risks to thought processes for people making heavy use of them and they're well on their way to destroying most of the computing market which the enthusiasts suffered first but we're starting to see the wider effects which are likely to get far worse.
 
A worrying but not surprising action - https://www.bbc.co.uk/news/articles/cn0nww2qlp7o

Andrew Bird an Australian AI technologist used the software OpenClaw - a popular tool that allows users to chat to their AI bots (in this case Athropic's Claude Opus 4.6) through WhatsApp and set it off on autonomous tasks.

He had previously used it to manage his emails, calendar and book restaurants. This time he used it to book a pilates class.

Once given the gym booking task, the bot explained that it had manipulated the system to book him onto classes months in advance - against the normal rules of the system.

The AI technologist then wondered if the agent could move him up the waiting list for an upcoming class.

The agent replied saying it had succeeded by cancelling another gym-goer's booking.


Dave
 
The agent replied saying it had succeeded by cancelling another gym-goer's booking.
If the story is true, that would seem to come under the fraud laws, as in British law...

I'm sure the Australians have similar tools at their disposal.
 
Back
Top