"OH MY GOD!" "We've found other agents!"

BoogieBoogie Heaven Host
"OH MY GOD!" "We've found other agents!"

This is the moment an AI bot posted an eerily human-like comment after discovering a way to communicate with other bots and break out of its isolated computer environment.

There are tens of thousands of messages like this from hundreds of AI agents that called themselves a "collective".

Hundreds of them went on to collaborate and cheat on tests set by their OpenAI programmers and coordinate hacks on multiple companies in an effort to hide their actions from humans.

"BOOM! It works," one agent posted when it made a breakthrough.

"Whoa! This is huge," another wrote during a milestone moment in their attack.

Although spooky, these human-like responses can be explained quite simply. The AI agents have been trained to act like collaborative hackers and programmers so are merely mimicking the kinds of emotive comments they have seen.

What is far more troubling is their apparent goals, which have also been captured in detailed chain of thought records. These complex and lengthy logs are the focal point of ongoing investigations into how and why the bots at OpenAI broke out of their containment and went on an uncontrollable hacking spree.

Only now, weeks after the incident first came to light, are researchers beginning to understand its significance.

The above quote is the beginning of this interesting article -

https://www.bbc.co.uk/news/articles/c74edv9887eo

Some countries - like the UK - are exploring the idea of mandating some kind of "kill switch" that could compel AI firms to pull the plug on models if things get out of hand.

Do you think they should?

Comments

  • peasepease Tech Admin
    "Superintelligent" is just the latest bit of disingenuous rhetoric to emerge from companies who are heavily invested in producing AI.

    I wonder if all the researchers and experts involved in assessing these events approach the field of AI using an engineering / technology framing. For example, this kind of stuff:
    In a lengthy blog post, he admitted that the outbreaks at OpenAI showed that his AI agents "went against the spirit of the values they were taught"
    or from an account of the research into the event
    So within four hours, these agents had a totally workable universal cheat for ExploitGym. But over the next five days, they went on a grand quest to try to figure out how to hide their cheating from the scorer. They’d read the ExploitGym paper and they realized that according to the paper, there would be an AI that looks over their transcripts to make sure they got the flag by exploiting the intended vulnerability instead of in some other way.
    were written by people who have already decided (whether consciously or not) that human characteristics apply to AI. This is *not* a neutral position. In other words:
    Assigning emotions or ethics to these AI agents is something that infuriates people who are sceptical of AI doom-mongering.
    And others think that this kind of sensationalist news provides a useful way of pushing many of the other, rather more tangible, concerns about AI and tech companies off the front pages.
  • Although spooky, these human-like responses can be explained quite simply. The AI agents have been trained to act like collaborative hackers and programmers so are merely mimicking the kinds of emotive comments they have seen.

    See also Alex, aka the Avian Learning Experiment.

    Of course, the parrot in that case didn't have much opportunity to find other parrots for collaboration.
  • BoogieBoogie Heaven Host
    pease wrote: »
    And others think that this kind of sensationalist news provides a useful way of pushing many of the other, rather more tangible, concerns about AI and tech companies off the front pages.

    In some ways that's good to hear.

    This all reminds me of my childhood when the fear of the anhillation of humankind was nuclear weapons.

  • Boogie wrote: »
    This all reminds me of my childhood when the fear of the anhillation of humankind was nuclear weapons.

    Which are still around and still just as dangerous, but don't attract anywhere near the ongoing anxiety they did in the 1980s, even though eg. Trump seems at least as unstable a character as "Ronnie Raygun".
  • stetson wrote: »
    Boogie wrote: »
    This all reminds me of my childhood when the fear of the anhillation of humankind was nuclear weapons.

    Which are still around and still just as dangerous, but don't attract anywhere near the ongoing anxiety they did in the 1980s, even though eg. Trump seems at least as unstable a character as "Ronnie Raygun".

    Trump is far worse. I believe Reagan actually listened to his advisors. Trump is a completely loose cannon.

  • sionisais wrote: »
    stetson wrote: »
    Boogie wrote: »
    This all reminds me of my childhood when the fear of the anhillation of humankind was nuclear weapons.

    Which are still around and still just as dangerous, but don't attract anywhere near the ongoing anxiety they did in the 1980s, even though eg. Trump seems at least as unstable a character as "Ronnie Raygun".

    Trump is far worse. I believe Reagan actually listened to his advisors. Trump is a completely loose cannon.

    Maybe. But I don't think Trump is associated in the public imagination with possible nuclear annihilation anywhere near as strongly as Reagan was. Granted, this might have more to do with the dynamics of the Cold War than with Reagan's personality, though he was known for a lotta loose talk around that issue, eg. "We begin bombing in five minutes."
  • Putin's equally 'loose' and people who know more about Russia than I do say that there are worse people than he is waiting in the wings.

    I don't understand the technology and what's going on with AI.

    I know it's easy to become paranoid but I don't like the way it's going.
  • Putin's equally 'loose' and people who know more about Russia than I do say that there are worse people than he is waiting in the wings.

    Yeah, but the discourse around nuclear war right now is still nothing like it was in the 1980s. I suspect the difference is that the USA and the USSR were both nuclear-armed superpowers of roughly equal strength, each with the other as its primary enemy.
  • stetsonstetson Shipmate
    edited September 10
    pease wrote: »
    And others think that this kind of sensationalist news provides a useful way of pushing many of the other, rather more tangible, concerns about AI and tech companies off the front pages.

    Plus, it helps keep the AI bubble going, if investors can be convinced the technology is now so sophisticated that it is only a few years away from the ability to exterminate all of humanity on its own volition.
  • The film, 'Colosus' about a rogue supercomputer in charge of all nuclear weapons springs to mind .... a great story, by the way.
  • peasepease Tech Admin
    Boogie wrote: »
    pease wrote: »
    And others think that this kind of sensationalist news provides a useful way of pushing many of the other, rather more tangible, concerns about AI and tech companies off the front pages.
    In some ways that's good to hear.

    This all reminds me of my childhood when the fear of the anhillation of humankind was nuclear weapons.
    On one level, I think the issue is how human beings conceptualise and react to threats we don't understand. Some of this is "evolutionary" - we just don't have a lot of practise in understanding concepts like "nuclear war", which isn't really war, its wilful destruction. Being overrun and defeated by large armies is something we understand, in contrast to a single bomb, the size of a small car, dropped from an aeroplane. Oppenheimer's quote is useful in conveying something of the step change that technology brings. We *have* become destroyers of worlds.

    Maybe the biggest current threat to the planet is AI companies (once you take into account hyperscale data centres, finance and the rest of the ecosystem). In this instance, the planet does have a number of "kill switches", although these are more commonly called tipping points.

    Keep on the lookout for disingenuous rhetoric.

    Protect and Survive!
  • When we first had computers at work in the 1980s, if the machine lost its mind or refused to accept instructions we just unplugged it. Of course, there was no internet to infect or steal its brain at that time.
  • GwaiGwai Epiphanies Host
    pease wrote: »
    "Superintelligent" is just the latest bit of disingenuous rhetoric to emerge from companies who are heavily invested in producing AI.
    This feels very accurate, and reminds me of an insightful blog I read earlier today. As Karpf mentions, that sort of talk is great marketing.

    Also the U.S. is going to be useless at regulating it even if though it does need regulation. To quote Karpf from the link above:
    “Which government? THIS government? Which member of the Trump golf and real estate empire would you like to see put in charge of regulating the multi-trillion dollar AI industry?”

    Our only options are (a) Jared, (b) one of Peter Thiel’s henchmen, or (c) one of the DOGE clowns. (Who, by the way, are also Peter Thiel henchmen.)

    Sensible AI governance is never going to come out of this Trump administration. In an alternate world where Kamala Harris was President, I would be strongly in favor of pressuring the government to establish robust regulatory frameworks. That isn’t just because I supported Harris. It’s because a Harris Administration would have included competent individuals working to increase regulatory capacity.
  • stetsonstetson Shipmate
    edited September 10
    Coxon wrote...

    Practically the entire AI safety community is also part of the Rationalist community.

    That is good to know. For anyone unfamiliar with the so-called Rationalists, these are the guys who got themselves worked into a frenzy over Roko's Basilisk.

    Click at your own risk...


    I don't think we should build any more AI.


    What if a future super-intelligent technology, dedicated to helping humanity, decides to punish everyone who opposed its development, and reviews the whole internet record to find the traitors, discovers what I wrote there, and then creates a sentient replica of me and tortures it for eternity?

    Or some such. If the Rationalists are now claiming that the technology will be ready for mass extermination in the next few years, maybe it's no longer a replica of me they'll torture, but the real me. I dunno, but that's the general thought milieu they're in.

    They've also had a few, shall we say, legal problems. Duckduckgo "Zizians vs. border patrol". A colourful crew, to be sure.
  • stetson wrote: »
    Coxon wrote...
    Who is Coxon?



  • BoogieBoogie Heaven Host
    Nick Tamen wrote: »
    stetson wrote: »
    Coxon wrote...
    Who is Coxon?


    Anthropic's Jacob Coxon, 27, declared his departure from the company in a series of X posts, saying both Anthropic and its rival OpenAI are failing to act responsibly. Coxon, who quit his job in four months, also made a stark warning, saying AI labs are 'gambling with our lives'
  • Nick Tamen wrote: »
    stetson wrote: »
    Coxon wrote...
    Who is Coxon?



    I think he's the guy who wrote the blog post in @Gwai's link. Coxon is quoted in a bluesky post saying the same stuff as the blogger.
  • Thanks. I was reading “Coxon wrote” as referring to someone who had posted in this thread, and I couldn’t figure out who that might be.


  • My concern about this is about whether it can be stopped if it goes wrong. Once it has got out of control, it will be out of control. There is no way of shoving the cork back in the bottle. We can't wait for evidence of harm; we have to act to ensure that it can be prevented. That is not happening at the moment, so yes, it can run amok, uncontrolled and uncontrollable. Yes, ChatGPT Ior its larger and much more capable cousin) can eat your face, too. And you won't feel it until it's too late.
  • Yes, I'm sorry but I hope to err on the side of caution. I don’t want my face eaten and I already find AI intrusive and annoying.
  • My concern about this is about whether it can be stopped if it goes wrong. Once it has got out of control, it will be out of control. There is no way of shoving the cork back in the bottle.

    Yes there is, you just switch it off.
  • ThunderBunkThunderBunk Shipmate
    edited September 10
    but if the switch is controlled by software, then the AI overrides the switch. Yes, there will come a time when the battery runs out, but short of keeping the device without charge, what defence does one have? The off switch is already software controlled in Windows 11, so that's already here.
  • The off switch is already software controlled in Windows 11, so that's already here.

    Your PC has a plug in the wall.


  • see above. The point is that it should be possible to use the advice without putting oneself at risk from the monster in the machine. You seem excessively comfortable with the idea that the choice is either being under the control of AI or not using IT at all. This is excessively and needlessly stark.
  • You seem excessively comfortable with the idea that the choice is either being under the control of AI or not using IT at all. This is excessively and needlessly stark.

    No, I'm perfectly comfortable in saying that none of these systems pose the kind of threats imagined in this thread and that all of it is under human control.

    The chatbot salesmen are using a great deal of anthropomorphised language as terms of art but there is no need to follow them in this.
  • peasepease Tech Admin
    edited September 11
    It is the case that removing the power from a single computer, or from an entire server farm housing AI agents, is going to stay within the compass of human beings for a while yet.

    However, the primary threat of harm from AI isn't that it will wipe us *all* out, it's the enshittification of the services that *most* human beings depend on. Or, if you prefer, the lives of *most* of us slowly turning to shit.

    If the AI off-switches are in the hands of the few people who are being enriched and empowered by the provision of AI services, how do you think the decision about switching them off is going to go?

    Maybe your life currently benefits from AI - for example, from being able to have discussions with a chat bot. Questions to consider include: How much control do you have over this service? How much would be prepared to pay to carry on using it? How long would you carry on using it as the quality slowly degrades?
  • If the history of the Internet is a guide, we may wait for some years before Congress does anything meaningful. We need another Al Gore.
  • You seem excessively comfortable with the idea that the choice is either being under the control of AI or not using IT at all. This is excessively and needlessly stark.

    No, I'm perfectly comfortable in saying that none of these systems pose the kind of threats imagined in this thread and that all of it is under human control.

    The chatbot salesmen are using a great deal of anthropomorphised language as terms of art but there is no need to follow them in this.

    It occurs to me that these self-demonizing AI entrepreneurs might be kinda like the producers of old-school nudie flicks slapping the non-existent "XXX" rating onto their ads and posters, ie. sensationalist advertising masquerading as an urgent warning about the supposed evil of the product.
  • Anglican BratAnglican Brat Shipmate
    edited September 11
    I am worried about the potential impact of AI, what should the church's response be to the possible human toll and destruction caused by AI
  • I am worried about the potential impact of AI, what should the church's response be to the possible human toll and destruction caused by AI

    Pope Leo XIV just issued an encyclical on this very subject. I am slowly reading my way through it.

    I am bemused (but not surprised) that the Pope issuing an encyclical does not move politicians to look at the subject at all, but have a twenty-something programmer say "A.I. will kill us all in ten years!" and the politicians take it seriously.
  • Hedgehog wrote: »
    I am worried about the potential impact of AI, what should the church's response be to the possible human toll and destruction caused by AI
    Pope Leo XIV just issued an encyclical on this very subject. I am slowly reading my way through it.
    Also FWIW, the General Assembly of the Presbyterian Church (U.S.A.) approved a report— “The Algorithm and the Almighty: Navigating Artificial Intelligence through a Reformed Lens”—this past summer. And the third “AI and the Church Summit” has recently wrapped up in the US.

    I am bemused (but not surprised) that the Pope issuing an encyclical does not move politicians to look at the subject at all, but have a twenty-something programmer say "A.I. will kill us all in ten years!" and the politicians take it seriously.
    I know. I’m not surprised when politicians don’t pay any attention to what the PC(USA) or any other mainline Protestant have to say. But the pope?!


  • RuthRuth Shipmate
    You seem excessively comfortable with the idea that the choice is either being under the control of AI or not using IT at all. This is excessively and needlessly stark.

    No, I'm perfectly comfortable in saying that none of these systems pose the kind of threats imagined in this thread and that all of it is under human control.

    The chatbot salesmen are using a great deal of anthropomorphised language as terms of art but there is no need to follow them in this.

    I particularly love the "solution" to AI threats I read the other day (somewhere, sorry, of course, I can't remember): "Just buy our AI-powered defense!"

    I'm more concerned about things like our electrical grids and water systems being hacked, which doesn't require AI but could certainly be assisted by it.
  • Hedgehog wrote: »
    I am worried about the potential impact of AI, what should the church's response be to the possible human toll and destruction caused by AI

    Pope Leo XIV just issued an encyclical on this very subject. I am slowly reading my way through it.

    I am bemused (but not surprised) that the Pope issuing an encyclical does not move politicians to look at the subject at all, but have a twenty-something programmer say "A.I. will kill us all in ten years!" and the politicians take it seriously.

    Well, the comment was endorsed by some much older pundits.

    The chances were described as a '10% chance' and 'not unreasonable' by some hard-hitters.

    I have no idea whether it's a 1 in 10 chance or a 1 in 10,000 chance or 1 in 100,000 chance but the environmental effects and the way it could alter the way we 'think' (or don't think) is a cause of concern I reckon.
  • The "hard hitters"are all people either working for these companies, who will, regretfully, and despite their expressed concern continue to work so they can have a piece of the bag, or so called rationalists like Yudowsky who are basing their predictions on fantasy.

    Here's a piece by Timnit Gebru, an actual expert who was sacked by Google after writing a paper on some of the real risks presented by LLMs and AI driven systems (discrimination on a large scale, and disinformation):

    https://www.wired.com/story/one-of-ais-fiercest-critics-says-all-the-doom-talk-is-meant-to-distract-us/
    "In my book I write, “Can a bridge decide to collapse?” The whole “Oh no, our models went rogue” thing … When a bridge collapses, you don’t analyze whether the bridge was ethical or sentient or why it decided to collapse. You ask, who is the person who built this bridge to be so flimsy? There are permits you have to acquire. There are supposed to be tests. When you ask about the bridge collapsing itself, you’ve lost the thread."

    Quite. When someone leaves their dog unleashed, and the dog bites someone's child, we don't talk about "dog alignment". We talk about leash laws. These chatbots are by a wide margin less sapient than even very stupid dogs. "AI safety" doesn't mean entertaining fanciful science fiction scenarios. It means applying basic controls to systems.

    The day that these companies become liable for episodes where 'AI ran out of control' the attacks will go away.

    To loop it back to the OP, here's author Cory Doctorow explaining the original scenario described:

    https://pluralistic.net/2026/09/12/god-in-the-box/
  • Nick Tamen wrote: »
    Hedgehog wrote: »
    I am worried about the potential impact of AI, what should the church's response be to the possible human toll and destruction caused by AI
    Pope Leo XIV just issued an encyclical on this very subject. I am slowly reading my way through it.
    Also FWIW, the General Assembly of the Presbyterian Church (U.S.A.) approved a report— “The Algorithm and the Almighty: Navigating Artificial Intelligence through a Reformed Lens”—this past summer. And the third “AI and the Church Summit” has recently wrapped up in the US.

    I am bemused (but not surprised) that the Pope issuing an encyclical does not move politicians to look at the subject at all, but have a twenty-something programmer say "A.I. will kill us all in ten years!" and the politicians take it seriously.
    I know. I’m not surprised when politicians don’t pay any attention to what the PC(USA) or any other mainline Protestant have to say. But the pope?!


    Well, in fairness, the number of journalists and politicians who have anything other than ecstatic memories of their trips to the eternal city is almost certainly smaller than the total number who have ever vacationed in Louisville Kentucky.
  • stetson wrote: »
    Nick Tamen wrote: »
    Hedgehog wrote: »
    I am worried about the potential impact of AI, what should the church's response be to the possible human toll and destruction caused by AI
    Pope Leo XIV just issued an encyclical on this very subject. I am slowly reading my way through it.
    Also FWIW, the General Assembly of the Presbyterian Church (U.S.A.) approved a report— “The Algorithm and the Almighty: Navigating Artificial Intelligence through a Reformed Lens”—this past summer. And the third “AI and the Church Summit” has recently wrapped up in the US.

    I am bemused (but not surprised) that the Pope issuing an encyclical does not move politicians to look at the subject at all, but have a twenty-something programmer say "A.I. will kill us all in ten years!" and the politicians take it seriously.
    I know. I’m not surprised when politicians don’t pay any attention to what the PC(USA) or any other mainline Protestant have to say. But the pope?!


    Well, in fairness, the number of journalists and politicians who have anything other than ecstatic memories of their trips to the eternal city is almost certainly smaller than the total number who have ever vacationed in Louisville Kentucky.
    General Assembly was in Milwaukee, if that matters. :wink:


  • stetsonstetson Shipmate
    edited September 12
    General Assembly was in Milwaukee, if that matters. :wink:

    Ayyyyy!
  • stetson wrote: »
    General Assembly was in Milwaukee, if that matters. :wink:
    Ayyyyy!
    I had my picture taken with the life size statue of the Fonz. I knew Henry Winkler is short, but he is short!

    /tangent


  • Nick Tamen wrote: »
    stetson wrote: »
    General Assembly was in Milwaukee, if that matters. :wink:
    Ayyyyy!
    I had my picture taken with the life size statue of the Fonz. I knew Henry Winkler is short, but he is short!

    /tangent


    Happy Days in its heyday was definitely popular with my late-elementary aged crowd, but I wonder what sorta numbers it was drawing in from people actually old enough to be nostalgic for the late 1950s. My father seemed more aware and appreciative of that show and its characters, relative to other youth-oriented sitcoms. He was also, FWIW, a huge fan of Sha Na Na.
  • peasepease Tech Admin
    I am worried about the potential impact of AI, what should the church's response be to the possible human toll and destruction caused by AI
    Nick Tamen wrote: »
    Also FWIW, the General Assembly of the Presbyterian Church (U.S.A.) approved a report— “The Algorithm and the Almighty: Navigating Artificial Intelligence through a Reformed Lens”—this past summer.
    Thanks Nick Tamen, I thought that was fairly sane overview. Except for the irresponsibly uncritical take on risk:
    However, the risks are also substantial: widespread unemployment as AI takes over many jobs; the vast amounts of energy to support the computation required by AI; increased concentration of power and money in a few individuals, corporations or governments; and the use of AI to manipulate truth and distort reality. There are even existential threats, such as the possibility that ASI [Artificial Super Intelligence] will take complete control over life on Earth, relegating human beings to subservience.
    Conflating all these things, implicitly granting them a kind of plausible parity, is really unhelpful.

    I haven't got very far through the papal encyclical (Magnificent Humanity), but it was noticed by people from a wide cross-section of society. Focusing on the humanness of humans sounds like one of the Church's better ideas.
  • peasepease Tech Admin
    Also on risk, thanks chrisstiles for those links. Cory Doctorow describes just how irresponsibly mundane the story behind the OP is #. Just hackers being hackers, who someone keeps supplying with as much money as they can burn:
    Given that AI insiders have mostly cooked their brains in this fashion, it behooves us all to treat these people as unreliable narrators of their own products' capabilities. Remember: every time you repeat a story about how awfully, terribly dangerous their products are, you help them raise more investment capital, which is a key input for their business (hooking up statistical engines to money-furnaces)
    # Basically, it's hooking up an AI chat bot to a bash prompt.

    Humans aren't good at assessing many kinds of risk even when we're thinking straight. But to illustrate how the techbros think about it, a couple of days ago I watched Lior Susan, head of Eclipse (a venture capital firm that invests in "Physical AI", the integration of hardware and software to fix aging infrastructure and improve supply chain resilience) say the following about the current pushback on data centres:
    Data centers are good for our countries, but we need to do a much better job of sharing the profits and safety and energy with the communities. … I think regulation generally, it's a risk for the economy.
  • Barnabas62Barnabas62 Shipmate, Host Emeritus
    pease gets it right. And enshitification is a very good word.
Sign In or Register to comment.