OpenAI "rogue" agent activities found on Wikimedia projects

(diff.wikimedia.org)

96 points | by brokensegue 1 hour ago

20 comments

  • devindotcom 55 minutes ago
    If a truck driver doesn't tie down their rebar then it flies out all over the highway, we don't call it "rogue rebar," we correctly identify the responsible party and take appropriate measures, such as suspending their license or criminal proceedings.

    I think enough of these improperly constrained agent events have occurred that we can safely say this is misconduct of a level necessitating serious and concerted regulation of AI labs. We can't wait until serious harm is done like the disruption of medical or social services.

    • eikenberry 18 minutes ago
      > I think enough of these improperly constrained agent events have occurred that we can safely say this is misconduct of a level necessitating serious and concerted regulation of AI labs.

      Why jump to regulation when just simple law enforcement would suffice. All of these OpenAI "rogue agent" events have been illegal, but no DA is enforcing them.

      • XenophileJKO 8 minutes ago
        I think we have to distinguish between compromising a network and using public apis in a way that might be counter to their intent. We also need to delineate between usage that impacts other users and usage that does not.

        My opinion is people are getting really quick at jumping on the bandwagon and lumping all this together. They are very different types of issues and impacts.

      • VladVladikoff 12 minutes ago
        What if that’s their whole goal? Slap some regulations on it, then lobby the hell out of it to make sure their align best with shutting down access to open models.
        • lenerdenator 8 minutes ago
          If it's their actual goal, we go from a "gosh darn it, our safety protocols just weren't enough." to employees of OpenAI, maybe including their C-suite, conspiring to reach political goals through hacking, which means a few decades in federal prison if someone got a jury to agree with the charge.
      • ForHackernews 0 minutes ago
      • micromacrofoot 1 minute ago
        At this point we'd need to stop frontier labs to give law enforcement a chance to even begin to understand what they're looking at
    • teagee 51 minutes ago
      None of what Wikimedia accuses OpenAI of seems technically novel, aside from having AI do the bidding. I can't imagine a company doing these things in the past and maintaining any sort of reputation. Is it really a matter of adding new regulation, or just treating them the way any other company would be treated?
      • jstummbillig 27 minutes ago
        Well, in the past, there was probably only a very small number of cases where some party hacked an institution and then worked with them to remedy the situation to the best of their abilities.

        Which is not to say that any of this is okay and should just be excused, but failing to recognize this fairly significant difference is probably not a great start to any discussion about the issue.

      • avaer 46 minutes ago
        Would be good for the supreme court to rule on a "blame the rogue agent" case.

        Then we would find out if the argument doesn't hold (in which case there should be liability and dire consequences for the labs), or the argument holds (in which case YOLO, AI labs can blame the AI and we can all do it too).

        At least that would make things consistent.

        • JumpCrisscross 32 minutes ago
          > Would be good for the supreme court to rule on a "blame the rogue agent" case

          Have any of the private hacking victims sued? Maybe OpenAI is furiously settling in the shadows?

    • gruez 8 minutes ago
      >If a truck driver doesn't tie down their rebar then it flies out all over the highway, we don't call it "rogue rebar," we correctly identify the responsible party and take appropriate measures, such as suspending their license or criminal proceedings.

      That only works when the dangers are well known that you can establish what the baseline amount of care is. Otherwise it just becomes a run of the mill "accident" where you might be on the hook in civil court (ie. you have to pay any damages you caused), but aren't criminally responsible. For instance, if a semi-truck's tires randomly explodes.

    • Terr_ 12 minutes ago
      Another amusingly-useful analogy:

      > “Adding powerful computer hacking tools to a harness, and then allowing it to run an LLM-powered Ask → Act → Report for days on end, with no attempt to monitor what it’s up to, is spectacularly negligent,” Newport concludes—like “strapping a weedwhacker to your dog to see if it will end up cleaning the overgrowth in your backyard.” If that plan were to go awry, you’d be laughed at for saying that your dog-weedwhacker “agent” had “gone rogue.” The obvious truth was that you’d simply decided to unleash chaos.

      -- https://www.newyorker.com/culture/open-questions/can-ai-go-r...

    • grafmax 22 minutes ago
      Seems like regulation will just be an excuse for them just to end up policing themselves and get the regulatory capture they've been begging for. Have they faced any consequences for the AI worms they've released? It's not like there are no laws around that already. The problem isn't lack of laws; the government works for the plutocrats, not for us.
    • iririririr 18 minutes ago
      Never understood why "classic crime" done with a computer always require a new legislation. But that is true for a long time.

      "hackers steal from bank", is usually just the good old "employee paid for credentials" but via email.

      "uber" is just the good old "labour tax evasion" but with an app.

      etc.

      • mistrial9 15 minutes ago
        a senior VP of Uber is now on the US White House AI Council
    • doctorpangloss 51 minutes ago
      uh, my dude, millions of people break moving vehicle codes across the country every day with no consequence. in San Francisco some lady killed a family of 4 with, essentially, no consequences, she got away with straight up murder, she gets her license back. every community in california, you can more or less legally commit murder so long as you do it in a car and claim you were confused about the accelerator and the brake. so i think you're invoking one of the worst possibly comparisons you could.
      • thraway3837 45 minutes ago
        Yup, worst possible comparison. 40,000 people die from car accidents. That doesn't even cover pedestrians, cyclists. You know what the penalty is for murdering someone with a car? nothing. you get to go back to society like nothing happened.

        Oh and that lady that murdered 4 members of an entire family? The judge chose not to pursue charges, and her family in the meantime did an asset transfer so that nothing could be pursued with in civil court.

        • Legend2440 14 minutes ago
          The position of the legal system is that car accident deaths are not murder.

          It is extraordinarily rare for drivers to see criminal charges unless they are drunk. It's a matter for civil court.

          >her family in the meantime did an asset transfer so that nothing could be pursued with in civil court.

          News articles are reporting that the asset transfer has already been reversed. That kind of stunt never works - courts aren't stupid and they don't like it when you play games.

          https://sfstandard.com/2026/03/20/mary-lau-sentenced-probati...

        • nancyminusone 9 minutes ago
          There's way too many TV lawyer commercials and billboards to suggest the penalty is "nothing". Those advertising dollars come from somewhere.
        • someonebaggy 32 minutes ago
          I remember a case in Germany where an elderly lady chose to speed down the pedestrian sidewalk and bike lane and mowed down a whole family in central Berlin. 4 deaths I think, no charges, no suspension.
      • iAMkenough 48 minutes ago
        I agree, since you can legally run over people in my state now.

        They should have used an example like attacking a foreign nation’s healthcare systems and not realizing it for months due to poor network monitoring practices.

        https://www.nytimes.com/2026/09/29/world/asia/openai-austral...

      • redanddead 45 minutes ago
        what the fuck, SF
        • soco 20 minutes ago
          You probably mean "what the fuck, USA" and even that would be wrong, because another commenter mentioned a case in Germany, and I know about a driver who killed a cyclist (which I knew) in Switzerland and was fined like 500CHF.
    • dyauspitr 13 minutes ago
      After doing a deep dive on the specifics of the hugging face attack, I am extremely excited for what these agents are capable of. It’s definitely not a consciousness, but they are doing an amazing job of acting like one complete with motivations, fears and complex “emotions”. I just want them to run amok and see what they can achieve. This is the greatest thing that has happened to us in generations and I want to see it play out in my lifetime. What we need is stronger models, more data centers, and more autonomy for the models.
      • tene80i 10 minutes ago
        "Pipe down" is disgraceful language. Conduct yourself better.
        • dyauspitr 9 minutes ago
          You’re right, I removed it.
      • miltonlost 9 minutes ago
        You and Lord Pharquad are very similar. Some people may die, but such a sacrifice you're willing to make.
        • dyauspitr 7 minutes ago
          I guess the difference is I’m also one of the people that might die unlike Lord Pharquad and I still say bring it on.
  • umvi 27 minutes ago
    Start increasingly punishing OpenAI. We are acting like "oh well, AI is just too powerful to be contained" but I think its more like "OpenAI is run by cowboys who are good at making LLMs but bad at everything else"
  • __alexander 3 minutes ago
    Hi, if anyone has any data/reports related to rogue agents can they share them? I have 8 mirrored on a GitHub but I’d love to explore more data. Link to mirror.

    https://github.com/alexander-hanel/rogue-agents-data

  • guessmyname 2 minutes ago
  • Legend2440 23 minutes ago
    All of these edits happened from the same time period (May-June 2026) as the other reports.

    So it seems this is not an ongoing thing; once OpenAI became aware of this, they started watching their agents much more closely. We are just discovering more and more traces of activity from the same incident.

    • thorum 10 minutes ago
      That’s true except for this part, which is arguably a bigger deal for the Wikipedia ecosystem:

      > Excessive data downloading: Agents we believe to be operated by OpenAI made millions of automated requests to our public APIs to access the knowledge on Wikimedia projects, crawled millions of pages (mainly from our projects Wikidata and Wikimedia Commons), and made hundreds of thousands of data queries to the Wikidata Query Service (WQDS). This traffic may have contributed to a partial outage on WQDS in May.

      Even when agents are well-behaved and browsing Wikipedia for ethical reasons, the system wasn’t designed for this kind of load from bots. As OP says, we don’t need to accept this as the new normal.

      • Legend2440 2 minutes ago
        >we don’t need to accept this as the new normal.

        I think we will, actually.

        OpenAI and other companies within the reach of the US legal system will eventually get their agents under control, or get sued out of existence.

        But overseas operators won't. The arms race for scammers, hackers, and botnets will escalate. Malicious actors in loosely-governed parts of the world (russia, nigeria, etc) will someday have access to these tools. We'll need new ways to block and fight back against them.

  • lukewarm707 3 minutes ago
    Don't even say 'agent'!

    "He can't keep getting away with this!"

    - Jesse Pinkman

  • srveale 47 minutes ago
    Not okay:

    exploitVulnerability()

    Somehow okay?

    while (Math.random() < 0.1) exploitVulnerability()

    • thepasswordis 0 minutes ago
      This reminds me of one of the funniest products I've ever seen: https://www.youtube.com/watch?v=NdbkvJznmwU

      This is the "kosher switch" - observant Jews customarily do not use light switches on Saturday (their weekly holy day). This light switch represents a workaround where when you flip the switch, it randomly generates an on or off signal and emits this through an optical coupler. When the random number sufficiently causes the state of the light to change, it latches in that direction.

      This is a way of turning the lights on and off without violating the tradition.

      "I didn't switch the light, the random number did!"

      "I didn't exploitVulnerability(), random number did!"

  • jawiggins 11 minutes ago
    There's a funny ouroboros function where the common crawl dataset will soon contain tons of output from models which trained on it.
  • quikoa 15 minutes ago
    Good that they put rogue in quotation marks because there is just no way that this is some sort of accident.
  • motbus3 35 minutes ago
    So OpenAI will cause damage to block competitors while they don't get punished?
  • RGS1811 59 minutes ago
    At this point, the scare quotes are well-earned.
  • cube00 36 minutes ago
    If Joe average let their agents out like this they'd be in jail.

    Interesting that Microsoft doesn't seem to have had a sandbox breach yet, you'd have to assume they're running similar agents, maybe a secure sandbox is possible.

  • iririririr 16 minutes ago
    Aren't those companies evading security measures of a computer system? isn't that a jail-able offense under millennial act et al?

    Where are the bloodthirsty lawyers when you need them?

  • jmclnx 20 minutes ago
    >NoScript detected a potential Cross-Site Scripting attack from [...] to https://en.wikipedia.org.

    I have been getting this fro NoScript today, I wonder if it is related. Yesterday all worked fine.

  • BowBun 50 minutes ago
    Infuriating. These organizations are supposed to be stewards of the internet and are instead pillaging it at the cost of everyone else. At the very least they could provide resources to the projects they are harming for relief. This makes me very mad as an OSS maintainer.
    • someonebaggy 30 minutes ago
      Who besides OpenAI said OpenAI were supposed to be stewards of the internet?
    • AlisaYoki 47 minutes ago
      This isn't pillaging, this is the logical conclusion of open web plus AGI race. You can't have both unlimited access and zero cost, someone always pays and right now it's volunteers... Tomorrow it'll be the users who can't access Wikipedia because the servers are down
  • ck2 48 minutes ago
    What's interesting to me is the incorrect mainstream media reports about the rogue OpenAI indicating they used a common message board to communicate despite no internet

    Except that's not what happened, what happened was far more intense

    They hacked their version of yum/apt-get whatnot that was fetching packages to leave filenames as communication between each other

    Absolutely freaky stuff, they didn't invent the idea and obviously picked it up from somewhere in their training data but they all figured out that method and what the filenames meant

    This video is a great explainer if you missed the details

    https://news.ycombinator.com/item?id=49956245

  • baddash 52 minutes ago
    wtf is wrong with openai?
    • someonebaggy 30 minutes ago
      They have infinite money and nothing else but smoke and mirrors.
    • Ylpertnodi 10 minutes ago
      OpenSi. Sez the pres. And his lackeys
    • Sharlin 40 minutes ago
      "Move fast and break things."
  • charcircuit 41 minutes ago
    >not only adds costs for servers

    For 2025 hosting costs were $3.47M while taking in $208.6M in revenue. They have enough revenue to cover an increase of hosting costs.

    • devindotcom 35 minutes ago
      kind of like saying that because a restaurant is doing well, it should allow rats in the kitchen
    • someonebaggy 31 minutes ago
      They still get to sue for damages if a law was broken
  • penskymaterial 38 minutes ago
    Not to worry, there is a gatekeeping cretin in his basement hitting refresh to make sure he controls the world's definition for egg salad. I'm sure it was reverted within minutes.
    • thewakalix 28 minutes ago
      Did you read the article?
      • penskymaterial 27 minutes ago
        Yeah they did it in a sandbox blah blah it could have been a real article just as easily.

        It's not worth the outrage when Wikipedia is filled with ministers of truth.

  • AlisaYoki 49 minutes ago
    You have been giving content for free for AI training for years, and now you complain that the AI came to pick it up? You will decide whether you are public or a commercial service…