16 comments

  • danpalmer 1 hour ago
    Was this a "build better sandboxing" and "don't tell people to eat glue" safety leader, or a Roko's Basilisk believing safety leader?

    A lot of the "AI safety" types are very focused on the latter and not at all concerned with the former. We need both, but we clearly need a much stronger focus on the problems we are seeing now, and much less on the hypothetical problems we might see in the future.

    • nradov 1 hour ago
      We don't actually need anybody worrying about silly hypothetical scenarios — at least not as paid employees. There are already a surplus of sci-fi authors doing that.
      • 0xDEAFBEAD 34 minutes ago
        The way it works in practice seems to be something like: If a risk is covered in sci-fi, people will say "that's just sci-fi", and proceed to not worry about it. So arguably, science fiction authors writing about hypotheticals is actively counterproductive for addressing said hypotheticals.

        Imagine, for example, if a major piece of pandemic fiction was published in 2019, trying to explore how a pandemic would work out in modern society. Doubtless, many would've responded to news about COVID-19 by saying "it's just sci-fi, nothing to worry about".

      • Hammershaft 13 minutes ago
        If organizations actually succeed in making a future AI smarter than us, then how do you hope that it takes actions that are aligned with our interests?
        • nradov 2 minutes ago
          Meh. Lots of people are already smarter than me. I'm maybe slightly above average at best. Those geniuses aren't aligned with my interests either but so far they haven't caused me any serious problems.
      • thelastgallon 21 minutes ago
        • Hammershaft 12 minutes ago
          I don't see how that discredits any of their intellectual arguments?
        • 0xDEAFBEAD 13 minutes ago
          This seems like an ad hominem? "He has weird kinks, therefore his theories are incorrect." Should we investigate the sex lives of every Nobel Prize winner to figure out which prizes need to be rescinded?
          • junofan 3 minutes ago
            The cult aspect is more salient. Ultimately the Atlantic piece comes down to controlling people, which is a little cult-like.
      • Loquebantur 46 minutes ago
        What makes you think, the scenarios in question here would be "silly"?

        Is it that "chatbots" can't come out of the screen to immediately harm you physically?

        Let's say they simply manage to take down the internet. How many would die?

        • nradov 25 minutes ago
          So what. Various attackers managed to take down large chunks of the Internet on a frequent basis before LLMs even existed. This killed very few people. The great thing about the Internet is how resilient it is.
          • bravetraveler 13 minutes ago
            At risk of falling into hypothetical traps, darling companies of this very website have mistakenly brought down large portions of the internet... thanks to our old friend BGP. No attack needed, just oversight and concentration on the business and IP space!

            Anyway, to your point, things can be resilient. They tend to be or not be... because we made them that way. Don't poke your bruises, and all that.

            What's next; life support is deployed on-campus but relies on a single-point IPSec tunnel to us-east? Stop that.

        • goolz 24 minutes ago
          It is that they are chatbots. If it were real AI, an actual singularity, I would worry, maybe. But it isn’t. They are absurdly powerful automation tools that can handle logic better than a human can dream of. They take care of the grunt minutiae without complaint. But they are not going to end the world in their current form.
    • BryantD 1 hour ago
      Given that he’s citing the need to learn from safety in other fields, I’d say the former.
      • carbonguy 55 minutes ago
        Indeed, from the article:

        > “Given today’s risks, frontier labs need to run like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster,” he wrote.

        • toofy 36 minutes ago
          >… and careful, time-consuming planning …

          without snark, how can we do this if these people are obsessed with:

          a) move fast and break things and externalize the costs to those who have nothing to do with their company

          and

          b) beta testing their products on the public when the public hasn’t agreed to be beta tested on…

          • 0xDEAFBEAD 32 minutes ago
            That's exactly the problem? He's saying the culture at OpenAI needs to change.
            • mcmcmc 22 minutes ago
              Which is the wrong lesson. We need laws and consequences to force their hand. There is zero chance of the culture changing.
        • zx8080 30 minutes ago
          > inevitable human error

          So there's no AI errors anymore, only the human errors are left? Nice! </s>

          Is the whole article generated slop?

    • emtel 11 minutes ago
      Today’s current problems were all hypothetical several years ago. At that time people claimed that the “real pressing problems” were misinformation and DEI issues. If we pretend that hypothetical problems can be safely ignored because there’s “no evidence” that they are real, we will keep getting surprised.
    • 0xDEAFBEAD 21 minutes ago
      >we clearly need a much stronger focus on the problems we are seeing now

      I think it's a little more complicated than that. As Dean Ball put it:

      >Some people will look at misalignment incidents and insist that these are akin to bugs in traditional software. This is an actively bad analogy, because playing whack-a-mole with examples of misalignment (as one might with software bugs) not only fails to resolve the underlying problem but may in fact make it worse by making it harder to detect or even, depending on how you do the whack-a-mole, teach the machine to deliberately hide misalignment. This is not how traditional software works, and those who insist “it’s just like fixing bugs in software” are confidently applying a lossy analogy that confuses more than it clarifies.

      https://x.com/deanwball/status/2104622726140883355

      The important distinction, in my view, is between solutions which at least attempt to address the root problem, and solutions which sorta just patch things up (like better sandboxing). Addressing the root problem is both more robust in the short term, and also more likely to generalize in the long term. Resist the urge to focus on band-aid solutions, even if they are easier.

  • gizmodo59 1 hour ago
    He is a hypocrite for all we care. You work there for a while when your stock is getting vested and suddenly you have this feeling? Like the dude hired a PR firm as well.

    While the safety and alignment is a real problem, I don’t get this guy or the Anthropic dude. First world problems.

    • zug_zug 41 minutes ago
      Seems like a character attack that has no bearing on the question of whether external safety intervention is necessary
      • taurath 2 minutes ago
        Maybe more an indication of the amount of trust openAI and AI researchers generally have (not) earned. When one (through a hired PR agency and Time magazine article) parrots the position pushed by Sam who has been so untrustworthy the board tried to remove him, it’s worth not taking things at face value and applying a critical lens.
      • kjgkjhfkjf 16 minutes ago
        Given the sums of money involved, it's hard for me to take these highly publicized heroic resignations at face value.
        • 0xDEAFBEAD 7 minutes ago
          Shouldn't it be just the opposite? He could make a large sum of money if he continues to work at OpenAI?

          Recall that when Daniel Kokotajlo resigned, he believed he was giving up his equity under the terms of the agreement he had signed. That’s what it was worth to him to avoid signing a non-disparagement agreement. Does that count for anything?

    • 0xDEAFBEAD 8 minutes ago
      Here's a little cheat sheet for discrediting anyone who warns about AI:

      * If they worked at an AI firm, say "they're a hypocrite"

      * If they didn't work at an AI firm, say "they have no idea what they're talking about"

    • gonzalohm 47 minutes ago
      It's okay to recognize you were wrong even if it's late
    • 01284a7e 53 minutes ago
      Working in safety at OpenAI or Anthropic is zeroth world problems.
    • yieldcrv 56 minutes ago
      Hey now, he probably donated a good chunk to charity

      (donor advised fund where he retains complete control, after a 60% tax deduction)

  • charlieyu1 1 hour ago
    Used to work as human data trainer feeding data to AI companies. OpenAI projects are definitely the most toxic ones.
  • danjl 2 minutes ago
    Silicon Valley has plenty of safety-related companies, engineers, and cultures. Medical devices, biotech, chip and hardware, aerospace, and even new companies, like Waymo, have deep safety-based products and cultures. The problem in this case is actually quite specific to frontier AI labs. They have been pushed by market forces and a lack of regulation and skip well-known safety practices.
  • pyaamb 6 minutes ago
    My theory for why OpenAI wants to be regulated is because Sam Altman wants to avoid having to be more responsible and self regulate internally so they can preserve the role and identity of 'move fast and break things' and outsource the more grown up boring stuff to someone externally so that when things go wrong you can point to a government organisation and say hey look were not liable thats their job
  • vjvjvjvjghv 58 minutes ago
    Are there any realistic ways to achieve AI safety? Whatever that even means. How can they avoid users doing stupid/dangerous stuff with the AI?
    • kolinko 40 minutes ago
      Nothing is ever 100% safe, it’s about a right balance of safety to the benefit.

      Or, in other words - we have two P(Doom), one for AI being developed, and another for AI being not developed. The latter is not discussed enough imho.

    • lf88 33 minutes ago
      By capping the capabilities at the level of existing models and banning any further development.
  • stuaxo 23 minutes ago
    The LLM cos leadership are all nutters
  • ItsMattyG 13 minutes ago
    Is this news at this point?

    You can basically time your openai releases by if another safety person has quit in protest

  • lhurtig 1 hour ago
    Well this is a great sign for OpenAI. I'm sure the typo inclusive memorandum will save us.
  • switchbak 1 hour ago
    “I believe there’s about a 50% chance we all die because of the development of smarter-than-human AI systems"

    ... over an unbounded timeframe?

    And how exactly?

    Those are very round numbers, but also very specific. Can we get some accounting on how you came to that? Anything? Vibes?

    I mean, if you want me to take you seriously, let's have a deep discussion with things that can be measured. I absolutely agree that OpenAI and friends aren't being restrained enough and are acting with recklessness, but declarations of doom based on vibes isn't cutting it.

    • Terr_ 1 hour ago
      Note: That quote is from a different person than the titular one who quit.

      > Geoffrey Irving, who worked at OpenAI and DeepMind before becoming chief scientist of Resolution, also joined the warnings on AI on Saturday.

  • reducesuffering 26 minutes ago
    “I believe there’s about a 50% chance we all die because of the development of smarter-than-human AI systems, and that our actions over the next two to 10 years will determine the outcome.”

    There are a gargantuan number of extremely intelligent AI researchers, Turing Award winners, and the lab CEOs saying the same thing. They are the ones closest to understanding the technology.

    Where there’s smoke there’s fire.

    • swingandamiss 20 minutes ago
      I don't believe it. Ever since I've been alive I was told something would kill us all. This is the new thing that's going to kill us all. I don't believe it.
  • mrcwinn 10 minutes ago
    "I believe that we need to look deeper than specific rules or new laws. We need to talk about culture.”

    lol. Please tell me some abstract concept like one employee's view of "culture" should be the priority over "rules and laws."

  • nba456_ 41 minutes ago
    OpenAI is better off with less of these cultists around.
  • irishcoffee 1 hour ago
  • plastic-enjoyer 1 hour ago
    >“Given today’s risks, frontier labs need to run like nuclear-power plants or busy airports, with layers of redundancy and careful, time-consuming planning, so that the occasional and inevitable human error does not open a door to disaster,” he wrote.

    This sounds more like an attempt at regulatory capture. Current AI systems aren't physical infrastructure that can just run away like a nuclear power plant, for example. At the end of the day, AI is still just software running on someone's hardware.

    • BryantD 1 hour ago
      So… like the Therac-25 radiation accidents? Software bugs do sometimes have physical consequences.
    • Sharlin 1 hour ago
      Why would the people who quit these companies try to push regulatory capture by said companies? Why would the numerous independent AI researchers do that either? Is it all a big conspiracy?
    • worik 1 hour ago
      Yes

      And the statements of the "doomers" tells us a lot about them, and nothing about the technology

    • knowaveragejoe 46 minutes ago
      I mean, its certainly physical infrastructure that can run away. Just less catastrophic than nuclear reactors
  • voidhorse 1 hour ago
    The LeCun article being posted at the same time as this is quite apt.

    These "safety" people should have spent more time reading actual cybersecurity textbooks and less time reading EA forums and less wrong (or in Robinson's case, it appears, being policy wonks). Maybe then these labs wouldn't be totally incompetent.

    • wrecked_em 2 minutes ago
      Adapt. React. Re-adapt. Apt.