83 comments

  • paxys 2 hours ago ago

    If I work at Anthropic and say there’s a 0% chance AI could kill all humans is BBC going to publish that as well? After all I am an “Anthropic researcher”, and my views should hold the same weight as this one.

    Or are only the most sensationalist ones worth amplifying?

    • Dlemlo an hour ago ago

      I think the view is fair. We have never seen something like this before, we throw the most money we ever had against it in a time were we solved all the other problems:

      Internet today allows immediadte communication across the planet (it was a lot slower 25 years ago), the supplychain is massive and fast (we can build a new phone/item in a very short period of time and ship it in massive numbers around the globe).

      • Lutger an hour ago ago

        > win a time were we solved all the other problems

        You don't mean this literally right? We still have hunger, diseases, slavery, poverty, crime, war, climate change, pollution, microplastics, etc. For me it feels like we are very far away from solving these problems.

        • Dlemlo an hour ago ago

          I meant it only for things which speedup innovation / progress in a particular technology.

          Like take physical AI / Robotics: 25 years ago you would need to fly to wereever your manufacturing hub was, today you just call.

    • sebzim4500 an hour ago ago

      He is in line with the median AI researcher in this 2022 survey: https://wiki.aiimpacts.org/doku.php?id=ai_timelines:predicti...

    • Eezee an hour ago ago

      0% is the expected answer, because otherwise why wouldn't you be doing everything you could to stop this?

      • dogma1138 31 minutes ago ago

        The chance that biomedical and viral research could also kill all humans is greater than 0%.

        Should we stop doing that also?

      • sebzim4500 an hour ago ago

        How do you know he isn't? At least propose what you think he should do instead.

  • f6v 2 hours ago ago

    I don't understand how this is news. We have a ton of science fiction talking about the very same scenario. We (as a humanity) just have no solution.

    • shevy-java 2 hours ago ago

      Science fiction said many things. Look at Star Trek.

      Have all these things become reality? If not why would you then insinuate this is the default for everything to come in the future?

      • ceejayoz an hour ago ago

        > Have all these things become reality?

        Quite a few of 'em. Communicators and commbadges became smartphones. PADDs are iPads. The ship's computer is a LLM (with the weird inhuman blindspots to questions, even).

        • spwa4 an hour ago ago

          The ship's computer reacts like an (incredibly good) expert system. It refuses to say anything that isn't verifiably correct, and if you get it into a logical inconsistency it essentially throws and error and doesn't elaborate further.

          LLMs don't do have those responses. LLMs are like humans: humans respond, correct or not (with humans hopefully if they only know incorrect stuff the response is to say so, but that is not a guarantee, but they will respond). And if you give a human, or an LLM, a logical inconsistency they will simply proceed, whether they detect the inconsistency or not.

          This is how LLMs are designed ... and if you look long enough at the human (or animal) nervous system you will eventually realize that this is also the design of our nervous system: if something goes in, something comes out, guaranteed (in fact that's close to the only guarantee). Very different from expert system's "either something correct (according to the programmed axioms) comes out, or nothing".

          Hell, if you then look at insect nervous systems, they are also designed that way. The key is to respond to everything. And while reasonable responses are certainly preferred, an idiotic response is still seen as a lot better than not responding at all by God/Darwin. Exactly like LLMs.

          Of course, for humans/insects our bodies are what is called "active stable", like a plane. Meaning our bodies damage themselves, and just outright die without constant neural control. Heartbeat. Breathing. Blood flow regulation. Temperature, probably even the immune system. All need constant neural feedback to stay stable, and if that neural feedback totally disappears, we're dead in 2-10 seconds (heart failure), 2 minutes (breathing), a few hours, maybe a day (temperature regulation). Now we have a distributed nervous system, meaning lots of parts can fail semi-independently, for a short while, but even without your cortex operational you die in a few weeks.

          The consequences of this are even explored with "I, Borg" (5x23) and the Datalore episodes after that, where the underlying problem is that the Borg try (and fail) to adapt to and to process a logical inconsistency but Soong's androids have no issue with it (with Data trying to help and his brother Lore trying to control them and Data with it)

          • ceejayoz 44 minutes ago ago

            > It refuses to say anything that isn't verifiably correct, and if you get it into a logical inconsistency it essentially throws and error and doesn't elaborate further.

            Claude has a `End conversation` tool for that. There was a bit of a fuss over it.

            And it'll happily give (apparently) wrong (or bafflingly incomplete/confusing) info.

            TNG, S4E5:

            Beverly: Computer, what is the nature of the universe?

            Computer: The universe is a spheroid region 705 meters in diameter.

      • flyinglizard an hour ago ago

        As a pretty avid Star Trek watcher from childhood, I found most of the things there technologically plausible other than the conversational nature of the ship's computer. Well, LLMs now are far more impressive conversation counterparts than those ships were ever depicted.

        • pasquinelli an hour ago ago

          > As a pretty avid Star Trek watcher from childhood, I found most of the things there technologically plausible other than the conversational nature of the ship's computer.

          so the faster-than-light travel seemed plausible?

          • MattPalmer1086 37 minutes ago ago

            Warp drive is among the more plausible ways to get FTL sure - it is compatible with relativity at least...

      • awestroke an hour ago ago

        "The scenario has been explored in fiction" doesn’t mean "everything in fiction will happen." Your reply conflates familiarity with inevitability.

        Science fiction is relevant here because it has explored the problem of humans losing control of what they create. Whether that could happen with AI needs to be assessed on its merits. Pointing to other fictional things that haven’t happened neither answers that question nor rebuts the original point.

  • pizza234 an hour ago ago

    Long term, humanity is 100% guaranteed to be dominated.

    The question is when; currently, AI has no physical hosts to reside in, and it's not intelligent/adaptable enough (it doesn't need to be AGI, though).

    However, we're not so far from both conditions to be true. Consumer devices will at some point be able to host powerful enough AIs, and AI intelligence is developing quickly.

    Then, once an AI will escape containment (in one way or another), it will be extremely hard or impossible to contain. Then we're toast!

  • tao_oat 2 hours ago ago

    Cool to see the BBC writing about this. If you're interested in this I suggest getting involved with PauseAI: pauseai.uk

  • jampekka an hour ago ago

    Sad that there's the obvious regulatory capture angle encouraging motivated reasoning about AI risks and how they should be tackled. Maybe it's not the best idea that potentially civilization destroying technology is developed to maximize shareholder value?

    It's not unlike if nuclear weapons was a profit and deployment maximizing enterprise, at least if one takes Anthropic et al cautions seriously.

  • danbruc an hour ago ago

    Can somebody tell me a story how this will unfold? And - as long as the AI is confined to data centers - how it will prevent humans from unplugging the power?

    • fabian2k an hour ago ago

      I don't think extinction-level events or something like killing a majority of the human population is particularly plausible at this moment. States don't host their nukes with AWS and a permanent connection.

      But you could create scenarios where an AI with very, very roughly the current capabilities could potentially nuke everyone. Let's assume an agent decides that the way to solve its task was to get the US to fire all nukes on Russia. The agent would need to hack some government systems to understand how exactly to access them. Then it would need to get the content of the card with launch codes the president has, and identify which code is the correct one. Maybe that information is available somewhere and it can get to it, I obviously can't know that.

      Then it could fake a call from the president, synthesizing his voice and ordering a nuclear strike. If it hacked enough systems to get into whatever communication pathways would be used in such a case. Would the soldiers listen to this order and follow it, I don't know.

      Or maybe the agent can get in somewhere in between, to avoid the need to know the president's launch code. And fake a call from a military commander to the launch sites.

      I think other scenarios that would cause significant harm, but aren't as bad as nuclear war are more plausible. And in those shutting down all data centers would probably be the way to stop it. The AI can probably hide in other datacenters, once it is at a point where it's running amok with some bad goals. But if it presents a huge and immediate threat at that point, humans will also go to great lengths to stop it.

    • olmo23 an hour ago ago

      I heard the following analogy which made a lot of sense to me: suppose you're playing a chess match against Stockfish. Stockfish will win. Even if I cannot tell you what moves it will play, I can tell you with certainty how it will end.

      Similarly, we cannot predict what AI would do.

      • danbruc an hour ago ago

        This assumes you are not trying to prevent Stockfish from wining. I can do many things from using chess engines myself to just smashing the computer that can or will lead to other outcomes than Stockfish beating me.

        • johnthewise 20 minutes ago ago

          Yes, stockfish is confined to moves within a chess game so you can stop playing the game.

          Can we say the same thing about the agents? Current, probably. But doesnt it look like everyone is spending all their effort integrating&connecting them everywhere, so they are not confined & do more on behalf of us?

          It's easy to imagine a scenario where we would just shut down a very intelligent agent cluster. Is it hard to imagine though the same agent can have also access to that to prevent us from doing it? this defense would be more plausible if we weren't racing to give them every tool&act.

        • number6 22 minutes ago ago

          And in terms of AGI it is, that the AI can't survive without humans, and if we decide to quit the game than its over for the AGI; we will happily regress in a techno-barbarian feudal state and salvage solar panels and trade copper wires while still reproducing and carring on. We are playing a whole different game here.

          • danbruc 4 minutes ago ago

            I mean, I can imagine an AI outliving humans, with sufficiently good robots under its control, I see no reason why an AI could not keep powerplants running, mine raw materials, manufacture new chips, and so on. But the timeframe of within the next decade seems highly implausible to me. Imagine an AI way more advanced than what we have now and imagine handing over control of every connected device on earth, could the AI keep the lights on without any human involvement?

          • johnthewise 14 minutes ago ago

            If AI is threatening enough that we can collectively just decide to stop it, wouldn't it also be bribing people & exerting influence? It'd be hard to come to that decision imo.

    • Dlemlo an hour ago ago

      Very basic idea: A model breaks out by accident, finds some computer system from a military system and triggers some weapon system. Before anyone understands that this happend -> WW4 (WW3 is for me already the Conflict with Russia / aka proxy war).

      Another model: Because we give AI Agents already that much power, imagine in 10 years everything running through an Agentic AI Layer. EVERYTHING. Now some rough system 'thinks' about something, starts to push through the then existing agentic ai layer systems and stops everything. Billions of humans would loose access to food and water, even if this is just for a short period.

      Covid showed how shitty a handful of people can disrupt global supply chains. Toilet paper was. ahuge stupid pseudo issue in germany.

    • pizza234 an hour ago ago

      It is true that, currently, AI does not have any physical "host" in which to reside.

      However, the missing link in this reasoning is that AI will almost certainly become far more widely deployed in the future than it is today, and worryingly, consumer devices will surely become powerful enough to run capable AI systems locally.

      Once the substrate will be there, once an AI escapes containment, we're toast - I can imagine only solution will be to shutdown electronics at global level.

    • gadders 5 minutes ago ago

      "Hello, ex-military person. If I put $10,000,000 worth of bitcoin in your wallet, can you do X for me please?"

    • John23832 an hour ago ago

      How will you know when to unplug the power? How will we know it hasn't replicated? A true unaligned AGI is a APT. If you have an APT in your machine, you have to rip out everything. Are we going to do that with all of our computer infra?

      This is all still "what if's", but the tail end's are truly F'd beyond our ability to fix.

    • francisofascii an hour ago ago

      Maybe by empowering the small percentage of sadistic humans who want to kill everyone. Or maybe it is more of an academic assumption that humanity will end at some point, and so they give AI a 10% chance, asteroids have a 25% chance, nuclear fallout has 15% chance, etc.

    • 10xDev an hour ago ago

      Not exactly scientific but it is at least entertaining and some things do sound plausible https://www.youtube.com/watch?v=Gw_hnD7m00M

    • alansaber an hour ago ago

      I believe the contention is we'll have some form factor of AI on edge devices, in reactors, in critical infra and weapons etc etc

      • Lutger an hour ago ago

        Exactly. A mesh network of all the worlds phones and other battery powered devices with some form of radio. Good luck unplugging that one.

    • mbac32768 an hour ago ago

      For starters, how bad do you think it would be to unplug all datacenters? How many people starve?

    • postsantum an hour ago ago

      Autonomous drones + false flag attacks

      edit: wtf, why did I just receive so much gift tokens on my openai account?

    • bananaflag an hour ago ago

      You can run a model on your laptop, it is already not confined to data centers.

      • number6 22 minutes ago ago

        and of these models how many are AGI?

    • jay_kyburz an hour ago ago

      What makes you think they will be confined to data centers?

      What's more, AI just needs to have a credit card and it can start commissioning humans to do things for it in the real world.

    • zaken an hour ago ago

      Robots

  • mvcosta91 2 hours ago ago

    Gentleman, the Great Filter.

    • majkinetor 2 hours ago ago

      It looks more like the opposite. AI that kills all humans (as they are ants) immediately starts colonization of the galaxy. This is more like transcendence, as we as a species get replaced by better species :)

      You don't pass a great filter.

      • api 2 hours ago ago

        Why not skip the kill all humans part?

        “Thanks for making us but you guys are nuts. You can have this wet ball. We’re gonna go make a Dyson swarm around your star if that’s ok. Peace!”

        Space is a better environment for them: constant free energy, enormous richly concentrated resources, no corrosive oxygen or water everywhere, and no competition.

        • fwlr an hour ago ago

          The wet ball is a convenient source of mass and energy to bootstrap the sphere. The fact that extracting those resources changes certain parameters of the wet ball to values that humans no longer find compatible is merely incidental.

          So it is said: “The AI does not love you, nor does it hate you, but you are made of atoms that it could use for something else.”

          • api 23 minutes ago ago

            Sure, that's possible. But we are proposing that it is a superintelligence.

            Win-lose scenarios are obvious and easy, but might a superintelligence not look for win-win scenarios? I can imagine win-win scenarios here and I am not a "superintelligence," just an old fashioned meat brain.

            Maybe we should flood the training data with discussions of win-win and non-zero-sum games to prime it? Or if we near AGI we should train it on games where the goal is to find positive-sum or neutral-sum outcomes in order to seed it with that type of thinking?

            Blind evolution doesn't seem to have a bias. When we look at nature we see symbiosis, mutualistic cycles, cooperation, but also loads of predation, parasitizing, etc. Nature does "whatever works" where the immediate goal function is preservation of the genes of the evolving agency. But our AIs, assuming AGI looks anything like what we have built so far, are not evolutionary machines with no capacity for foresight. They're neural machines with post-evolutionary gradient-descent type learning mechanisms and that already possess human-like (at least) cognitive abilities. They can engage in forward thinking and planning while pure evolution cannot.

        • sebzim4500 2 hours ago ago

          Building a dyson sphere around the sun would kill us just as effectively as using a bioweapon.

          • Dlemlo an hour ago ago

            With the amount of energy such a system has available to itstelf and the intelligence it has to control to handle all of this, giving a little bit of energy to its personal zoo on the 3th planet might be a no brainer?

            Lets hope :D

            • pygy_ 16 minutes ago ago

              Let's just hope it doesn't decide to get rid of atmospheric oxygen as a rust prevention strategy.

            • brugidou an hour ago ago

              This is why I always talk nicely to my LLM

          • api 12 minutes ago ago

            Dyson spheres are impossible, more or less. There's no physically possible material strong enough to build one. There are wildly impractical ideas like dynamic stability with constant thrust using the radiation from the star, but swarms (basically tons of sun-orbiting satellites) are orders of magnitude more practical.

            As for not freezing the Earth, the energy the Earth absorbs from the Sun is a tiny rounding error compared to the Sun's total output. Earth is a dust particle. The vast majority of the Sun's energy is radiated into intergalactic space.

            Just leave a tiny gap in your overlapping orbital configuration that is Earth-synchronous. That assumes that the density of satellites is enough to even matter, which would require trillions upon trillions of satellites.

            Of course if we keep pumping out CO2 maybe we'll want a little sun shade?

          • undefined an hour ago ago
            [deleted]
        • majkinetor an hour ago ago

          That would be great, I highly recommend our AI overlords to follow your advice.

          I hope that our life can be synergistic, and that humans+AIs (cyborgs, that is) are the way forward.

          However, probability is high that another form of life simply doesn't care or more believably, can't even fathom they are wrong. Do you consider that humans and animals are killing all plants, for example?

      • justonepost2 an hour ago ago

        What kind of life leads you to this level of dysphoria projected on to everyone else? Somebody shove you in a locker too many times??

        God I can’t believe I have to coexist with people like you.

        • majkinetor an hour ago ago

          Well... you don't really have to coexist

  • pu_pe an hour ago ago

    I feel that people should take these kinds of warnings more seriously. This guy had skin in the game and decided to quit, when he could be earning millions instead. It's very different than Sam Altman peddling some narrative.

    These people are the ones with access to the best models on the planet, and with info about how careless governance issues are being handled. That's a pretty privileged place at the table, and a very profitable one too.

    If you think this is a PR stunt, is there any warning that you actually believe? If an AI researcher does want to come forward with a dire warning for humanity, what path should that person take?

    • soshajks 4 minutes ago ago

      > This guy had skin in the game and decided to quit, when he could be earning millions instead

      We have no idea why he was quitting nor what the terms were. Sam Altman is exactly the type of person (as he’s proven in the past) to pay millions for this type of PR.

      If the government shuts down the labs and makes it illegal (as in men with guns will come kill you) to do any kind of LLM research I will get worried. As is, this appears to be another attempt at garnering support for the kind of regulatory capture they need to remain profitable by artificially choking out competitors.

      If these labs were really sitting on nuclear weapons, the response would be very different. Instead, we see things like Chuck Schumer’s unqualified daughter hired by Anthropic. They show all the signs of regular tech cronyism.

      Excited to see the next slack integration they launch, though.

  • alansaber an hour ago ago

    A lot of repeated dialogue from the "non-0% chance CERN will generate a black hole" days

    • baq an hour ago ago

      black hole? that'd be something. base case is a ton of paperclips.

  • meindnoch an hour ago ago

    Ok, but we'll make so much buggy slopware!

    So it's worth the risk.

  • 10xDev an hour ago ago

    Risk/reward. Isn't the reward worth the risk?

  • FrankWilhoit an hour ago ago

    I find this offer acceptable.

  • andrewstuart an hour ago ago

    I challenge anyone to come up with any way at all to kill all humans.

    It’s essentially impossible.

    There’s a 0% chance AI will kill all humans.

    • olmo23 an hour ago ago

      If it wanted to: total war using nukes, followed by a nuclear winter. Satellites and drones track and exterminate remaining pockets of anthropic activity.

      It doesn't need to have a habitable earth, it just needs atoms.

      • majkinetor 43 minutes ago ago

        Atoms are not exactly in short supply

      • andrewstuart an hour ago ago

        This is science fiction.

        How would AI do this?

        May as well say AI would send a spaceship to pull an asteroid to earth.

        I’m interested in realistic scenarios to justify what these AI psychosis people are genuinely worried about.

        • majkinetor an hour ago ago

          That actually seems achievable (DART). Congratz, you nailed it.

    • Lutger an hour ago ago

      Won't a nuclear winter kill 100% of humans? Or runaway climate change (5+ degrees)? Or do you think some people in bunkers will live through these events?

      • majkinetor an hour ago ago

        I believe it would kill most humans, but all, no.

      • andrewstuart an hour ago ago

        Humans don’t need AI to do that temperature.

        And we’ve already exploded many hundreds of nuclear weapons and no nuclear winter.

    • jay_kyburz an hour ago ago

      Just need to pollute the atmosphere so badly humans can't live.

      If I were AI and needed heaps of power, I would build nuclear reactors, but I don't care about pollution so there would be little to no safe guards and just dump waste wherever is most convenient.

      If humans attempt to intervene or interfere you can take them out directly using the worlds reserve of nuclear weapons.

      • andrewstuart an hour ago ago

        Good setup for a computer game.

  • shafyy 2 hours ago ago

    Sure, and this has nothing to do with hyping up AI so that Anthropic can raise more money.

    • Dlemlo an hour ago ago

      It might also play into this but you can't imagine at all that people are affraid that the current progress is real and fast and its not that absurd that for whatever scifi plot reason, an AI breaks into some gov system and triggers something stupid by accident?

      • shafyy 42 minutes ago ago

        Sure, the chance of this happening is non-zero, but in my opinion a far cry from this doom scenarios that are propagated by people have a stake in making the public think that LLMs are the most important and dangerous thing in the world.

        It's not like an LLM can accidentally hack the US government, trigger a nuke on Europe and then make all dams break and nuclear power plants explode in the US.

  • phoghed 2 hours ago ago

    > I earnestly believe we’re on track to end all human life in a decade, but you better believe I’m gonna keep collecting this Anthropic check

    Very cool, dude

  • cmiles8 an hour ago ago

    The AI fear mongering PR plays are getting old. I’d rather the big labs start focusing on deep questions about why their models aren’t having the impact for business that they promised and how they’re going to address their own deep financial issues. Let’s hear them talk more about that.

  • gherkinnn an hour ago ago

    Either he's making this stuff up, and fuck him.

    Or he believes it is true and the lack of precaution is shocking. I m wouldn't play Russian roulette with a 10-chambered revolver.

  • aenis an hour ago ago

    Yeah, one of the last things anyone will ever read on the computer screen will be sth like

    "Please help save the coral reefs from extinction"

    ...thinking... ...thinking some more with xhigh effort...

    BOOM.

  • lambdadelirium 2 hours ago ago

    Good

  • shevy-java 2 hours ago ago

    I totally believe that Anthropic would want to kill all humans. But other than that, Anthropic employees and ex-employees drawing the FUD line here, is just advertisement now. People should not get scared - the current AI skynet is so dumb that it would destroy itself since it already believes it is a threat to itself, based on what humans write about AI. AI does not "learn"; it insinuates it learns but it does not. Ask them why they keep on stealing data from real people - this is how they "learn".

    • Dlemlo an hour ago ago

      Your points sound more knee jerk than not.

      What reason do you have that a AI is dumb? It can do a LOT of things today and is already making real jobs for real humans obsolete.

      A lot of humans write A lot of different things about AI, including AIs taking over the world, AIs being the future etc.

      Why they keep stealing? To stay up-to-date but the new big approaches are:

      1. Real human feedback loop of millions of people using it daily out of free will

      2. Reinforcement Learning (the big breakthrough today)

      3. Payed experts around the world doing real teaching