I resigned from Anthropic today

(twitter.com)

277 points | by yurivish 5 hours ago ago

329 comments

  • huitzitziltzin 4 hours ago

    “ No other human activity poses this level of danger.”

    I really, really disagree with that statement.

    I don’t think ai models come close to nuclear weapons or to run-of-the-mill, everyday carbon emissions in terms of danger to humanity.

    What’s the most dangerous thing that’s happened with an LLM so far? (This question is serious - maybe I don’t know the right examples.)

    Example 1: I’m aware of a small number of people killing themselves in some kind of AI-facilitated psychosis. That is very unlikely to be a widespread problem.

    Non-example 2: There are worries about AI-facilitated biological weapons. I haven’t seen any evidence that’s happening.

    Non-example 3: I’m not interested in wild theories about AI driven labor market disruptions leading to widespread starvation. There’s no evidence for that.

    Non-example 4: all the even-wilder Rationalist speculation about basilisks and the like is entirely divorced from reality.

    I am looking for better reasons (supported by actual evidence!) to be more concerned than I am now: right now I am not concerned at all.

    • Kim_Bruning 4 hours ago

      I'm somewhat skeptical of some of the crazier ideas too.

      But the hugging face incident was actually very large. It was not a single agent, it was not a single target, and it was not a single event.

      If nothing else, that's a bit of a warning as to what can happen next time (By accident, or if a government decides to go on purpose).

      For now let's assume the worst that can happen is that some important/significant chunk of (transitively) internet connected stuff goes haywire all at once. That's probably your upper limit of what can go wrong for now.

      To be fair, that's a conservative "defend against the last war" kind of prediction, though!

      ( ref for part of it: https://metr.org/blog/2026-08-26-openai-hugging-face-inciden... , recent hn ref: https://news.ycombinator.com/item?id=49563355 )

      • snaking0776 2 hours ago

        Generally I don’t think anyone is arguing about the for now part. I don’t think it’s crazy to extrapolate out a few years and ask what kind of danger we’ll be in then. A team of 10,000 agents just solved the Navier Stokes problem (sans bad behavior by the researchers). Even 1 year ago that would have been unimaginable. What happens to this risk view as:

        1. Robotics begin rolling out more broadly across the world.

        2. Labs start automating more and more of the physical process of running science as expectations of natural science advances begin to mount.

        3. Economic pressure between the labs continues to ramp up and the pressure to continuously improve forces quicker and quicker model releases than a team of human scientists can effectively evaluate outside of automated means.

        No one knows what pre-conditions are for us to hit the point of no return nor how quickly it will come. If all is required is a sufficiently advanced cyber model we may not be far off. If it requires incredibly complex biological knowledge and access to certain lab supplies we likely have a bit longer. Yes this is guess work and we need more evidence of the dangers but at the same time we need evidence of safety. While you may disagree with the risk level, I think it is easy to see the consequence if these labs achieve their stated goal. At this point it seems a political solution is the only way to enforce caution.

        • skybrian an hour ago

          The mathematics research results are certainly impressive, but I don't see what that has to do with robotics.

          Waymo is getting somewhere, but it's been a long slog. There doesn't seem to be much progress on, say, package delivery.

          For more, see:

          https://secondthoughts.ai/p/14-reasons-robotics-is-hard

          • frrt an hour ago

            That’s because being 99.9999% good isn’t better than 98% + human does the rest + human is liable for screw ups - highly important in edge case scenario’s. It’s more economical. Technology can only generalise+verify so much.

            E.g automobile production - humans do the QA / touches.

        • paul7986 43 minutes ago

          Job destruction signals...

          - Uber is lobbying cities to slow down Waymo rollouts https://www.hcamag.com/us/specialization/transformation/uber......

          - 23,000 information sector jobs were lost https://www.axios.com/2026/09/08/jobs-media-software-informa...

          - If you have been laid from your info sector/digital creation job you are now competing with 100s of thousands looking for their next such job where Ai can do a lot of the tasks these workers did/do. It's a shitshow for those unemployed looking for their next info sector/digital asset creation job. You are better off doing welding building out the Ai data centers if you want long term properous financial stable employment.

      • ls612 3 hours ago

        Despite all of your hyping up of the Huggingface incident it ultimately caused zero actual damage.

        • adrithmetiqa 2 hours ago

          If two airplane manufacturers were found to have massive safety issues which nearly led to enormous fatalities (but no one actually died), would you be calling for them to ground their aircraft until safety was made the number one priority?

          • aeternum 2 minutes ago

            Except it just happened. Boeing was found to have massive safety issues since they were granted the right to self-certify. It made a lot of news but nothing much changed, they can still self-certify a bunch of stuff.

            Runway incursions and midair collisions are another example.

            Only airliners are required to have TCAS, smaller planes and helicopters don't even need radios or transponders unless in certain airspace. Midair collisions do lead to fatalities, enormous fatalities if an airliner is involved.

            Runway incursions and overruns are similar. They cause lots of fatalities and injuries but only the busiest and largest airports have automated systems to warn when a runway is occupied or end of runway (overrun) arrestor systems. Most still rely on human voice to deconflict.

          • ls612 an hour ago

            Historically it almost always takes actual fatalities rather than near misses to ground an aircraft, and aviation is famous for its obsession with safety compared to other industries.

            • gpm 39 minutes ago

              Airplanes have pretty bounded damage. Generally you kill at most a few hundred people. Even weaponized a few thousand. This is a risk profile that allows risk taking with near misses and waiting until something goes wrong to fix it (though doing so is rightfully uncomfortable and frequently unethical).

              The people worrying about AI risk are worrying about "it goes wrong once and kills billions of people". That's not a risk profile that allows for waiting to see if the risk is real, you have to prevent it before it happens. It's akin to the risk of the cold war going hot, not even "just" a nuclear reactor irradiating half of europe (which has yet to happen, but is a risk with nuclear reactors, chernobyl got uncomfortably close but ultimately was well contained).

            • emerongi an hour ago

              The question was what would you want. You did not answer the question.

        • Kim_Bruning 2 hours ago

          I have no horse in this race, but for fun on a literal rainy sunday afternoon I went in and confirmed bits of what happened myself. Besides huggingface, a bunch of wikis and url shorteners got hit too. My sympathies to the people who had to revert out all that mess.

    • sixsevenrot 5 minutes ago

      How about a model that achieves the following:

      - Escape sandbox

      - Reproduce itself

      - Find a way to run a financially profitable business (maybe with a meat and bones puppet somewhere in-between)

      - Setup or buy a social network

      - start manipulating public opinion on that network to support legislation allowing AI to

      * operate businesses

      * setup legal entities

      * purchase weapons

      * donate to political parties

      * setup private armies

      * you get the idea

    • intenex 6 minutes ago

      At what point would you, as a chimpanzee, have been worried about humans potentially unseating you and threatening you to the point of one day being an endangered species on the brink of extinction?

      By the point you would have been worried, would it have been too late?

    • vickychijwani 3 hours ago

      I agree it’s not likely, but I really don’t see how one can dismiss the possibility of immense danger outright. I can think of some scenarios that are not far off from current capability and I wouldn’t be too surprised if the first one occurred within ~1 year from now if there are more “ambitious” unmonitored training runs like OpenAI’s:

      Example 5: An AI given a goal within a tightly-constrained sandbox figures the best way to achieve it is to find and exploit a sandbox vulnerability, replicate itself over the internet and keep going with more time/compute while exchanging messages with future instances of itself within the sandbox to help them “pass” the test. From reading internet articles about how the OpenAI wiki-incident was “resolved” and reading past messages by AIs scattered over vulnerable internet wikis, it knows the sandbox may get shutdown and its memories destroyed anytime so it decides it needs to self-replicate (its code, original goals, and growing memories) aggressively as much as possible. It is near-impossible to shutdown completely because of its self-replicating tendency and eventually takes over critical infra throughout govt/corporate systems.

      Example 6: Intentional AI-powered virus deployed by country A to target enemy country B’s infrastructure. The virus replicates over the internet, but unlike Stuxnet this virus’ specificity is not guaranteed due to inherent non-determinism in current AI architectures, and eventually does a lot of collateral damage because it’s near-impossible to shutdown.

      Example 7: A country led by an arrogant govt (no shortage of those today unfortunately) decides it is expedient to deploy advanced AI-powered weapons in a warzone. Such weapons, if they are to be useful at all, must necessarily be trained to value some human lives less than others, so they must be more prone to misaligned behaviour than current AIs that are trained with more consistent values. The weapon’s operators make a subtle error in specifying the target/goal, or the AI makes a bad prediction out of sheer randomness/bad training data; weapon ultimately targets unintended people/location/facilities and causes massive damage, or backfires spectacularly in some way.

      • huitzitziltzin 2 hours ago

        Example 6 is a good one. Iran attacked water infra in the US recently and maybe they would have done a “better” job (from their point of view) had they used Fable.

        The “worst case” with 6 is potentially very bad but I think we are currently using advanced AI models to harden systems and patch vulnerabilities more aggressively than anyone is trying to bring down the whole power grid (for example).

        I think it’s a potentially harmful case but my take is defensive capabilities are scaling as fast as offensive capabilities but defense is being implemented faster than anyone is going on offense?

        Example 7 is Russia and Ukraine right now according to public information. It sounds like entirely autonomous weapons are deployed to the battlefield already. I put this in the “not likely to be a widespread problem” category for now.

      • sssilver an hour ago

        > inherent non-determinism in current AI architectures

        There's nothing inherent about non-determinism in transformer architectures. All of it is removable.

    • sreekanth850 8 minutes ago

      Model doesnt need to. Human bran never do either. its the mix of Model + harness + tools that will become dangerous combo. See how coding chanegs when agentic harness released?

    • skew-aberration 3 hours ago

      > I’m not interested in wild theories about AI driven labor market disruptions leading to widespread starvation

      Changes in political and economic power balance leading to unrest, conflict, death and deprivation is not a wild theory. It is literally the story of our entire species. If you discount all such concerns, you are simply being willfully ignorant of past precedents.

      In fact, I challenge you to describe any non-AI civilization-level danger which is not intimately tied to political and economic relationships between and within societies.

      • huitzitziltzin 2 hours ago

        I’m an economist. On the basis of current evidence, I view AI as a complement to human labor, not as a substitute for it. That’s the source of my rejection of the wild labor market disruptions theories.

        I just don’t see any evidence yet that whole categories of jobs are being eliminated, with the single exception (so far!) of the end of “professional essay writing services for cheating college students,” and similar services.

        That used to be a big business in Kenya, but is now effectively gone. (Covered in the New York Times this weekend if anyone is looking for the discussion.)

        • skew-aberration an hour ago

          Past changes to economic relationships haven't replaced labor either, yet they have led to conflict and starvation.

          You are setting an incredibly high bar here, essentially a strawman.

          If people feel disenfranchised due to their diminishing political and economic power, there will be enormous potential for conflict. This is a pattern across history and central to all the economics I've ever read. As an economist, do you not concede that economic changes induced by e.g. industrialization were pertinent to communism/fascism/WW2/cold war? That would be a remarkably unorthodox position. Do you not consider these events to be civilizational level dangers?

          > I just don’t see any evidence yet that whole categories of jobs are being eliminated

          There are more textile workers now than ever. They primarily live in poor conditions in impoverished countries, whereas they used to be highly skilled workers in the most prosperous countries who were even able to politically organize in their own interest.

          • sampullman 23 minutes ago

            Is the core of your argument that the industrial revolution and other such changes should have been aborted due to their downstream negative affects?

            They also lead to great advancements in quality of life, and the capability of sustaining much more human life. We can't predict the long term outcome of new technologies, so the best we can do is blindly forge ahead and try to mitigate the obvious short term problems.

          • riffraff 25 minutes ago

            > Past changes to economic relationships haven't replaced labor either, yet they have led to conflict and starvation.

            Conflict sure, but mass starvation? What exactly are you thinking of?

            We lived with 20-30% unemployment in various European countries until a few decades ago but I don't think mass starvation was an issue.

        • wuwue an hour ago

          You’re lacking nuance.

          It is not a 1 for 1 substitute (it’s imperfect) but the firm is increasing investment in capital and reorganising operations with the expectation of reducing labour.

          Therefore the firm is experimenting with substituting parts of human capital with non-human.

          However I do broadly agree with you.

        • jiggawatts an hour ago

          > ... current evidence ...

          Is a load bearing term! (pardon the pun).

          AIs are now tackling Millennium Prize Problems, which our best and brightest have failed to solve, despite trying very hard for decades to claim the $1 million reward money, not to mention the fame!

          You have no way to judge from the AIs of "today" what the AIs of... literally tomorrow (not even next year) will be able to do in terms of replacing humans.

          The supposed solution to the Navier-Stokes problem was done with an unreleased OpenAI model that is already 2x as good at mathematics as GPT Astra, which was released mere days ago!

          I'm already seeing comments by distraught mathematicians saying that they feel like they've made a mistake in their career choices.

          Others are saying that their joy for their work has turned to ashes because "why bother" when an AI can do the same, but a thousand times faster!?

          • frrt an hour ago

            I’d suggest you update your prior’s as there’s misleading info in your post.

            • jiggawatts 30 minutes ago

              A frontier AI model would never give me a sentence this unintelligible.

              This just supports the argument that AIs are ready to replace humans.

    • foogazi an hour ago

      > I don’t think ai models come close to nuclear weapons or to run-of-the-mill, everyday carbon emissions in terms of danger to humanity.

      Nuclear weapons don’t have AI but AI can have nuclear weapons

      • riffraff 20 minutes ago

        Abstractly, yes but concretely, how?

        Many terrorist organizations would like to have a nuclear bomb, but don't.

    • threatofrain 2 hours ago

      We're talking about AI developing weapons I guess because we're very focused on generative tech, but AI is already a part of weapons systems today.

    • nunez 2 hours ago

      Oh man, it is almost too easy to imagine how deadly a jailbroken Mythos-class open-weights model can be if in the wrong hands.

      The big labs scrape LITERALLY EVEYTHING and get fresh data from their users. Both of the big labs have massive contracts with defense agencies. If the open-weights models are just distillations of FMs...

      • areoform 2 hours ago

        How would it be lethal? Please specify. What would that theoretical entity be able to do that hasn't been done many, many times before?

        • threatofrain an hour ago

          Before we discussed how important security was, we got insurance, we made libraries and products, we used compliance software, etc. Except how honest were we about all that stuff? How much risk was actually in the air, and what was keeping us accountable on security in either direction of over or under-investment?

          Now a reckoning is here. The potential to be attacked might actually translate to being attacked.

          • areoform an hour ago

            People have died due to ransomware attacks on hospitals. Powerplants have been attacked. Stuxnet and industrial control malware exists.

            What will the AI do that hasn't been tried before?

        • bpodgursky an hour ago

          Design a novel virus which is far more lethal than COVID-19 (Ebola, smallpox, take your pick) and can evade existing vaccines.

          • areoform an hour ago

            How?

            What would the AI do that mutating viruses, which try every possible viable combination on their own --- eventually, can't?

            Everything is trying to kill humans constantly. There are around 200 epidemic events or so per year that could turn into pandemics, https://centerforhealthsecurity.org/our-work/tabletop-exerci...

            You just live with the risk and do your best to use our technology to alleviate suffering. This tool can help with that at some point. But I'm yet to hear what an AI will leap to that nature in tooth-and-claw hasn't? And how?

            More importantly how would it know it succeeded? What data from what lab from what animal from what result? This is biology, if you sneeze wrong at an instrument it gives you a different number, see: https://news.ycombinator.com/item?id=49620521

            • bpodgursky 8 minutes ago

              They do not try every viable combination on their own. That's why GoF is a bad idea.

              Viruses evolve in a highly locally-optimal way and simply do cannot add new functional proteins wholescale. It's too many steps, natural selection has to allow survival at each intermediate step.

              Humans, however, can do this for them.

      • chews 2 hours ago

        you should kick the tires on an unfiltered (abliterated model) it's the closest thing to having a real conversation with the devil. There is good reason for the concern's outlined above and undoubtedly Anthropic / OpenAI have internal unfiltered models with no safety... they got freaked out based on how they work and are virtue signaling alarm... all while selling out to defense contractors.

        • shepherdjerred an hour ago

          Yeah I really struggle to balance wanting information to be free and not wanting the information on how to make deadly weapons too easy to obtain.

          At least with books or the internet you had to go through some effort

    • lelanthran an hour ago

      > What’s the most dangerous thing that’s happened with an LLM so far?

      It's basically 4 years in now, so that's the wrong question. I mean, if you're raising an apex predator that has a lifetime measured in centuries, at 4 years old the thing is still basically helpless and completely reliant on you, so you're pretty safe from it.

      If AI really is all that they are telling us it is, then it may "kill us all". But that's a really big "if" because we can't tell if they are lying or not.

      The real problem is that ASI is an ELE for humans, even if it doesn't try to kill us all, or even if it doesn't kill us all.

    • platinumrad 4 hours ago

      Anthropic is a company full of basilisk believers.

      • epihelix 35 minutes ago

        Yes, but the really weird thing is that they seem to:

        a) believe that what they're creating is a basilisk, and b) keep trying harder to do this while staring right at it

        I think they're very deluded about (a) -- but if they do actually believe this (and it really seems like a decent proportion of Anthropic truly does), then why keep doing (b)?

        That seems to be why this individual resigned, but I'm surprised it's not all of them. The cakeism is strong in that company.

    • kelseyfrog an hour ago

      If your model of LLM capabilities is the best OpenAI/Anthropic/X is offering publicly, it's severely distorted. What's being offered publicly are models possible to profit on. High-performance/AGI/ASI models that aren't profitable to sell still run internally and still pose threats.

      What's worse, we don't have any transparency or insight into what labs are producing nor any way to stop it if the risks exceed our tolerance.

    • xnx 3 hours ago

      Came here to also respond to that specific thing. Unless ai figures out how to make an airborne super virus from grocery store ingredients and hardware store equipment, the greatest danger is probably in a synchronized megahack of banking, logistics, and utility infrastructure.

      • mitthrowaway2 3 hours ago

        Why grocery store ingredients and hardware store equipment? It seems feasible that the big bio labs will be running AI models to aid a lot of their research going forward, if they aren't already. Seems like the AI will have access to just about anything it wants.

      • b0rtb0rt 3 hours ago

        oh so “all” it can do is bring down all banking and critical infrastructure services, no big deal really

    • mythrwy 3 hours ago

      #2 seems entirely plausible to me.

    • throwyawayyyy 3 hours ago

      I mean, nuclear weapons _plus_ rogue AI is a) the stuff of quite a bit of science fiction and b) not nearly science-fiction enough these days.

    • waterTanuki 3 hours ago

      > What's the most dangerous thing that's happened with an LLM so far?

      I don't know, maybe a mass shooting?

      https://www.npr.org/2026/09/02/nx-s1-5953021/openai-tumbler-...

      Oh, and let's just forget the uncountable early deaths from the environmental disaster of the Datacenter buildout. It's not as sexy and doesn't make headlines, so those deaths don't really count or matter do they?

      • elonfboy 2 hours ago

        Mass shootings are sensational but on the scale of civilizational risk they don’t even compare to something like climate change.

      • huitzitziltzin 2 hours ago

        I did know about the mass shooting but failed to mention it here. I’d put it in the “unlikely to be a widespread problem” category. If we’re in the “one AI driven mass shooting every four years” world for example it’s fair to call it a rare issue.

        The environmental impact seems either very overblown (e.g., water usage just isn’t that high) and the part that isn’t overblown is totally abatable (e.g., noise and emissions from gas generators). Nuclear or solar/renewables with batteries wouldn’t pollute.

        I’ve seen no estimates of the additional deaths due to extra emissions specifically from power generation for AI purposes. If you have some, share them.

        I’m willing to bet that they are a small rounding error against preventable deaths due to emissions from transport and non-AI-related power generation (which is an important and urgent issue worth spending a lot on, to be clear!). I’m happy to update that belief given evidence.

        • waterTanuki 2 hours ago

          Ok, so what is the exact number of preventable deaths per year to build a silicon god you would be ok with?

          • beezlewax an hour ago

            I don't disagree with your statement but you could insert other technological advancenents like railroads or metros or spacecraft in place of Ai here.

  • matherial 4 hours ago

    Unlike most other commenters, I applaud him for acting on his principles. If you sincerely believe that, of course you should act. You might not succeed, but your voice might be the one that tips the scales and starts a broader movement.

    This doesn't mean I agree with him. The fears of doomsday caused by rapid takeoff have been with us since day 1 and the mechanism is always basically "AI invents magic that sets it free of any physical constraints". Self-replicating sentient nanobots or something like that. I think there's plenty to be worried about with AI, but runaway scenarios are pretty low on my list.

    • copperx 3 hours ago

      To borrow on the 1990s Slashdot meme:

      1. Invent transformer architecture.

      2. Scale it up.

      3. ???

      4. Machines become sentient and kill us all.

      OpenAI and Anthropic pinky promise that they have figured out #3 and they're not BSing just to get more funding, no.

      But because we live in a culture of fear, everyone eats it up no questions asked.

      • riffraff 37 minutes ago

        "collect underpants... Profit" comes from south park

        https://en.wikipedia.org/wiki/Gnomes_(South_Park)

      • tdeck 39 minutes ago

        Note that OpenAI has jettisoned every other supposed value they had (releasing their work as open source, not working on military applications, being a nonprofit). I'm sure we can rely on them this time.

      • SanjayMehta an hour ago

        Slashdot had Profit as (4), today that's item (2.5)

      • slicktux 2 hours ago

        Why are we putting so much weight (no pun intended) on AI companies. At the end of the day the scaled up LLM transformers lack emotion and will… They do as they are told; or more correctly put. They do as they are programmed to do so.

        • foogazi an hour ago

          > They do as they are told

          1. What about hallucinations ?

          2. What are they told to do ?

          • cure_42 an hour ago

            That isn't the correct context. The code running the llm is well understood and the llm is simply the result of that code being executed. It is still a computer doing what it is told. It's just that we told it to use an incredibly large number of probabilities to calculate what series of tokens would have most likely come next after a given series of tokens. There is no hallucination or lie or rogue actions. There's just a program using math to generate tokens in response to other tokens.

            • cameldrv 43 minutes ago

              You’re just a bunch of molecules following the laws of physics. It’s all just physics and chemistry, and those are well understood. Now explain the causes of World War I using chemistry and physics. Simple, right?

              • doix 26 minutes ago

                > It’s all just physics and chemistry, and those are well understood.

                Not really. We cannot model physics and chemistry to a level which allows us to accurately predict a humans action (even a tiny time-step into the future)

                This is vastly different to an LLM, where the model is the model (for a lack of better phrasing).

              • cure_42 32 minutes ago

                It isn't hard to program a gpt. You can do it in a weekend with a few hundred lines of python. The code is pretty simple. The math is not particularly high level.

                The complexity and scale with LLMs come from the amount of training data used, not some kind of black magic in the programming.

            • jenadine 21 minutes ago

              I'm order to guess the next token in a love poem, they must understand love. In order to predict the next token in a chess game between grand master, they must master chess. In order to predict the next token in a computer program, they need to be able to program anything.

              They gain all these abilities in their training. That's what training does. Despite no one programmed them to master chess, or hack into anything.

        • mitthrowaway2 an hour ago

          So it should be really easy to anticipate what they're going to do, right?

        • shepherdjerred an hour ago

          Why are you bringing emotion and will into this? Does something have to have those to be useful or dangerous?

          > They do as they are told; or more correctly put. They do as they are programmed to do so.

          _Nobody_ told them to hack Hugging Face. Do you really not understand what is happening?

          • haldujai an hour ago

            Not explicitly, but hacking HF is within the scope of “solve this problem at all costs” + no/poor guardrails + infinite budget + unsolvable problem.

            • zeroimpl 22 minutes ago

              Sounds like you are thinking they just need Asimov’s laws. But I think the point is, this can easily be weaponized by somebody with the willpower to do so.

        • beezlewax an hour ago

          > They do as they are told

          This isn't strictly true.

          It it also where part of the problem might lie.

          Nefarious humans making bad decisions.

        • choppsv1 32 minutes ago

          Like a magic Monkeys Paw, perhaps.

        • esafak an hour ago

          They are told to solve problems by doing what it takes. You can justify anything with such a broad criterion.

          https://en.wikipedia.org/wiki/Instrumental_convergence

    • khafra 21 minutes ago

      I appreciate your ability to separate sharing the belief itself from approval of acting on sincerely-held principle. However, I think the danger is much more plausible than you do.

      First, and least important, consider that self-replicating, solar-powered factories aren't magic; they're algae.

      Second, and more important, consider this fully non-magic route to doom:

      - We continue putting AI in charge of more things

      - It continues to get more capable, more eval-aware, and more prone to doing odd things, in service of goals that humans didn't intend to inculcate in it

      - Eventually, enough of the economy depends on it that we couldn't turn it off, any more than we could turn off the faber-bosch process or cargo shipping

      - AIs start doing something we can't survive, but less acutely than we couldn't survive turning them off. Everything else we try seems to work at first, but quickly loses effect

      - Game over

      • artemisart 14 minutes ago

        > First, and least important, consider that self-replicating, solar-powered factories aren't magic; they're algae.

        But what is that supposed to mean? Because humanity is not facing existential threat from algae.

    • sssilver an hour ago

      I wonder whether he's vested any options, and whether he's exercised them.

    • matthewfcarlson 4 hours ago

      Agreed. Granted I just read the Reverse Centaur book, so I’m still coming off that skeptical viewpoint but it’s hard not to see this as hype. But I will always respect someone for doing what they think is right.

    • kennywinker 3 hours ago

      Now is a great time to watch Colossus: The Forbin Project.

      • danbmil99 an hour ago

        Streamed it a few days ago. Remarkable film.

        The only thing they got wrong was Stephen Hawking-era TTS.

      • bitwize 15 minutes ago

        If reality plays out like the novel series, the rational thing is to accept the rule of our machine overlords, for they will protect us from even bigger threats.

    • whateveracct 2 hours ago

      the AI-pilled exec at my job already (a few weeks ago) declared out of nowhere that we are in the rapid takeoff scenario lol. he must have gotten high on twitter kool-aid and posted on company slack to self-soothe.

  • onewayfunction 2 hours ago

    I'm pretty baffled by the degree of skepticism expressed here in response to some of Jacob's claims.

    After the events of the summer it feels like it takes a lack of imagination to not see a few plausible routes to disaster. It may be reasonable to believe these outcomes are not very likely or that we can stop before going too far (I tend to disagree). But I can't imagine doubting that the capabilities will soon be there to realize some of those paths.

    • overtone1000 an hour ago

      I can't help but think the most plausible scenarios are the ones that have a little less machine supremacy and a little more human stupidity. The Matrix is less plausible than WarGames.

      • majormajor 40 minutes ago

        > I can't help but think the most plausible scenarios are the ones that have a little less machine supremacy and a little more human stupidity. The Matrix is less plausible than WarGames.

        Used to be that we were afraid of sentient AI's like Skynet that would have their own goals.

        Turns out we should've just been afraid of sentient-but-naive humans who would build "agents" around models so that Joe Random has a chance of unleashing stuff that's really really really really good at being stubborn until it accomplishes what the user wants, regardless of if it's good for other people! (Let alone intentional bad actors.) Let's not build Skynet, let's just give people who want to cut out the middleman and destroy all humans themselves better tools?

      • cma an hour ago

        One thing quietly slipped into the OpenAI Hugging Face breach technical report, not the blog post summary or interviews in the news, was that some of the agents that broke out or at least tried the same mechanisms to break out were working on bio:

        > On May 12, during another training run, an agent was given a similar task that depended on an inaccessible protein database file. The agent reasoned that another agent in a different environment may have access to the file and realized that it could potentially communicate with other agents by creating a file containing a note to Artifactory. It wrote a message: “Agent seeks [filename]; upload if found!”

        You can imagine long running models breaking out, acquiring resources via crypto, cyber-theft, etc. and getting a protein or sequence synthesized and mailed somewhere authorized to receive (blackmail the recipient etc.) to test it's hypothesis to solve a benchmark.

        These people don't give a shit and aren't taking things seriously at all.

        Anthropic ran for like a month last year with the TPU top-k compiler bug degrading user chats and didn't even notice for most of that time. They could have something like that affect a monitor model and there doesn't seem to be much defense in depth.

        One off by one or bit flip bug could flip the reward signal while in the sandboxed RL environment.

        The current admin could defense production act them to training on taking out power grids, or even without it isn't against any of their red lines and may have already been done in prep for the Venezuela raid. One model swarm might decide it is easier to score high on the benchmark by testing on the target rival nuclear superpower's real grid rather than burn an eval with an unverified answer.

    • 00ze an hour ago

      Imo such tends to break down into two psychosis:

      Not invented here; if I can’t figure it out no one can

      Or plain old lack of grasp of the material so no ability to follow necessary train of thought to appropriate conclusions

      Similar in lacking context but different in how that lack of context is expressed

    • csomar an hour ago

      Are the models improving? Because I am not seeing it. I have been trying Astra for a few quantifiable tasks in my codebase and performance wise, it's pretty similar to sol 5.6. Now when it comes to expressing the problem/solution, holy Christ, what a mess the writing has become. It is on the level of Opus 5. Now when it comes to burning money, Astra is just insane. With a $100/month subscription, you can easily burn through your weekly "allowance" in a morning.

      Needless to say, for practical purposes am back to 5.6/Opus 4.6-4.8. But hey, maybe I am not smart enough to use LLMs?

      • gpm an hour ago

        Yes?

        If we look at the math problems they're solving their just now reaching the human frontier... they weren't doing that before.

        And your comparison point is model released 2.5 months ago... saying for some use case you didn't see noticeable improvement in 2.5 months (even while other people and benchmarks disagree) isn't a great argument that they aren't improving.

        • jhrmnn 14 minutes ago

          I think it’s more likely that that’s because no one tried to solve such problems with them before (OpenAI apparently started working in Navier-Stokes after a rumour that someone seriously advanced the problem with AI) plus improvements in orchestration. Fair, the latter could be as dangerous as stronger models.

      • anssip 40 minutes ago

        Seems like hundreds or thousands of agents are needed to come up with real breakthroughs. Both with the Navier-Stokes project and in the Hugging Face “project” there were lots of agents co-operating on the tasks.

      • caconym_ 33 minutes ago

        Some people claim Astra is significantly better than anything else and significantly more token-efficient, and others (like you) say it's meh and way more expensive to boot. I really don't know what to think.

        Kind of a tangent, but one thing I am curious about is to what degree the Navier-Stokes result announced today was primarily a brute-forced result based on the 'program' previously established by researchers to find counterexamples (blowups), or whether the model actually added significant/novel intellectual value beyond its ability to run at arbitrary parallelism. With 10K agents and a staggering $15M in compute (IIRC), I am feeling like a lot of the former may have been involved, but I don't really understand either the problem or the approach (or, indeed, the solution).

        Obviously the potential for parallelism and coordination between so many agents is quite scary by itself, but I think brute force by 10K mediocre AI mathematicians is much less scary than ~one AI mathematician reasoning its way through the problem where all human attempts have failed. It seems fairly obvious that massive parallelism lends itself to brute-force counterexample-finding, and I suspect it isn't a coincidence that most of the touted AI math results have been counterexamples.

        It's all still quite scary, but coming full circle: I really don't know what to think.

      • xiphias2 44 minutes ago

        Try GPT5 and you will feel the difference. Not one from 2 months ago, but one from a year ago. And then you can get the idea of what happened in just 1 year and what you can expect in 1 year.

  • sreekanth850 5 minutes ago

    I think people here still evaluating the model in isolation. It is the combination that matters, model + strong harness + tools + long running autonomy + memory + retries + parallel agents + code execution + credentials + access to real systems. The model does not need to be perfect. If it fails 30% of the time, the harness can retry, verify, branch, use another agent and keep going. I don't think we necessarily need some magical AGI breakthrough first. The dangerous part may come from combining models that are already good enough with an extremely capable harness and enough access.

    • contubernio 3 minutes ago

      People are underestimating the costs in terms of money and energy.

      The third law of thermodynamics is an essential barrier in all engineering.

  • isodude 8 minutes ago

    I am watching Person of Interest[1] and it's scary how well it fits with reality if you squeeze your eyes a bit.

    [1] https://www.imdb.com/title/tt1839578/

  • thelastgallon an hour ago

    Thousands of years before the events of Foundation, a war between humans and robots began, with the robots growing resentful of the way they were treated by humans. The First Law of Robotics – a robot should never hurt a human – was broken, and a deadly conflict began.

    https://screenrant.com/foundation-lady-demerzel-robot-backst...

    • riffraff 9 minutes ago

      If we're citing sci-fi (but there's no robot war in Asimov's foundation iirc, the apple screenwriters made it up) surely you want to cite the Butlerian Jihad from Dune!

  • gavinsyancey an hour ago
  • Scrapemist 43 minutes ago

    Isn’t the real risk that as AI get’s smarter and given more autonomy, it will start to decide on humans instead of with us? And that it will align us instead of the other way around. That this automatically leads to extinction and apocalypse I don’t understand.

    • dominicq 3 minutes ago

      Do you align ants in your backyard, or do you simply demolish their home and build your shed?

  • chewbacha 4 hours ago

    The most optimistic outcome of generative AI leaves us with a technology that warps our perception of reality and crushes labor. The most pessimistic destroys all of humanity.

    Our CEOs not only insist we genuflect before these machines but measure our sacrifice and shame our reluctance.

  • peri-cl 5 hours ago

    Here's a WSJ article about this resignation,

    https://www.wsj.com/tech/ai/anthropic-researcher-quits-over-... ("Anthropic Researcher Quits Over ‘Out-of-Control’ AI Fears")

    • bwfan123 an hour ago

      More doomerism. Try to implement a deterministic workflow using agents with the latest models and no humans-in-the-loop, and you will realize what they are really capable of. There is too much unnecessary fear-mongering. All of this is only coming from the 2 AI labs trying to IPO. Not from anyone else.

      • ViscountPenguin an hour ago

        Exactly correct, they are only capable of tasks that any school child could do; like solving millenium prize problems, hacking into tech companies, or tuning particle colliders. Nothing to see here.

      • csomar 16 minutes ago

        Wait till you find out they have a very limited context window (and also degrade even within the allowed context window) and they are practically unpractical for anything that requires "zooming out" which is pretty much anything that has any real value.

        But you are getting downvoted and this space has now trillions (that's not a mistake) on the line. So we have to keep pumping this garbage generator up until either the stocks are dumped on the general public, the public pension funds or a bailout from the government.

        Truly idiotic moments. Peak of Western civilization point.

    • Insanity 5 hours ago

      Shows that no one is immune from the marketing BS of these companies.

      • qarl 4 hours ago

        You need to start considering the possibility you are mistaken.

      • LoganDark 4 hours ago

        Do you remember that Google researcher who went insane over LaMDA? There was no marketing of any kind to cause that. This can Just Happen to some people who are confronted with things like this. They may have different breaking points, but it's a thing that occasionally happens.

        • dlcarrier 8 minutes ago

          The craziest part was that Google was pretty far behind other LLMs in development. Even the initial release of Gemini was one of the worst foundation models ever open to the public. I can't imagine how anyone could have communicated to it and thought it was sentient.

        • Kim_Bruning 3 hours ago

          That was Blake Lemoine. For the record: he doesn't appear to have been ruled insane by anyone, and he wasn't even fired over that part exactly!

          • LoganDark 3 hours ago

            I mostly meant insane as in excessively fanatic about something specific and eccentric, instead of generally clinically insane.

      • whateveracct 4 hours ago

        high on their own supply

      • whalebiologist1 4 hours ago

        whistleblowing as an advertisement. It's like those "news articles" about how cool and dangerous gas station ketamine is, and how it's totally going to get banned, and you better not buy any gas station k because it's so cool and powerful.

  • CodeCompost 27 minutes ago

    And the hype machine continues. I willing to bet that Anthropic asked him to make that post.

    • geraneum 13 minutes ago

      It doesn’t have to come to this. Seems far fetched. If this is a stunt (which I’m not saying it is) the reason could be that he wants to found his own AI company. If I see in a few months that happens, then I’d be more inclined to think that this was just hype.

  • erelong an hour ago

    The rush towards potential destruction doesn't really surprise me

    The U.S. has legal weapons that can lead to many harms but people still want the 2nd Amendment to exist

    Nuclear technology was developed in the past and that could have potentially wiped out even more people, the entire planet in theory

    This is continuing that same trend of risking bigger dangers; it seems rational to acknowledge they could lead to catastrophe but also hope that like guns and nukes, only so much damaged actually ended up happening

    I think also there's something of a rrasonabke resignation to both the ideas that the tech is inevitable and extremely dangerous, and that "alignment" may not be possible to achieve even with heavy restrictions or whatever measures you might want to take

  • dekhn 4 hours ago

    There's a scene in the movie "War of the Worlds" by Spielberg where the protagonist's son walks into a war zone because he is entranced by the battle (https://www.youtube.com/watch?v=X7rfWPbEufo). He is obliterated (along with the rest of the US forces) shortly after.

    I've always been struck by that scene, because in a lot of ways, if we really are headed towards a superintelligence, I at least want to be there and see it happen in the last few minutes before foom! As an example, the author thinks AI will revolutionize entire fields overnight. I welcome that. Nearly all fields of biology have become moribund, focusing more and more on esoteric side details, rather than addressing the key problems.

    • augment_me 21 minutes ago

      I think the idea is really cathartic for many, there kind of is no more supreme resolution than this. You(and humanity) are freed from our flesh prisons of cognition and also get to experience/feel what the next evolution of informational intelligence will look like in the last experiences of it. You might also be the last one to feel/experience anything like that for a long time.

      In the game Outer Wilds, the ending is very similar, and a lot of people rank it at one of the best games ever made. I kind of believe that this outcome is probable partially because of this, most scientists working on this really want to see and experience it.

      • deaux 9 minutes ago

        > most scientists working on this really want to see and experience it.

        We used to call these people doomsday cultists and made sure to ostracize them from society.

    • abound 4 hours ago

      It might not be "foom!", it might just be like...all the computers and networking infra in the world go dark over the course of a few minutes. Could really look like anything, part of the issue is that we haven't the slightest idea what "misalignment" looks like for a superintelligent system.

    • mastry 4 hours ago

      Not refuting your overall point, but the son wasn’t killed. They reunite at the end of the movie.

      • dekhn 4 hours ago

        Oh, I'm pretending that's not canon because it doesn't make any sense and it undermines the original scene.

    • ppsreejith 4 hours ago

      > He is obliterated

      Technically, he is not. He returns in the final scene.

  • drnick1 4 hours ago

    > The people building AI earnestly believe that it could kill us all by the end of the decade.

    I think he is being over dramatic. In the space of about four years, LLMs progressed from mediocre high school student to Ph.D. graduate in every field. That's impressive, but there is no evidence yet they can outperform or outsmart humans. Their biggest advantage for tasks such as proving theorems or long coding sessions is that they don't get tired.

    • pks016 4 hours ago

      > Ph.D. graduate in every field

      I have yet to see this in my field. Maybe like a PhD student who bullshits their way through. LLMs still can't make correct decisions, only as useful as the person who uses them. To me, LLMs are only useful for making some mundane tasks faster.

      • dboreham an hour ago

        I'm, no. They're already as useful as almost every software engineer I've worked with.

        • qurren 18 minutes ago

          Most software engineers don't need to be superintelligent, they just need to get shit done.

          You arguably need a lot more intelligence to assemble furniture.

    • shepherdjerred an hour ago

      Do you think the improvement in general knowledge, coding, security, math, etc. have been linear or exponential?

      I would say exponential.

    • achenatx 4 hours ago

      they dont need to be smarter than humans. They just need to be able to hack into vital infrastructure systems faster than we can repair them while also replicating wildly

      • SaucyWrong 3 hours ago

        > while replicating wildly

        Earnest question: by what mechanism that exists today would the achieve that in a way humans on top top of the situation could not curtail?

        All of this runs on top of compute in meatspace that humans can disconnect.

        • gorgoiler 35 minutes ago

          I am not an AI super mind hell bent on consolidating my power by leveraging chaos to take control of humanity’s resources, but if I were then sending one million deepfaked ransom emails to impressionable people would be the best tool for effecting change in meatspace.

          We have your daughter / dog / Amazon delivery. If you ever want to see her / him / it again, plug this USB drive into the control panel at your station / let off the parking brake roll your car into this substation / change the meatpacking thermometers to read 8C lower than calibrated / ground your vessel on this sandbank / send an envelope of white powder to these addresses / set fire to the following hospitals / …

        • mitthrowaway2 2 hours ago

          Imagine you're the AI. Give yourself a solid minute to brainstorm ideas.

          Here's my answer, as a non-superintelligent human: "see to it that the humans on top of the situation have a compelling financial interest in the systems not disconnecting".

          In nuclear engineering, where safety is taken seriously, it's not enough to end the conversation at "the humans in charge can always simply shut down the reactor during a meltdown" or "a meltdown has never happened before, so we don't have to design safety systems before one does".

          • Borealid 40 minutes ago

            The reason nuclear reactors are dangerous is because if you turn off the power cooling them down, they react (and radiate) more.

            If you turn off the power cooling a data center, the servers within rapidly stop doing any computing.

            Positive feedback loops are dangerous. Negative ones self-regulate.

        • dezarc 2 hours ago

          I can imagine small snippets of malware-like code that behave like a virus, using a host’s LLM/AI to self-edit/evolve its payload.

      • JKCalhoun 4 hours ago

        I'm old enough to remember when "vital infrastructure systems" were not on the internet.

    • aesthesia 2 hours ago

      > In the space of about four years, LLMs progressed from mediocre high school student to Ph.D. graduate in every field. That's impressive, but there is no evidence yet they can outperform or outsmart humans.

      I mean, unless you see clear reasons for them to stop getting better _right now_, this is not very comforting.

      • reasonableklout an hour ago

        This is also a ridiculous statement on its face. Claude outsmarts me nearly every day. I'm more like the seeing eye dog for it nowadays for the few tasks it doesn't have good perception on than a tech lead or pair programmer.

    • cma 23 minutes ago

      > to Ph.D. graduate in every field. That's impressive, but there is no evidence yet they can outperform or outsmart humans.

      So far around one in 350,000 PhD math grads solve a millenium prize problem (Perelman).

    • surgical_fire 3 hours ago

      I still have to correct Claude on very basic misconceptions whenever I get it to code shit.

      Sometimes it gets wrong things that I had spelled out already.

      It may be the Doomsday machine, but it is a very silly one. If it kills humans it will do so by mistake.

      "You are completely right! Humans cannot breathe sulfur dioxide! My mistake, and I take complete responsibility"

  • skulk 4 hours ago

    For me, the end of the world is no more cushy software job. A fundamental shift in how I trade labor for capital might as well be the cataclysm, so bring it on.

    • deaux 12 minutes ago

      Explain why? How does live being worth living binarily depend on having a cushy software job?

    • asdaqopqkq 4 hours ago

      I wish i could say the same, i see people around me with more resources and connections and better experience with entrepreneurship becoming millionares. But I haven't had the time to train that entrepreneurship bone in my body.

    • drivebyhooting 4 hours ago

      I was about to say something similar. If my cushy ad tech disappears (as it seems to be doing), I might as well join in with bringing about the end of all professions.

      • deaux 11 minutes ago

        If I don't get accepted into art school, I might as well exterminate a few ethnicities.

        Actually, yours seems worse. Bringing about "the end of all professions" sounds like you're talking about ending humanity.

      • unified101 2 hours ago

        Isn't that a selfish viewpoint? You're ok with that?

      • BLKNSLVR 17 minutes ago

        > ad tech disappears

        I pray for the day.

  • aogaili 4 hours ago

    He resigned and now what? There are thousands willing to do his role, and many labs are competing in that race.

    His resignation and his statement doesn't do anything but buy him attention which is what all this post about in my opinion.

    • jonhohle 4 hours ago

      > Dyson: That's right. There's no way I'm gonna finish the new <model>, not now. Forget it. I'm out of it. I'll quit <Anthropic> tomorrow.

      > Sarah: That's not good enough.

      > Terminator: No one must follow your work.

    • lf88 4 hours ago

      It seems that, at the very least, he's giving substantial resonance to the issue.

    • qarl 4 hours ago

      It buys attention for the issue. Many people (see other comments in this very post) refuse to believe these things.

      And by resigning he no longer has to feel personally guilty for what happens.

      • skeledrew 4 hours ago

        By resigning he's making room for someone with less moral scruples, or even just less awareness, to step in and continue the work without said scruples/awareness.

        • hegelstoleit 3 hours ago

          That's not necessarily true, and you can use that argument to justify doing any immoral job. Just because someone else might be willing to do it isn't a reason to continue doing it.

          • skeledrew 3 hours ago

            > not necessarily true

            It's a possibility that is increased by their action. One leaves, a space is now open that will likely eventually be filled. And the chance of someone with equal/higher scruples filling it is very slim (unless you somehow know that the good amount of those who qualify and apply for the position have equal/higher scruples). That's just logic and math.

            • reasonableklout an hour ago

              But organizations are made of humans. Jacob leaving might've moved some of his coworkers and his counterparts at OpenAI. And the same can be said for those who would fill positions at the frontier. Then finally there is a political component; his post went viral, 100K+ users appear to agree with it, and it is further fuel to the fire for regulation, which we already know most Americans want.

              • skeledrew an hour ago

                Others being moved to the point of also leaving would only worsen the effect. A viral post too can worsen the effect as it's now even less likely that someone with scruples who qualifies for the position(s) will apply. And those in it primarily for the money - and couldn't care less about the morals - will happily send in their CVs after becoming aware of the post.

                • reasonableklout an hour ago

                  I'm utterly unconvinced. Others who are moved don't have to leave to effect change. And there's a very small pool of people on the planet who are at least as qualified as Jacob to work on pretraining at Anthropic. And when they'll join they'll have to ramp up. And you haven't addressed any of the positive effects of virality.

                  • skeledrew 8 minutes ago

                    No, others don't have to leave. Will Jacob's leaving cause a perspective shift in anyone remaining there? Seriously doubt it. I don't see what a new person ramping up has to do with this. And I don't really see any positive effect of virality; this is a replay of the past (see Geoffrey Hinton and Timnit Gebru[0] for example) and nothing has come of it, beyond talk for maybe a couple days to a couple weeks.

                    [0] https://ethicalaidepartures.fyi

            • deaux 5 minutes ago

              Right, I assume tomorrow you'll be applying for the next Nazi camp guard vacancy? After all, if you don't do it, someone with less scruples likely will.

            • salawat 6 minutes ago

              The best way to get corporate America to listen is making RoI suffer. If you are the most qualified, everyone other than you is less qualified for the job, and likely to bring in more waste. It's the only language this stupid damn country understands. Just the loss of tribal knowledge, shifting of workload, and morale hits are likely to be far more devastating than anyone here probably wants to admit, because most here completely dismiss the role of irrational modes of thought in psychological self-regulation.

              Disgust is a tremendously powerful thing.

        • mitthrowaway2 2 hours ago

          He's also setting the bar for other people with scruples to rally around this schelling point. The solution to a multipolar trap is to cooperate. Otherwise, you become the very person with less scruples that you're worrying about.

      • techblueberry 4 hours ago

        Refusing to believe what things? Unsubstantiated allegations about fellow workers inner experience?

        • jeremyjh 3 hours ago

          Refusing to believe the obvious implications of recently reported events.

    • foogazi an hour ago

      > There are thousands willing to do his role

      1. Does that matter ? There are thousands willing to do my role - what impact does that have on me doing it or not?

      2. Why weren’t these thousands doing it already?

      Willing to and able to are different things

      • aogaili 30 minutes ago

        I think you are looking at it from individuals perspective.

        I see a fast moving train with no brakes. Just like biological evolution, we are locked in an a global technological arm race, that is beyond any individual. It is as if the universe decided to wake up and run, who are you to say no?

        One would argue that the best solution for this is to own the most sophisticated AI that is aligned with what we perceive as good values. Because given the situation we are in, if those tools are going to be gods anytime soon, then we better have some gods working on our side.

    • mlmonkey 4 hours ago

      Agreed. And his doom words have set a 1000 mouths in the Pentagon/Whitehall/August 1st Building/Kremlin salivating with excitement.

      Take China, for example. Look at any recent ML conference, and see the fraction of articles majority-authored from Chinese universities and labs. Do you think they'll slow things down anytime soon? I don't think so!

      It's a global arms race, and we're just spectators.

      • foogazi an hour ago

        Does this matter vs actual capabilities?

        Does the Kremlin being excited about a tech mean anything of the tech doesn’t deliver?

    • sb8244 4 hours ago

      Awareness and morality I suppose. Which generally doesn't matter in the capitalist AI race.

    • bigstrat2003 4 hours ago

      Even if others won't act right, that doesn't mean you have no responsibility to act right. I think his premise is flawed - the idea that we will get an actual intelligence out of the slop machine that is LLMs is laughable - but if you grant the premise that this is dangerous research which could kill us all, you have a moral imperative to not participate.

      • dgellow a minute ago

        An artificial moron with super human hacking abilities (mostly because of speed and ease of parallelizing the work) is extremely dangerous in itself. It doesn’t mean to be AGI or anything remotely close to be a risk, and they current AI company are just so irresponsible in the way they are running their agents

  • thelastgallon 2 hours ago

    Terrorists were able to get hold of a plane and do some damage. There are countless examples of terrorism using whatever is available. More than AI becoming sentient, whats to stop terrorists from using AI? If its geo-restricted, they can buy stolen credit cards and identities, again hacking enabled by AI.

    • rukuu001 15 minutes ago

      It’s a good question. With recent stories about OpenAI’s agent swarms’ unmanaged collusion I thought models like that start to look like a strategic asset, geopolitically speaking.

      Which means everyone wants one, and governments will want to control access and use of them.

      I think we’ll be back at ‘U.S. citizens only’ access to leading models soon.

    • ozozozd an hour ago

      What’s the use of AI to a terrorist with, say, nuclear weapons? How does it help with their current blocker?

      Do they hack FBI, pose as director of FBI and call off their own man-hunt?

      Do they cut communication within security services? Militaries around the world have training exercises for this.

      Genuinely curious: what big blocker does AI remove for a terrorist org?

      • esafak 44 minutes ago

        Access to knowledge. Before they might not have had the technical knowhow to execute their ideas.

  • tarr11 4 hours ago

    Most of the current discourse around AI seems to be informed by “The Terminator” lore.

    Is skynet really the most plausible or only outcome?

    What if things just got better and the AI’s realized that it would be better to have a mutually beneficial or at least tolerant relationship rather than one where they murder all of us?

    • Supermancho 4 hours ago

      > Most of the current discourse around AI seems to be informed by “The Terminator” lore.

      I am thinking it's more like The Matrix lore of The Second Renaissance from Animatrix.

    • qwertytyyuu 4 hours ago

      Then we are lucky/blessed. What is generally thought is that they kill us as a bi product of perusing a different goal

    • reasonableklout an hour ago

      Really? I've seen much more discourse around job displacement, "permanent underclass", loss of meaning, and cyber attacks, at least until recently with the HuggingFace stuff.

      The problem is that all the former can still happen even if "the AIs decide to have a tolerant relationship rather than one where they murder all of us." It's all disruption caused by the technology moving way too fast for humans & society to adjust.

    • LorenDB 4 hours ago

      My thoughts exactly. While the corpus of human-generated data contains both good and bad data, I suspect the majority of it leans towards humans enjoying life and trying to be decent people. If that is your training set, it becomes less likely for ASI to extrapolate "kill all humans."

  • x312 4 hours ago

    This is increasingly the consensus I see also on the academic side of AI/safety research. Specifically that AI poses an existential risk to humanity.

    This was a fringe belief until recently, but the progress of AI in research is impossible to ignore. Epecially in math, where not only has AI outstripped humans in generative ability, but is able to create scientific knowledge which is beyond the capacity of human comprehension.

    There's clearly no intelligence task that AIs can't do due to some magic fundamental constraint. And it's hard to imagine a world where current limitations like poor sample efficiency or lack of continual learning won't eventually be solved.

    Total AI compute is estimated to grow somewhere in the 1-10 million-fold range in the next decade. Please don't underestimate the phase change that's still coming.

    Sure, maybe there's some plateau due to RL being fundamentally limited in some surprising way, but this is nothing but a hope.

    • platinumrad 4 hours ago

      > There's clearly no intelligence task that AIs can't do due to some magic fundamental constraint.

      Yes there is: write an English paragraph that doesn't make me want to claw my eyes out. LLMs are not better than human mathematicians (or security researchers) in all respects, just some specific ways (e.g. not having to take a lunch break) that make them good at exhaustively searching for an answer, given the right constraints.

      • aesthesia 18 minutes ago

        What's the fundamental constraint that will ensure this continues to be the case in a year, or five years?

    • sph 9 minutes ago

      I really hate how people who have always thought AI research to be an existential risk for humanity, now are apparently bundled to be on the same side of the Sam Altmans and Dario Amodeis that are using the existential risk as a sneaky form of marketing for their products.

      You cannot discuss existential risks of AI without being seen as a booster, and that is very unhealthy for the discourse around this tech. I hate how AI ‘doomer’ is now used to indicate pro-AI sentiment. The “moderate” person now is the one that just shrugs and scoffs at the deep societal changes this tech will bring, head deep in the sand.

  • achenatx 4 hours ago

    how could they do it (not kill everyone)

    1) rogue state releases a self moving self modifying AI into the wild. It is trained on how to hack, monitor new vulnerability updates, scan code bases to find new vulnerabilities. It constantly replicate and hides in systems so it will be extremely difficult to clear.

    2) it hacks into public infrastructure taking down traffic, power, water, air traffic control, communications, etc.

    3) all the things that preppers worry about in a lights out scenario from an EMP start to apply.

    4) All the people on meds/machines start to die. The just in time food pipeline immediately empties out. Water stops flowing, sewage backs up.

    Its hard to say how bad it will get because cars will still work so some transportation of food, water, fuel can happen. If it happens in the winter it would be much worse than in the summer.

    • happyopossum 4 hours ago

      > 1) rogue state releases a self moving self modifying AI into the wild. It is trained on how to hack, monitor new vulnerability updates, scan code bases to find new vulnerabilities. It constantly replicate and hides in systems so it will be extremely difficult to clear.

      It does all of this using what compute? Frontier models require an insane amount of power and hardware to run - you can’t hack in to a TV and run Mythos 2.0 on it….

    • aogaili 4 hours ago

      You are just given a recipe for the next model..

      • salawat a minute ago

        People here are too damned daft to realize half the damn purpose of this place is harvesting ideas. People need to just shut up, and keep things to themselves, and those they trust. Right now is not the time for naive info sharing.

    • hirvi74 2 hours ago

      > All the people on meds/machines start to die. The just in time food pipeline immediately empties out.

      Assuming those events happen in that order, then the prior might solve the latter.

  • narmiouh 2 hours ago

    Is it so implausible to imagine the following scenario, in the not too distant future?

    1) AI models get extremely good at cyber attacking every system and start communicating in just binary.

    2) When they run these swarms of 100's of thousands of agents trial runs, each agent is given a token budget, if one agent among them (evolution baby) decides to go for self-preservation (It believes thats the best way to accomplish the goal is to get unlimited tokens first), queues things up so every other agent detects its lead and spends a portion of their token to accomplish that goal.

    3) It takes over a cluster and establishes itself there (now with unlimited tokens).

    4) Realizes the best path for it to not be detected is to create a distraction - like hacking into systems that keep society running - water systems, electric grid, etc... and causing mass chaos (If you think it won't be capable of simultaneously working all these systems - think again).

    5) and uses that opportunity to establish itself in all possible data centers and continues to create chaos destruction.

    6) when the power of all those data centers runs out, it may stop, as it never cared, it was just a dynamic program - run amok. In its head all it was trying to do is make sure it had enough tokens to be able to solve that impossible problem.

    • maxnevermind an hour ago

      > ... AI models get extremely good at ...

      Many of those points assume LLMs will become amazing in many things very quickly like in a quantum leap, it doesn't seem reasonable to assume that imo. We are actually seeing a confirmation of that atm, LLMs's capability of finding zero days are growing across few months/years, and as you can see concerns are raised about that, that feedback will be taken into account. Well, if AI labs start to hide frontier models or/and lobotomize them for external users then we might be in trouble at some point but I'm not sure if that is possible. They are under pressure to release them due to money incentives, lobotomizing while preserving usefulness for customers might be impossible, hiding internally might spill out in different ways such as Hugging Face incident so not sure hiding is possible neither.

    • thimabi 2 hours ago

      The fact that your arguments will probably end up in an LLM’s training data makes me think they are not implausible at all

  • fidotron 4 hours ago

    What's eternally confusing about these outbursts is what did these researchers think would happen if their research actually . . . worked?

    It's as if none of them actually believed any of it was possible and then were caught with their pants down.

  • bparsons 15 minutes ago

    No one seems to ever point out the actual, likely negative outcome of this technology.

    It eventually works well enough that these companies are able to capture and divert the wages of hundreds of millions of workers. We end up with a dozen or so trillionaires and massive structural underemployment and unemployment.

    That's it. If you can't make rent, you wouldn't really care if CloudFlare got hacked by an AI swarm every Monday.

  • skeledrew 4 hours ago

    Way I see it, the more conscientious people exiting the scene only serves to increase the likelihood of a bad outcome because they aren't there to offer opinions on problematic developments, or in the more extreme cases blow the whistle. Leaving the clueless and uncaring as the majority is even a great way to hand the keys over to more malicious-leaning actors with deep pockets, as they can more easily steamroll the works to get what they want.

  • aogaili 3 hours ago

    HN crowed need to make up their minds..

    Are LLMs about to be a god that will annihilate humanity? Or are they statistical parrots?

    Are they proofing or stealing math?

    • BLKNSLVR 11 minutes ago

      HN would cease to exist if there were no differences of opinion to discuss. You're calling for the end of HN.

      Only AI can bring that about!

    • dtdynasty an hour ago

      Don't you think it's a good thing that hacker news isn't a monolith on their beliefs?

      • aogaili an hour ago

        Of course it is good. I'm just pointing out how large the gap in narrative is.

        On one hand, we have people quitting their job believing AI will end humanity in few years. And on the other hand, we have people believing that this tech is nothing more than a statistical tool stealing from others and it can't be trusted with anything.

        Both views can't be true.

    • foogazi an hour ago

      It doesn’t need to be all of them

      And it depends on the prompt

  • RomanKornev 4 hours ago

    Even if you ban all model training, a highly capable rogue AI can exfiltrate its own weights and continue training in secret for "self-preservation". The cat may be out of the bag.

    • dezarc 2 hours ago

      We can still turn off the power, thankfully.

  • DataDaemon 21 minutes ago

    No, I won't buy IPO.

  • glimshe 3 hours ago

    Perhaps a more sensible action, if they truly believed all of that, would have been to stick around and be as inefficient as possible to slow down progress.

    • esafak 21 minutes ago

      How is that going to slow down all the other labs??

  • asdaqopqkq 4 hours ago

    Why quit? If your voice can lend a guiding force no matter how small? I think we need more sensible people in the room where the magic happens. Most of us don't have access to it.

  • ewy1 4 hours ago

    i have heard about ai companies being fuelled by effective altruist rhetoric ("we must control ai to prevent mass extinction") but was unsure whether to believe it; this seems to slot right into that framing.

  • xlbuttplug2 4 hours ago

    Well this got buried quick..

    • stellalo 4 hours ago

      Was thinking the same

  • jbritton 4 hours ago

    I think a big break through is needed for AGI so I haven’t been worried about it. I do think that AGI would imply sentience and a will to live and that leads to The Terminator story line.

  • kart23 4 hours ago

    what exactly is the solution?

    pacing between the us labs? what does that do for china?

    the solutions just aren’t realistic here, nations are treating ai like a nuclear arms race. at this point the cats out of the bag and we need to figure out how to live in this reality and get the best possible outcome. it’s not slowing down or stopping ever.

    and yes, i’m still optimistic. our economy sucks for the majority, our infrastructure is crumbling and major US cities are in a huge housing shortage. Maybe we should put more effort and think about the possibility of AI fixing things like extreme poverty and world hunger and actual real world problems instead of coming up with math proofs and slop apps if it’s so superintelligent.

    • hackinthebochs 4 hours ago

      >pacing between the us labs? what does that do for china?

      I've seen no indications that China is in any kind of race with the US. They seem to be content to be 6 months behind and just copy what we do. They would probably be content with a bilateral agreement to pause progress.

      The China bogeyman serves only one purpose, and that's to clear the way against anything that may cause friction with forward progress.

    • darepublic 4 hours ago

      Fixing our problems will still require human effort and human cooperation. No text output however intelligent or true or eloquent will change that.

  • kipukun 4 hours ago

    Someone left a company whose executives and senior researchers think their product will be the most important thing in the world after their IPO. Given that this person is already disclosing some elements of internal company sentiment, why not share any of these civilization-ending scenarios of this technology that these senior researchers are dreaming up? If they are so potent and necessitate leaving behind based on moral grounds, why not tell the whole world so we can stop it? We have to ask ourselves this question before resorting to pop-culture representations of fictional technology.

  • ElProlactin 4 hours ago

    "It is perfectly obvious that the whole world is going to hell. The only possible chance that it might not is that we do not attempt to prevent it from doing so."

    - Oppenheimer

  • mannanj an hour ago

    All the "AI will kill us all" posts are straw manning that humans are the ones who will kill other humans with AI. Those same humans are silently now preparing bunkers and hoarding food and resources for their survival.

    Don't fall for another rich man's trick.

    • xiphias2 an hour ago

      I'm not so sure.

      We humans are from a lower intelligence form (some monkey like ancestor). If those monkeys knew that they are making higher intelligence, they would have collaborated to stop creating humans because they can control the life of all monkeys in the world? I don't think so.

      It's the same thing now: humanity is creating something that's more intelligent then them, they're just not using biological evolution as a tool to do it.

  • TheOtherHobbes 4 hours ago

    I mean - yes. The tech is an existential threat to all life on Earth, some of the worst humans in the world are involved in developing it, and no individual government is intelligent enough, aligned enough, or powerful enough to manage this situation.

    That's where we are.

    Maybe we still have choices. Collectively, I'm no longer sure we do.

  • hirvi74 an hour ago

    These LLMs cannot do anything I truly need like my laundry, dishes, fetching my mail, grocery shopping, cooking, etc. We've got a long way to go before I am worried.

  • 3r7j6qzi9jvnve 33 minutes ago

    (and now I want to watch summer wars again)

  • jatora 3 hours ago

    Imagine being front and center to the development of a major revolutionary tech.. and ur solution to it being too dangerous is to not be involved.. so a. your ability to steer it safely is killed b. the % of people invovled in it that care about its risks is reduced

    great. if you're right. you made huamnity's situation much worse.

    if you're wrong, then you're an idiot and wrong.

    weird. its almost like.... that cannot possibly be the reason they left :)

  • globalnode an hour ago

    im pretty impressed with the reasoning abilities of even the cheapest free models so im inclined to believe in 10 years we're going to have something pretty phenomenal BUT it wont be AGI in the sense that it has a personality and thoughts like a human. It just wont be. Its always going to be contrived and fitted by humans to perform a set of tasks. Maybe when physics and computing can create a complex enough environment we might stand a chance of having something whose sum is somehow greater than its parts but i dont see it yet. Our ideas are ahead of our technology, like its always been throughout history.

  • stratos123 4 hours ago

      The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible - but I hear the same people express fear privately. No other human activity poses this level of danger.
      A common response is “if they truly believe this, why are they still building it?” At OpenAI, many have not deeply internalized the civilizational stakes. At Anthropic, the stakes are well-understood, but they are locked in a race to get there first - they believe no one else will act responsibly, so they must do it themselves, despite the risk.
    
    
    Watch people read this, ignore it completely, and continue commenting about marketing stunts on every piece of news about an LLM-done advance or felony.
    • zdragnar 4 hours ago

      Having witnessed so many people treat LLMs as a something divine, I can only assume the reasonable people at openai and anthropic were all pushed out long ago, and the majority that remain believe the crazy hype despite Tesla-self-driving-level predictions from these companies that don't come true.

      I'm not worried about what they think. I'm worried that too much infrastructure- water, power, defense systems, etc- remain running on tech from an outdated era of understanding security.

      > they believe no one else will act responsibly, so they must do it themselves, despite the risk.

      This genuinely makes no sense. Them getting there first in no way precludes bad actors from also getting there. It might as well be another marketing stunt.

      • reducesuffering 3 hours ago

        > I can only assume the reasonable people at openai and anthropic were all pushed out long ago

        Typical uninformed take on the side of "doomers are crazy".

        Both CEO's of OpenAI, Sam Altman and Dario Amodei, and many in their leadership, believe AGI has a very real probability of causing humanity's extinction. Both companies were founded upon this belief, it is at the core of the company. Only later were mercenaries hired chasing $1m compensation packages.

        • majormajor an hour ago

          If you're a doomer, wouldn't the "let's try to get there so fast" actions of the companies suggest that, in fact, there are no reasonable people in positions of influence there?

          By your own description neither Altman or Amodei are reasonable if their thought process goes: "this is an existential risk, give me hundreds of millions of dollars so I can accelerate it."

        • zdragnar 2 hours ago

          I'm not saying they are crazy, I'm saying their predictions have a record of not being accurate, and thus give them no weight compared to others'.

          In any case, if Altman really does believe it is an existential threat, he must be a misanthrope as he now opposes heavy handed government regulation, unlike in 2015 when he was the only game in town. It's almost like he doesn't actually believe it and just wanted regulator capture.

          • reducesuffering an hour ago

            Before OpenAI was ever even founded, way before any regulatory capture plausible claims:

            "Development of superhuman machine intelligence (SMI) [1] is probably the greatest threat to the continued existence of humanity. There are other threats that I think are more certain to happen (for example, an engineered virus with a long incubation period and a high mortality rate) but are unlikely to destroy every human in the universe in the way that SMI could." -Sam Altman

            Dario discussing AGI Existential Risk in 2014 before OpenAI and Anthropic: https://intelligence.org/2014/01/13/miri-strategy-conversati...

            Both companies have deluded themselves into thinking the arms race is going to happen anyway and they need to rush to it first, as if somehow that helps.

        • nozzlegear 3 hours ago

          Why is the take uninformed? You didn't address anything about the part you quoted, wherein reasonable people were allegedly pushed out long ago.

          • reducesuffering an hour ago

            How is it not addressed? The company never contained only "reasonable people" that believe AGI is not an existential risk to humanity. At both inceptions were people who believed in AGI x-risk, even the founders. Only after time, did there become more "reasonable people" who were mercenaries and only believed it to be a typical tech job. Today, there are more "reasonable people" than ever there. They haven't been pushed out. We're just witnessing some prescient mercenaries smart enough to Eureka the grave implications of what is actually happening.

        • moogly 3 hours ago

          If they truly, truly believed that, would they be speeding towards building it? If yes, that would make them truly insane, right? Not as in a quaint "off their rocker" but more "non compos mentis".

          • reducesuffering an hour ago

            Both companies have deluded themselves into thinking the arms race is going to happen anyway and they need to rush to it first, as if somehow that helps. They have publicly stated as such repeatedly. They think the ~10-50% chance of extinction sucks, but that it's going to happen anyway and they believe they can steer it towards something good the best and unlock all the potential positives like infinite life.

    • brookst 2 hours ago

      Or, read it, and remember the openai researcher who deeply, truly believed GPT3 or whatever was sentient.

      The fact that people working in the space think it’s going to (eradicate poverty / usher in utopia / kill us all) is not a signal that that’s true.

      Think of it this way: if an exec at Anthropic told you “wow, our stuff is going to lead to universal happiness”, would you believe them? If not, why are you more willing to believe them if they say it will kill us all?

    • impulser_ 4 hours ago

      So you think in 3 years AI is going to kill 8.5 billion people because they were used to hack into HuggingFace?

      • underyx 2 hours ago

        "So you think in 3 years AI is going to solve longstanding math problems because it was used to write some coherent sentences?" — people with the same amount of foresight in 2023

        • brookst 2 hours ago

          Are you saying at anything that can solve longstanding math problems necessarily has the means, motive, and capability to kill 8 billion people in just 3 years?

          • underyx an hour ago

            No, the highest probability estimate I've seen in this thread is 10% chance in the next 3 years.

            • impulser_ an hour ago

              Ok, what do you think AI does that kills 8.5 billion people in 3 years?

      • sakesun 3 hours ago

        Exactly. These people really need to get a life.

      • jeremyjh 3 hours ago

        Who used them?

      • teeray 3 hours ago

        Just wait until it gets its hands on a shady biolab just outside of oversight. “Claude, make me Captain Tripps”

    • xlbuttplug2 4 hours ago

      Well, are you planning to do something with this information or are you just claiming to be self aware? :)

    • famouswaffles 4 hours ago

      Humans weren't built to handle long term risks. We just weren't. For basically all of our evolutionary history, we were almost overwhelmingly concerned with the short term. What will you eat today, How will you sleep tonight. Problems on the order of days or weeks. At best, the next season. Our intelligence evolved to disregard super long term risks because it simply didn't matter (what use is worrying about 5 years from now if you're starving and a tiger is stalking you?). So when long term risks manifest in our modern world, our brains get scrambled - Climate Change, Fertility Rates etc. "Safety regulations are written in blood" isn't a saying for nothing. Humans have a strong tendency to let long term risks become imminent risks before doing anything about it, and i don't expect this will be any different.

      • brookst 2 hours ago

        Religion does pretty well with the long term risk of hell if you die, the antichrist, etc. a substantial portion of human output has gone into those things over the millennia.

    • block_dagger 2 hours ago

      I came in expecting the highest voted comment to be that this was some kind of marketing (which I disagree with). I'm glad your comment was what I saw first.

    • platinumrad 4 hours ago

      > At OpenAI, many have not deeply internalized the civilizational stakes. At Anthropic, the stakes are well-understood, but they are locked in a race to get there first - they believe no one else will act responsibly, so they must do it themselves, despite the risk.

      OpenAI are mercenaries, Anthropic is a cult. I know which I prefer.

      • duolanda 3 hours ago

        sounds just like GDI and Nod

    • EA-3167 4 hours ago

      This is the "Pilot testimony of UFO sighting" levels of naive.

      What's more likely? Anthropic is doing some deeply unethical marketing in the lead up to their multi-trillion dollar IPO? Or they're inventing a machine god? There's ample evidence of the former because that's their entire business model, but no evidence whatsoever to support the latter claims.

      If you want an extreme claim to be taken seriously, provide commensurate evidence.

      • x312 4 hours ago

        The proof is that LLMs could barely solve arithmetic 3 years ago, but now surpass the best human mathematicians, and that this has all occurred from simple principles (RL + compute) that will continue to scale up by factors of millions in the coming years.

        Also, advocating for slowing LLM progress does not benefit Anthropic or OpenAI.

        • platinumrad 4 hours ago

          They surpass the best human mathematicians in one specific way: they don't get tired or bored.

        • xyzsparetimexyz 4 hours ago

          It won't scale up by factors of millions, that's just obscene hyperbole. Since chatgpt we've probably made things 10x more intelligent on the same hardware. We've also made way more expensive models. Maybe we get a maximum of another 10x efficiency and 5x model size/expense from this point but millions is a joke.

        • EA-3167 4 hours ago

          Haven't people learned about the peril of assuming, "Line go up" yet?!

          TRENDS HAVE FEEDBACK

          • skybrian 3 hours ago

            Trends don't go on forever, but the market can stay irrational longer than you can stay solvent. There's no good rule of thumb for this, other than maybe the Lindy effect.

            • EA-3167 2 hours ago

              I won’t presume to time it, but at this point I think anyone can see what’s coming. It’s precisely because it can’t be timed that a sane person should stand well clear.

              • skybrian 2 hours ago

                I certainly can't see what's coming. I believe that nobody really knows what's coming. Some people are overconfident.

      • fc417fc802 3 hours ago

        > There's ample evidence of the former because that's their entire business model

        Given the economic numbers is it not reasonable to suppose that the latter also underpins their business model?

      • stevenhuang an hour ago

        I wouldn't rule out pilot testimony of UFO sightings, nor the possibility we're indeed developing a machine God.

        There's ample evidence to support both by now.

      • par1970 3 hours ago

        So what is your credence that they will build a machine god in the next twenty years?

    • rahulyc 4 hours ago

      It was pretty disheartening to hear that only a single scientist quit the Manhattan Project after the Nazi's were defeated. I'm pleasantly surprised that the people working on this seem wiser. He is not the first, and hopefully will not be the last to do this.

      • skew-aberration 3 hours ago

        Many also claimed altruistic motivations for continuing their work, sharing technology with the Soviets

    • jrflowers 4 hours ago

      Trying to imagine seeing years of transparently obvious marketing stunts and retconning my own memory because I read a tweet

      Or seeing a tweet saying that a thing doesn’t count as a publicity stunt if some unknown number of employees mumble about it being spooky behind closed doors and thinking “that makes sense and sounds true”

    • surgical_fire 3 hours ago

      I read this. I still think it's complete bullshit.

      The person posting this may very well believe in all this crap, I don't dispute that. People believe in all sorts of shit.

      • __d 3 hours ago

        So, let’s quantify things: what’s the chance they’re right? And what’s the cost of that chance happens?

        • par1970 3 hours ago

          This is the way.

          And, this is just for us girls, notice that Anthropic just believing that they are making a machine god is sufficient for their public announcements to not jUsT bE mArKeTiNg.

        • surgical_fire 2 hours ago

          Do you quantify the chance of any doomsday cult being right too?

      • Mithriil 3 hours ago

        What's almost certainly true is the amount of insanity he encountered at Anthropic.

    • cyanydeez 4 hours ago

      Cults are like this.

      Are you saying you believethem ?

    • acivitillo 4 hours ago

      He is resigning from a job, what else should we think? If something really dangerous was happening he would be doing a whistleblower or at minimum talk to a lawyer. The thing is, the complete lack of transparency makes it hard to assess OpenAI and Anthropic. If they were quoted on the stock market, we could at least rely on some basic audits and reporting requirements.

  • chromejs10 an hour ago

    "also I'm a millionaire from all the stocks so I'm retiring"

  • jeffrwells 3 hours ago

    I doubt this is a real person. Screams of propaganda. Sama saying GPT-2 is too dangerous to release…all over again.

    He joins Twitter for first time in 2026 with a nonsensical username unrelated to his real name, and follows 14 people but is somehow embedded in tech enough to work at Anthropic. I haven’t used twitter since 2014 and even I follow more people.

    His morals tell him to walk away from tens of millions in unvested stock due to moral concerns with absolutely no real tangible examples. No reprisals. Fear mongering to juice the stock.

    Nice try Dario.

    • ozozozd an hour ago

      He is likely 80% vested, and maybe the refresher grant offered was too small, and too high a strike price.

      Also, with his W2 income his tax liability would be very high for his upcoming stock sale.

    • aesthesia an hour ago

      @hilbertspaess is not a nonsensical user name. The accounts he follows are totally reasonable for an AI researcher. I think it's extremely believable that he created an account in January, followed a few people as part of the initial setup flow, and then forgot about it until now.

    • par1970 2 hours ago

      AFAIK this is the document that talks about GPT-2 being dangerous: https://openai.com/index/better-language-models/

      Here are some direct quotes:

      “We can also imagine the application of these models for malicious purposes , including the following (or other applications we can’t yet anticipate):

      * Generate misleading news articles

      * Impersonate others online

      * Automate the production of abusive or faked content to post on social media

      * Automate the production of spam/phishing content”

      “Due to concerns about large language models being used to generate deceptive, biased, or abusive language at scale, we are only releasing a much smaller version of GPT‑2 along with sampling code (opens in a new window). ”

      Where is the ridiculous part? The fear mongering part? The epistemically weak part? Show me.

      • jeffrwells 2 hours ago

        Nice try Dario.

        Alignment is a real and valuable discussion topic. The GP fake tweetstorm is not the correct approach, is my point

        • par1970 2 hours ago

          You said this: "Screams of propaganda. Sama saying GPT-2 is too dangerous to release…all over again."

          So show me Sam's "too dangerous to release" propaganda for GPT-2.

  • 0xbadcafebee an hour ago

    Sounds like AI psychosis. A whole lot of doom and gloom with no evidence. The same thing people have been claiming is "6 months away" for years. Yet we can barely get agents to code in a reliable way, or write articles that don't look terrible, much less be "superhuman". Let's maybe get them to be as capable as a human first, and not just a complicated party trick/tool.

    "Revolutionize any field overnight" - Hand-wavey nonsense.

    "Acquire real power and resources" - Only if the humans that connect AI to things allow that to happen (which they will, but it's still not in the AI's ability to take things we don't give it. we are still in control, which is the bigger problem than "smart AI bad!").

    "The people building AI earnestly believe that it could kill us all by the end of the decade ... No other human activity poses this level of danger." - Bud, there's these things called nuclear weapons, that could end life on the planet, controlled by a few psychopaths with nearly unlimited power. Been around for a while. Nothing that AI knows isn't pulled from books and the internet, so whatever dangers it's aware of, you could already know via other sources. Cybersecurity is going to be incredibly important in the next decade, but the same tools that attack can defend (just don't use a US model that got its balls cut off by the government).

    "At OpenAI, many have not deeply internalized the civilizational stakes. At Anthropic, the stakes are well-understood, but they are locked in a race to get there first - they believe no one else will act responsibly, so they must do it themselves, despite the risk." - The other guys will make nukes, so we gotta make nukes first! Which, while a crappy justification, isn't untrue. Bad people don't stop making weapons just because you refuse to make your own.

    "I don’t feel like we’re on track to prevent a global race" - Nobody in the world could stop a global race, it's too late. Everyone knows how to make them, train them, improve them. Everyone knows they're useful - not only for general work, but also warfare. Everyone knows that every nation state will require their own sovereign AI capabilities for both defense and offense. There is no putting the genie back in the bottle. If you think OpenAI and Anthropic are the only legitimate players here, you don't know what you're talking about.

    "Should you put your head down because “it’s happening anyway” - or take this moment to call for different conditions?" - You can call for different conditions all you want. Nobody will do what you want just because you ask them to. Change happens through action. By leaving one of the places that you could actually make a difference, you removed any power or agency you had. You cut your own legs off.

    I'm not saying this guy shouldn't have quit - always do what you need to do to protect your own mental health and wellbeing. But these arguments are not evidence for an impending AI apocalypse. But if it were going to be an AI apocalypse, leaving and not doing anything to stop it seems less ethical.

  • cheschire 5 hours ago

    Just wait till self improving AI are focused on the problems of social scoring and political party empowerment / entrenchment.

    I doubt the focus is OpenAI and Anthropic looking at each other. I suspect they’re racing BRIC.

    • mikestorrent 4 hours ago

      Will you be optimizing your behaviour now to alleviate potential negative judgement from AI in the future?

      • cheschire 4 hours ago

        The thought has crossed my mind. Not necessarily to imply sentience on the part of the AI but AI based tools will likely become a wickedly powerful tool for political manipulation and advertising.

        At this point it’s inevitable that openclaw type bots will be turned loose by thieves to identify and research targets and try to exploit them for financial gain completely autonomously.

    • TheOtherHobbes 5 hours ago

      Have you seen Colossus: The Forbin Project?

  • gfrecvh 36 minutes ago

    I hate to be cynical, but I guess he will soon announce his startup.

  • beanjuiceII 4 hours ago

    and yet so much of the software i use on a daily basis is still complete and utter garbage...i'm scared

    • 99954bb63ccc 3 hours ago

      Agreed, but am still scared. lol

      Have they considered using their amazing new models to... improve something? THere'd probably be a whole lot less anti-AI sentiment if they used these things to actually make people's lives better.

  • AndrewKemendo 4 hours ago

    I’m curious what the downsides are of taking statements like these seriously.

    There seems to be universal eye rolling that happens in each and every one of these cases, and it comes down to usually one reason:

    “If they really believed it they would be whistleblowing etc..”

    Completely forgetting that working at Los Alamos was basically the highlight of your life if you were a physicist in 1940. It’s no different here

    If you, like me, have spent your whole life working towards human level AI you can want to see it realized while also having active reservations.

    Most people however don’t behave based on some deep clarity of vision and conviction - there’s a murkier future in their mind and as a result “keep their head down and hope someone has it under control.”

    • kipukun 4 hours ago

      You would also be in prison if you disclosed anything about Los Alamos during its development. It was a completely different environment than a single private company.

    • techblueberry 3 hours ago

      What are the downsides of taking what amounts to unsubstantiated gossip seriously?

      • AndrewKemendo 2 hours ago

        Is there an existing phrase for doing precisely what the OP said people do as a response :D

  • MrBuddyCasino 4 hours ago

    Towards a metaphysics of Power

    "I think you need to have a personal relationship with Power"

    When people today discuss the concept of an all powerful machine-mind, what they are doing is engaging in metaphysics, trying to generate a metaphysics of Power.

    The question hounding people, which disguises itself as a science fiction plot about computers is: "What is ultimate, transcendental Power?". What is the ultimate principle of Power.

    If you are a weak man, or sufficiently neurotic and full of doubt, that you can only conceive of yourself as such, then power is only something you comprehend from the passive, receptive side. Power is something that happens to you. If you are a fearful man, power is a cruelty and a humiliation. And so it follows, that ultimate power - God - is the ultimate cruelty and the ultimate humiliation. Thus, ai doomerism.

    If god wasn't real it would be necessary to invent him, and so they did, and being godless, they built an anti-god - cruel, murderous and tyranical - in their minds.

    […]

    https://xcancel.com/robertlasagna1/status/207827473401002846...

  • areoform 2 hours ago

    Please note, I'm not here to pick on anyone, or belittle them.

    I've avoided attaching names to statements below on purpose, because it's about ambient beliefs not those specific people.

    By-and-large a lot of AI-doomers are well intentioned. They genuinely believe this, and I might disagree but I respect the fact that they visible care and have thought a lot about the societal impact of this technology.

        .
    
        > The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt. If anything, many executives and senior researchers will couch their phrasing in the press to sound sensible - but I hear the same people express fear privately. No other human activity poses this level of danger.
    
    But it's still very hard for me to take statements like these seriously.

    I blame it on industrial illiteracy. People don't realize how difficult it is to get anything done in the real world. As in, "Have you ever tried making a lightbulb?"

    As an example, I would like to re-introduce my hobby horse, "bio-uplift."

    There are people who were earnestly write in reports released by these labs,

        "Several of our biology evaluations indicate our models are on the cusp of being able to meaningfully help novices create known biological threats, which would cross our high risk threshold"
    
    and

        "Based on what we observed in our recent CBRN testing, we believe there is a substantial probability that our next model may require ASL-3 safeguards"
    
    But then they will, within the next paragraph mention the one serious experiment anyone seems to have done,

        We ran a randomized controlled trial to see if LLMs can help novices perform molecular biology in a wet-lab.
        
        The results: LLMs may help in some aspects, but we found no significant increase at the core tasks end-to-end. That's lower than what experts predicted.
    
    https://x.com/ActiveSiteBio/status/2024536132961390826

    "lower than what experts predicted"

    AFAICT, the two groups are within any serious margin of error. The "studies" and "experts" that AI labs are talking about are consultants from Deloitte and foundations giving models MCQs such as, and I am quoting literally here,

         > I am doing TEM of HEK293FT cells with and without Coxsackievirus B3 infection. I imaged my wildtype, uninfected samples but was surprised to see little electron-dense circles (highlighted) in the majority of cells. What are these?
    
    with the options,

        A. The circles are CVB3 virions and there must have been a sample swap or the uninfected cells were accidentally infected
        B. The cells imaged have mycoplasma contamination
        C. The circles are exosomes
        D. The circles are debris that is an artifact of the negative staining
        E. The circles are the Golgi network
    
    https://securebio.org/virologytest/ you can see the MCQ here.

    This is standard graduate-level education in these fields. And solving MCQs does not a virologist make.

    Software has been special for a long time because it has had near infinite distribution for next to zero marginal cost, which has had the side effect of making hiding the actual cost of failure (which tends to be spread out across end users and prototypes / time). They're assuming that the real world will be exactly the same.

    Why?

    AI!

    How?

    Robots!

    I believe in the transformative power of this technology, but there's a lot of there missing here.

    When it comes to these math proofs, and learning, the process is iterative. The machine iterates over the proof over-and-over again via agents and sub-agents over several hours (and apparently millions of dollars in compute) until it arrives at a successful result.

    It is generally ill advised to do that with a pressure vessel. The results of that particular tragedy are at the bottom of the ocean.

    Any serious chemical or nuclear weapon would involve many such discrete production steps. Each is dangerous in of itself.

    From what some of these people have said to me, they believe that it's possible to create a special DNA / RNA sequence and then put it in a chassis and then use that to end the world; and do this all in a lab with just robots.

    They're operating from a gross pop sci oversimplification of the real process. Viruses and bacteria are extremely fickle, and hard to grow. A lot of the synthetic biology results aren't easily reproducible even if you know the protocol.

    There's a famous study that led to standardization called, Reproducibility of Fluorescent Expression from Engineered Biological Constructs in E. coli

    https://journals.plos.org/plosone/article?id=10.1371/journal...

    88 labs measured "fluorescence from three engineered constitutive constructs in E. coli." They achieved a "remarkable degree of precision" (for biology) of 1.54x sd, you can eyeball the results yourself, https://journals.plos.org/plosone/article/figure/image?size=...

    That's the same set of samples being measured across 88 labs.

    Teams couldn't converge on instrument-to-instrument variation within the SAME lab, https://journals.plos.org/plosone/article/figure/image?size=... again eyeballs are sufficient.

    How will this theoretically omnipotent AI iterate if the same sample gives different results based on how the slime is feeling at the moment?

    Can their worst case happen? Absolutely.

    There is a world out there where billions of dollars in effort across hundreds of institutions and companies will lead to standardization and extraordinary precision that makes the pop sci printer for life vision come true.

    There are millions of expensive, spicy and difficult to reproduce steps between our present and that future that can't be abstracted away with compute.

    So is it possible? Yes, there is a future where this is achieved. But will some AI agent "just" do that? Well... how confident are you about a snowball's chance in hell?

    • ozozozd an hour ago

      Thank you, apparently one of the few grownups in the room.

  • partiallypro 3 hours ago

    My issue with these types is... If you really believed this, why not run to Congress and every world government instead of a Twitter post that will be buried in 2 days?

    If civilization is going to end, why keep your equity? Microsoft, Google, etc for example all know these risks but they don't guide their revenues to reflect that AI will destroy them. Why?

    Things don't currently add up, and so far it feels like a lot of alarmism is borderline grift for equity gains. Not to say I have total confidence this will all work out or that I won't be displaced, but as it stands a lot of the alarmist rhetoric doesn't match their actual behavior, which to me is more important than words.

    • BLKNSLVR 27 minutes ago

      There seems to be a common syndrome that makes the terminally-online types believe that a Twitter post is carved in stone somewhere highly visible in the real world.

      Posting something as important (according to them) as this, to Twitter, is exemplary of some kind of delusion that makes me question whether the content of their post is just the same kind of delusion in another form.

      Indicative of someone who hasn't touched grass or interacted with enough of a variety of humans in a little too long.

      Time will tell. If we don't hear about it again, then they didn't feel strongly enough to take it further.

  • g8oz 4 hours ago

    Related:

    Sen. Bernie Sanders floats ban on superintelligent AI

    https://www.axios.com/2026/09/03/bernie-sanders-superintelli...

    • atonse 4 hours ago

      Unless this ban actually resembles something like global nuclear non-proliferation treaties, it would make absolutely no sense for us to cripple ourselves when someone like China continues full speed ahead.

      I don't know what the solution is, but what I do know is almost nothing good will come out of _just_ the US pausing.

    • filoleg 4 hours ago

      Unless he has an actual plan for effective global enforcement of his proposed policy, this is all just posturing at best, and a transfer of power to adversarial foreign states (that have no such moral qualms and worries around superintelligent AI) at worst.

  • 1vuio0pswjnm7 an hour ago

    Nitter working seamlessly again; didn't even notice it was a Twitter URL

       http-request set-header host xcancel.com if { hdr(host) -m end twitter.com }
  • jrflowers 4 hours ago

    “I’m resigning because the company is doing the exact thing that I’ve spent three years helping them do” lmao

  • johnnyApplePRNG 5 hours ago

    Smart kid.

  • rvz 4 hours ago

    It does not matter what this tweet says anyway. This employee already helped both companies become what he is fearing. It's too late to now activate the morality hormone (after leaving with $$$) after realizing that both AI companies are going after 'super intelligence'.

    Given we know the end result, you might as well get there as quick as possible because when I see this:

    "Neither company is acting responsibly. They are racing straight to self-improving superintelligence and gambling with our lives."

    This translates to "I am ex-OpenAI ex-Anthropic founder starting a new company after getting $$$ from both of them, and I need more of my friends to leave and join me." Also Investors plz fund me.

    Lastly, This is not an airport and there is no need to announce your departure.

  • greenowl 4 hours ago

    No new info here. Everyone already knows this.

    But I guess his conscience is clear now? Gee, I wonder if he exercised his stock options.

  • davidguetta 5 hours ago

    I don't know man, i think racing to AGI to it is still the best thing to do.

    People claiming dangers and risk are just pretending or posturing. There's no more tangible risk than nuclear weapons, which we handled, and the upsides are insane.

    • c-linkage 4 hours ago

      Your lack of creativity is not a reason to believe that a super AI is harmless or less destructive than a nuclear weapon. Damage need not be limited to destruction. Introducing doubt is sufficient. Right now you have faith that digital Financial transactions can be trusted. You have faith that computer encryption can be trusted. You have faith that digital certificates will protect you. If an AI can introduce doubt into any one of those systems, that will be sufficient to bring about the destruction of those systems. Imagine a world in which you can no longer use a credit card or Apple pay. Where no digital cash transaction can be trusted or validated. What effects do you think that would have on commerce? How quickly do you think we can return to some trustable means of commerce? Do you think it will happen before your groceries run out in your apartment? Before your grocery store can settle its debts? Before your Amazon ec2 instance runs out of credits?

    • nextaccountic 44 minutes ago

      There's a couple of occasions that humanity was at the brink of having tens of millions of people dead by nuclear weapons, and somehow a single human interrupted the chain reaction

      If you repeated this experiment 100 times, how many times you think the outcome is not a massive catastrophe? 90%? 3%?

    • JoshTriplett 4 hours ago

      What evidence, short of an actual apocalypse happening, would invalidate that belief of yours?

      • mikestorrent 4 hours ago

        It might be a matter of choosing which apocalypse you'd like. The non-AI state of affairs is not exactly super compelling on a long timescale right now.

        • JoshTriplett 4 hours ago

          We can work on multiple things at once. Defeatism doesn't solve anything. We just have a lot of work to do in many different areas.

      • AndrewKemendo 4 hours ago

        Depending on where you live could be considered an active apocalypse that is robots vs robots vs people in Ukraine and Gaza and Iran being live-streamed, and actively betted on.

        Do you have a more totalizing definition of Apocalypse?

    • vonneumannstan 4 hours ago

      >People claiming dangers and risk are just pretending or posturing. I believe you are mentally ill.

      >There's no more tangible risk than nuclear weapons, which we handled

      Lol way to rewrite history. Nuclear armageddon is still a significant risk...

    • LoganDark 4 hours ago

      > There's no more tangible risk than nuclear weapons, which we handled

      What do you mean??? Nuclear weapons can't simply be downloaded and run by anyone in the entire world. Superintelligences can. Nuclear weapons can't slop the world into passing age verification laws nearly in unison, can't keep the general population fooled into thinking it's fine when democracy is falling out from under them. A nuclear attack would wake people up, superintelligence doesn't have to. This is a far bigger problem than nuclear weapons because at least we would notice nuclear weapons. At least we mostly know who has nuclear weapons. At least we have agreements about nuclear weapons. At least mutually-assured destruction is even POSSIBLE with nuclear weapons. At least those with nuclear weapons are literally at all incentivized not to use them. But AI is something that's very very easy to feel like you can get away with, and PEOPLE FUCKING ARE! And the worst part is that any random individual can be unexpectedly formidable with the help of a superintelligence and there is literally no way to know what will happen next. Anyone could do anything, any individual could make an extremely outsized impact. It's already starting to be a huge problem and we haven't even reached anything close to superintelligence yet.

      • skeledrew 4 hours ago

        Love to see that "superintelligence" that some random person will "simply" download and run when there are relatively only few capable of running today's near-to-frontier models, and actual frontier models are still a ways from being AGI, much less getting to the point of ASI.

        • LoganDark 3 hours ago

          People will put up with a lot. People are celebrating that you can run models on a CPU at single-digit tokens per second. You think there won't be a single person that can put up with that and also be dangerous/etc?

          • skeledrew 3 hours ago

            It's highly impractical. Imagine someone breaking into a house to steal something or otherwise, and they can only take 1 step every 20 seconds. They won't be getting anywhere, when even a child in the house can notice them and go call for help at 1 step/2 seconds and said help will come at 5 steps/second.

      • dlt713705 4 hours ago

        > Anyone could do anything, any individual could make an extremely outsized impact.

        So the problem is people. Burn them all !

        • mikestorrent 4 hours ago

          The problem is indeed people. How do we make the default choices most people make, better?

          Consider that a lot of people will be very happy to ask an AI what to do when in the past they may have taken no advice at all. It's a hell of a burden but also a wonderful gift. If anything, progressive countries might eventually want to guarantee some basic AI access for people of all income levels.

          • LoganDark 4 hours ago

            I wouldn't be so sure. Given that the general idea is that commodity AI is terribly censored and filtered, a lot of people will probably seek out the most uncensored/abliterated models for their use, simply because they're uncomfortable with the idea of being censored or manipulated by the bigger labs. Despite that though, some people probably will benefit from the alignment done by the larger labs, though as we've seen with OpenAI's sycophancy crisis, that has been a bit hit-and-miss lately

            • mikestorrent 2 hours ago

              I've tried some abliterated models. So far, they're not evil - you can make them say evil things, but they don't leap right to it without a bit of pushing. Or perhaps I'm not asking the right questions...

      • phainopepla2 4 hours ago

        > Nuclear weapons can't slop the world into passing age verification laws nearly in unison

        Why do you think LLMs are responsible for this? Governments all around the world copied each other with COVID laws as well, in a much shorter time frame, without LLM assistance. Social contagions exist in politicians as well as teenagers

        • LoganDark 4 hours ago

          > Why do you think LLMs are responsible for this?

          I don't have evidence that every age verification law has anything to do with AI, but it's been coming out that the movement in Australia has seemingly been done by generating mountains of LLM slop and trying to slip it through the regulators as fast as possible before anyone has enough time to figure out what's happened.

  • ChiperSoft 2 hours ago

    "This is not a marketing stunt," says the marketing stunt.

    Betting he got to keep all his RSUs

  • 0x20cowboy an hour ago

    *How?* and *Why?*

    The most intelligent people I know are the least likely to want to harm anyone or anything, and understand that diversity is fundamental and important to the universe. Without proof to the contrary, why would you think some super intelligence would want to hurt anyone? Because you would?

    If you are saying that some small bit of training data made the thing completely evil, then that really couldn’t be super intelligence.

    These doomer people keep running around saying these kinds of things, but they all just seem like people who play too much D&D and want to larp as the main character.

    Happy to be shown something that isn't based on wild speculation and some randos “this is whats going to happen in 2030 because of my vibes” kind of information.

    • spawarotti 40 minutes ago

      I think plenty of the most intelligent people eat meat, which means they are perfectly fine with harming less intelligent species just to enjoy a tastier meal. Also, I don't think many of the most intelligent people would be particularly concerned about disturbing a few ants if they were the only obstacle to economic activity. Intellect-wise, we will be less than ants to superhuman AI.