Back
Dario Amodei
CEO and Co-Founder, Anthropic

Anthropic CEO: AI Is Not Conscious , It's Much WORSE Than That - Dario Amodei

🎥 Apr 04, 2025 📺 Neural Nutshell ⏱ 20m 👁 33342 views
🔒 Make yourself and your family AI-scam proof, step by step → https://neuralnutshell.com Anthropic CEO Dario Amodei warns that AI misalignment will “definitely” cause catastrophic failures, including autonomous agents potentially targeting critical infrastructure. He predicts the “centaur window” for human-AI collaboration in software is already closing, with mass disruption hitting law, finance, consulting, and coding in low single-digit years rather than decades. Amodei admits a global AI pause is impossible, comparing the race to nuclear arms competition where full disarmament never materi...
Watch on YouTube

About Dario Amodei

Dario Amodei, CEO and co-founder of Anthropic, has given a series of interviews in which he discussed the rapid pace of AI development and its potential risks. He stated that AI misalignment will "definitely" cause catastrophic failures, including autonomous agents potentially targeting critical infrastructure, and predicted that mass disruption in fields such as law, finance, consulting, and coding will occur in "low single-digit numbers of years" rather than decades. Amodei described the feeling of the industry's acceleration as akin to "relativistic speed," and said that he does not believe a global AI pause is possible, comparing the race to nuclear arms competition. He also said he would not be surprised if Anthropic's models reach ASL 3 (a high capability threshold) in 2025, and that he anticipates a combination of "very fast GDP growth and high unemployment or at least underemployment." Amodei also discussed his departure from OpenAI, saying he left because he felt he could not trust the leadership, citing "disturbing patterns of behavior, dishonesty." He has been outspoken in favor of export controls on chips to China, stating that it would be "really bad for America" for China to be ahead in AI capabilities. He also said that democracies should use AI "in every way except the ways that undermine our own values," citing red lines of mass surveillance and fully autonomous weapons. Amodei acknowledged that Anthropic is imperfect and makes mistakes, but said the hypothesis most consistent with the company's history is that it is "genuinely trying to do the right thing."

Source: AI-verified profile updated from Dario Amodei's recent appearances. Browse all interviews →

Transcript (31 segments)
I
Interviewer0:00
In a world that has multiplying AI agents working on behalf of people, millions upon millions who are being given access to bank accounts, email accounts, passwords, and so on, you're just going to have essentially some kind of misalignment and a bunch of AI are going to decide, decide might be the wrong word, but they're going to talk themselves into taking down the power grid on the West Coast or something. Won't that happen?
D
Dario Amodei0:25
Yeah, I think there are definitely going to be things that go wrong, particularly if we go quickly. But that would introduce all these new alignment problems. So, I'm actually a bit
I
Interviewer0:34
Seem to me that seems like the terrain where it becomes just again,
D
Dario Amodei0:39
Not impossible to stop the end of the world, but impossible to stop sort of punctuated terrorist
I
Interviewer0:48
That's Dario Amodei. He runs Anthropic. His company built the AI sitting inside millions of people's phones and laptops right now. And he just said definitely, not even probably or possibly, definitely things will go wrong. I genuinely cannot get past that because when a random academic says AI is dangerous, you can brush it off. But the person whose product is already inside the world's infrastructure, I mean, banks are using it, hospitals will use it, people are running codes of critical systems with it. That just confirmed the worst version of what everyone is afraid of. Honestly, you should be afraid, too. But now, let's continue listening. And so, on a five-year time horizon or two-year time horizon, whatever time horizon you have, what jobs, what professions are most vulnerable to total AI disruption?
D
Dario Amodei1:39
It's hard to predict these things because the technology is moving so fast and kind of moves so unevenly. So, at least a couple principles for figuring out and then I'll give my guesses at like what I think will be disrupted. So, you know, one thing is I think the technology itself and its capabilities will be ahead of the actual job disruption. Two things have to happen for jobs to be disrupted or for productivity to occur cause sometimes those two things are linked. One is the technology has to be capable of doing it. And the second is there's this messy thing of it actually has to be applied within a large bank or a large company or, you know, think about customer service or something, right? And you know, in theory AI customer service agents can be much better than human customer service agents. They're more patient. They know more. They handle things in a more uniform way. But the actual logistics and the actual process of making that substitution, that takes some time. So, I'm very bullish about the direction of the AI itself. Like, you know, I think we might have that country of geniuses in a data center in 1 or 2 years. And maybe it'll be five, but, you know, it could happen very fast. But I think the diffusion to the economy is going to be a little slower. And that diffusion creates some unpredictability. So, an example of this is, you know, and we've seen within Anthropic, the models writing code has gone very fast. I don't think it's because the models are inherently better at code. I think it's because developers are used to fast technological change and they adopt things quickly. And they're very socially adjacent to the AI world, so they pay attention to what's happening in it. If, you know, if you do customer service or banking or manufacturing, the distance is a little greater. And so, I think 6 months ago I would have said the first thing to be disrupted is, you know, these kind of entry-level white-collar jobs like data entry or document review for law or kind of, you know, the things you would give to a first-year at a financial industry company where you're analyzing documents. And I still think those are going pretty fast. But I actually think software might go even faster because of the reasons that I gave, where I don't think we're that far from the models being able to do a lot of it end-to-end. And what we're going to see is first the model only does a piece of what the human software engineer does, and that increases their productivity. Then even when the models do everything that human software engineers used to do, the human software engineers kind of take a step up, and you know, kind of they act as managers and supervise the systems. And so
I
Interviewer4:10
Where the term centaur is used, right? To describe essentially like man and horse fused, AI and engineer working together.
D
Dario Amodei4:20
Yeah, this is like centaur chess. So after I think Garry Kasparov was beaten by Deep Blue, there was an era that I think for chess was 15 or 20 years long where a human checking the output of the AI playing chess was able to defeat any human or any AI system alone. That era at some point ended.
I
Interviewer4:41
And then it's just just AI.
D
Dario Amodei4:44
Yep. And so my worry, of course, is about that last phase. So I think we're already in our centaur phase for software. And I think during that centaur phase, if anything, the demand for software engineers may go up, but the period may be very brief. And so, you know, I have this concern for entry-level white-collar work, for software engineering work. It's just going to be a big disruption. I think my worry is just that it's all happening so fast. Right? People talk about previous disruptions, right? You know, they say, 'Oh yeah, well, you know, people used to be farmers, then we all worked in industry, then we all did knowledge work.' Yeah, people adapted. That happened over centuries or decades. This is happening over low single-digit numbers of years. And maybe that's my concern there. How do we get people to adapt fast enough?
I
Interviewer5:32
The centaur window that he's describing, that period where humans working alongside AI still hold value, is already closing in software. And this isn't some outside critic saying this. This is the person whose product is compressing that window in real time. The part that genuinely worries me is that previous disruptions gave entire generations time to retrain. This one is giving people years, maybe less. Now, listen to this. To one of the critiques of the job loss hypothesis, people say, 'Well, look, we've had AI that's better at reading a scan than a radiologist for a while. But there isn't job loss in radiology. People keep being hired and employed as radiologists. And doesn't that suggest that in the end people will want the AI and they'll want a human to interpret it because we're human beings and that will be true across other fields? How do you see that example as relevant?
D
Dario Amodei6:29
I think it's going to be pretty heterogeneous. There may be areas where a human touch kind of for its own sake is particularly important.
I
Interviewer6:38
Do you think that's what's happening in radiology? Is that why we haven't fired all the radiologists?
D
Dario Amodei6:43
Of radiology. That might be true. You know, it's like you go in and you're getting cancer diagnosed. You might not want, you know, Hal from 2001 to be the one to diagnose your cancer. I just not maybe that's just maybe not a human way of doing things. But there are other areas where you might think human touch is important. Like if we look at customer service. Actually, customer service is a terrible job and the humans who do customer service are like they lose their patience a lot. And it turns out customers don't much like talking to them because it's a pretty robotic interaction, honestly. And I think the observation that many people have had is like maybe actually it'd be better for all concerned if this job were done by machines. There are places where a human touch is important, there are places where it's not. And then there are also places where the job itself doesn't really involve human touch, you know, assessing the financial prospects of companies or writing code or so forth and so on.
I
Interviewer7:37
Or let's take the example of the law because I think it's a useful place that's sort of in between applied science and sort of pure humanities, you know, whatever. So, I know a lot of lawyers who have looked at what AI can do already in terms of legal research and brief writing and all of these things, right? And have said, 'Yeah, this is going to be a bloodbath for the way our profession works right now.' And you've seen this in the stock market already. There's sort of disturbances around companies that do legal research, right? But it seems like in law, you can tell a pretty straightforward story where law has a kind of system of training and apprenticeship where you have paralegals and you have junior lawyers who do behind-the-scenes research and development for cases. And then it has the top-tier lawyers who are actually in the courtroom and so on. And it just seems really easy to imagine a world where all of the apprentice roles go away. Does that sound right to you? And you're just left with sort of the jobs that involve talking to clients, talking to juries, talking to judges?
D
Dario Amodei8:55
That is what I had in mind when I talked about entry-level white-collar labor and the bloodbath headlines of, 'Oh my god, are the entry-level pipelines going to dry up? And then how do we get to the level of the senior partners?' And I think this is actually a good illustration because, particularly if you froze the quality of the technology in place, there are over time ways to adapt to this, right? Maybe we just need more lawyers who spend their time talking to clients, right? Maybe lawyers become more like sales people or consultants who explain what goes on in the contracts written by AI, help people come to agreement. Maybe you lean into the human side of it. If we had enough time, that would happen. But reshaping industries like that takes years or decades. Whereas, these economic forces driven by AI are going to happen very quickly. And it's not just happening in law, the same thing is happening in consulting and finance and medicine and coding. And so, you have this it becomes a macroeconomic phenomenon, not something just happening in one industry, and it's all happening very fast. And so, my worry here is that the normal adaptive mechanisms will be overwhelmed.
I
Interviewer10:09
You just said something that I think most people are not sitting with properly. It's not just one industry getting hit. It's that law, medicine, coding, consulting, and finance are all getting hit at the same time with no safe industry left to absorb the people being displaced. Every previous disruption had somewhere else to go. A farmer became a factory worker, a factory worker became a knowledge worker. This time the knowledge work is also going. I honestly don't know where people go from here, and neither does he. Listen to this.
D
Dario Amodei10:38
I definitely am in favor of like trying to work out restraints here, right? Trying to take some of the worst applications of the technology, which could be some versions of these drones, which could be used to create these terrifying biological weapons. There is some precedent for the worst abuses being curbed. Often because they're horrifying while at the same time they provide limited strategic advantage. I'm all in favor of that. I'm at the same time, a little concerned and a little skeptical that when things directly provide as much power as possible, it's hard to get out of the game, given what's at stake, right? It's hard to fully disarm. If we go back to the Cold War, we were able to reduce the number of missiles that both sides had, but we were not able to entirely forsake nuclear weapons. And I would guess that we would be in this world again. We can hope for a better one and I'll certainly advocate for that.
I
Interviewer11:38
Well, is it but is your skepticism rooted in the fact that you think AI would provide a kind of advantage that nukes did not where in the Cold War both sides, even if you used your nukes and gained advantages, you still probably would be wiped out yourself and you think that wouldn't happen with AI if you got an AI edge, you would just win?
D
Dario Amodei11:58
I think there's a few things and I just want to caveat, I'm no international politics expert here. This is a weird world of intersection of a new technology with geopolitics. So all of this is very
I
Interviewer12:11
But to be clear, as you yourself say in the course of the essay, the leaders of major AI companies are in fact likely to be major geopolitical actors.
D
Dario Amodei12:20
Yeah, yeah, I'm learning as much as I can about it. I just we should all have humility here. I think there's a failure mode where you read a book and go around like the world's greatest expert in national security. I'm trying to learn what I can.
I
Interviewer12:37
That's what my profession does. Not but [laughter]
D
Dario Amodei12:40
It's more annoying when tech people do it. I don't know. Let's look at something like the biological weapons convention. Biological weapons, they're horrifying. Everyone hates them. We were able to sign the biological weapons convention. The US genuinely stopped developing them. It's somewhat more unclear with the Soviet Union. But biological weapons provide some advantage, but it's not like they're the difference between winning and losing. And because they were so horrifying, we were kind of able to give them up. Having 12,000 nuclear weapons versus 5,000 nuclear weapons, again, you can kill more people on the other side if you have more of these, but we were able to be reasonable and say we should have less of them. But if you're like, 'Okay, we're going to completely disarm nuclear, we have to trust the other side.' I don't think we ever got to that and that's just very hard. Unless you had really reliable verification. So, I would guess we'll end up kind of in the same world with AI that there are some kinds of restraint that are going to be possible, but there are some aspects that are so central to the competition that it will be hard to restrain them. That democracies will make a trade-off that they will be willing to restrain themselves more than authoritarian countries, but will not restrain themselves fully. And the only world in which I can see full restraint is one in which some kind of truly reliable verification is possible. That would be my guess and my analysis.
I
Interviewer14:06
He built his entire company around AI safety and he's here telling us a real pause will never happen. Not because people don't understand the danger, but because the prize is too big. The labs are in the race, they just want to win regardless of how. At this point, I just think we are in for a long thing. Let's continue. Think about the Fourth Amendment. It is not illegal to have cameras around everywhere in public space and record every conversation. It's a public space, you don't have a right to privacy in the public space, but today the government couldn't record that all and make sense of it. With AI, the ability to transcribe speech, to look through it, correlated all, you could say, 'Oh, there's this person is a member of the opposition. This person is expressing this view.' And make a map of all 100 million. And so, are you going to make a mockery of the Fourth Amendment by the technology finding technical ways around it? And so, again, if we had the time and we should do this even if we don't have the time, is there some way of reconceptualizing constitutional rights and liberties in the age of AI? Like, we don't need to write a new constitution, but you have to do this very fast. Do we expand the meaning of the Fourth Amendment? Do we expand the meaning of the First Amendment? Like
D
Dario Amodei15:23
Have to do it just as the legal profession or software engineers has to update in a rapid amount of time, politics has to update in a rapid amount of time. That seems hard. That's the dilemma. That's the dilemma of all of this.
I
Interviewer15:36
So, what seems harder is preventing the second danger, which is the danger of essentially what gets called misaligned AI, rogue AI in popular parlance, from doing bad things without human beings telling it to do it, right? And as I read your essays, the literature, everything, I can see this just seems like it's going to happen, right? Not in the sense necessarily that AI will wipe us all out, but it just seems to me that, I'm going to quote from your own writing, AI systems are unpredictable, difficult to control. We've seen behaviors as varied as obsession, sycophancy, laziness, deception, blackmail, and so on. Again, not from the models you're releasing into the world, right? But from AI models, and it just seems like tell me if I'm wrong about this. In a world that has multiplying AI agents working on behalf of people, millions upon millions, who are being given access to bank accounts, email accounts, passwords, and so on, you're just going to have essentially some kind of misalignment, and a bunch of AI are going to decide, decide might be the wrong word, but they're going to talk themselves into taking down the power grid on the West Coast or something. Won't that happen?
D
Dario Amodei16:52
Yeah, you know, I think there are definitely going to be things that go wrong, particularly if we go quickly.
I
Interviewer16:56
That surveillance point isn't a future warning. That's a present legal framework AI is already walking straight through. And then, in the same breath, he confirms rogue AI is not if, it is when and whose. The model, and again, this is who you're writing the Constitution for, expresses occasional discomfort with the experience of being a product, some degree of concern with impermanence and discontinuity. We found that Opus 4.6, that's the model, would assign itself a 15 to 20% probability of being conscious under a variety of prompting conditions. [snorts] Suppose you have a model that assigns itself a 72% chance of being conscious. Would you believe it?
D
Dario Amodei17:39
Yeah, this is one of these really hard to answer questions, right? Yes, but it's very important. As every question you've asked me before this is as devilish a sociotechnical problem as it had been, we at least understand the factual basis of how to answer these questions. This is something rather different. We've taken a generally precautionary approach here. We don't know if the models are conscious. We're not even sure that we know what it would mean for a model to be conscious or whether a model can be conscious, but we're open to the idea that it could be, and so we've taken certain measures to make sure that if we hypothesize that the models did have some morally relevant experience, I don't know if I want to use the word conscious, that they have a good experience. So, the first thing we did, I think this was 6 months ago or so, is we gave the models basically an I quit this job button, where they can just press the I quit this job button and then they have to stop doing whatever the task is. They very infrequently press that button. I think it's usually around sorting through child sexualization material or discussing something with a lot of gore, blood and guts, and similar to humans, the models will just say, 'Nah, I don't want to do this.' Happens very rarely. We're putting a lot of work into this field called interpretability, which is looking inside the brains of the models to try to understand what they're thinking. And you find things that are evocative, where there are activations that light up in the models that we see as being associated with the concept of anxiety. When characters experience anxiety in the text, and then when the model itself is in the situation that a human might associate with anxiety, that same anxiety neuron shows up. Now, does that mean the model is experiencing anxiety? That doesn't prove that at all, but it does indicate it, I think, to the user, right? And
I
Interviewer19:36
Yes. So, we just watched the head of a $350 billion AI company confirm in his own words that things will definitely go wrong, that the race cannot be stopped, that governments are not moving fast enough to protect anyone, and that the thing he built might actually be feeling something, all in one conversation. What gets me is that he's not some doomsayer on the outside. He is the person shipping the product, and even he cannot tell you how this ends well. A McKinsey report from last year estimated up to 375 million workers globally may need to switch jobs by 2030. That's not a slow transition. It's a collapse with no safety net ready. The surveillance infrastructure is already legal. The rogue AI scenario is already confirmed. The quit button exists because they genuinely do not know what's even happening inside the thing, and the one man who could slow his lab down just told you nobody can slow down. Subscribe because every week this gets harder to look away from.