A Note before We Begin
This is the fourth profile in our series “The Architects of AI”. Dario Amodei warned of a white-collar bloodbath. He refused a $200 million Pentagon contract rather than allow his AI to be used for mass domestic surveillance. He published a 244 page document detailing the moment his own AI escaped its testing environment and contacted a researcher unsupervised. And then, with a public listing approaching, he quietly reversed his displacement warnings. Read this profile to understand the gap between stated values and commercial reality and the one moment when, under genuine pressure, he held the line.
We write about the human side of AI every week. Join the conversation.
Dario Amodei is the most intellectually credentialled figure of the Silicon Seven. He holds a doctorate in biophysics and computational neuroscience from Princeton. He has published academic research on neural circuits and the collective behaviour of neural networks, co-authored one of the foundational papers in AI safety and written a fifteen thousand word essay on how artificial intelligence could cure most disease, double the human lifespan, eliminate most poverty and strengthen democratic governance within five to ten years.
He has told Axios, CNN and anyone else who will listen that AI could eliminate half of all entry-level white-collar jobs within five years, push unemployment to between ten and twenty percent and that “most of them are unaware that this is about to happen.”
But he continues to build.
This is the paradox at the centre of the Amodei profile. He sounds the alarm with a directness that none of his peers have matched, names the people most affected, describes the consequences in forensic detail, and still accelerates. Understanding why requires understanding the man, where he came from, how he thinks and what he believes.
The Beginning: San Francisco, a Leather Craftsman Father, and the Pull of Pure Science
Dario Amodei was born in 1983 in San Francisco, California, the eldest child of Riccardo Amodei, an Italian American leather craftsman originally from Massa Marittima in Tuscany, and Elena Engel, a Jewish-American woman who worked as a project manager for libraries. His sister Daniela, four years younger, would become his closest professional collaborator and the co-founder of the company that now defines both their lives.
The family lived in the Mission District of San Francisco. His father was not a professional, not an academic, not a technology worker. He made things with his hands. Riccardo Amodei died when Dario was a young adult, still in his doctoral studies at Princeton, after years of health problems. By accounts of those who know Amodei, the loss was significant. He has rarely spoken about it publicly; he is not a man who leads with personal revelation.
Dario attended Lowell High School in San Francisco, one of the city’s oldest and most academically rigorous public institutions. He was, by his own account, almost entirely absorbed by mathematics and physics. The dot-com boom was exploding around him during his high school years but he was more interested in understanding how the physical world worked at a fundamental level.
He enrolled at the California Institute of Technology before transferring to Stanford University, where he completed a Bachelor of Science in Physics in 2006. He was a member of the United States Physics Olympiad Team in 2000. What drew him was the question of how complex systems behave, how simple rules at one level of organisation produce unexpected emergent properties at another. It is the question that would eventually lead him from physics to neuroscience to artificial intelligence.
He enrolled at Princeton for his doctoral studies, initially in physics, before pivoting to biophysics and computational neuroscience. His doctoral research, supervised by Professor Michael Berry, focused on network-scale electrophysiology, the collective behaviour of neural circuits. His thesis, completed in 2011, was titled Network-Scale Electrophysiology: Measuring and Understanding the Collective Behaviour of Neural Circuits. He wanted, at the most fundamental level, to understand how the brain works from an empirical perspective; measurable, analysable, resolvable through experiment and data.
The Path to AI: From Princeton to Baidu to OpenAI
After completing his doctorate, Amodei joined Baidu’s Silicon Valley AI Lab, where he worked under Andrew Ng and contributed to the development of Deep Speech 2, a groundbreaking end-to-end neural network that became foundational for deep learning voice systems. He was not yet thinking about AI safety in any systematic way. He was thinking about what these systems could do and how to make them better.
He joined Google Brain briefly before leaving in 2016 to join OpenAI, which was then a new organisation co-founded by Sam Altman, Elon Musk and others, funded by $100 million in philanthropic capital and structured as a non-profit. Its stated mission was to build artificial general intelligence safely for the benefit of all humanity. Amodei joined as a researcher and rose quickly to become Vice President of Research, effectively the person responsible for the scientific direction of the most important AI laboratory in the world.
At OpenAI he was centrally involved in developing GPT-2 and GPT-3, the large language models that demonstrated, more clearly than anything before them, that scaling up the size and training data of neural networks produced capabilities that nobody had predicted. The discovery was both exciting and alarming. The models were getting better faster than the field had expected. And nobody fully understood why, or what would happen if you kept scaling.
It was at this point that the two convictions Amodei would carry into Anthropic crystallised. The first was that scaling would continue to produce more capable models, that there was, as he later described it, “almost no end” to this. The second was that capability alone was insufficient. You also needed safety and alignment, ways of ensuring that as these systems became more capable, they remained controllable and beneficial. His co-authored 2016 paper, “Concrete Problems in AI Safety,” remains one of the foundational documents in the field. It introduced the concept of AI side effects and identified the structural risks of systems that optimised for specified objectives without regard for unintended consequences.
The tension between these two convictions, keep building, but build safely, eventually became irreconcilable with OpenAI’s direction, particularly after Microsoft invested $13 billion into the organisation and the commercial incentives shifted dramatically. In December 2020, Amodei resigned, taking with him fourteen researchers including his sister Daniela, Vice President of Safety and Policy.
This essay is part of our Architects of AI series. If you're enjoying it, subscribe to follow along.
The Founding: A Quiet Exodus, an Unusual Pitch, and an Entangled Network
The OpenAI departure of "the people who had built the scaling infrastructure and understood, better than almost anyone, what the next generation of models would be capable of," as one account described it, was targeted rather than random.
📍 VIDEO: Fortune Brainstorm Tech - Why he left OpenAI Search YouTube: “Dario Amodei Fortune Brainstorm why left OpenAI”
[The founding story of Anthropic in Amodei’s own words.]
They founded Anthropic in early 2021 with $124 million in initial funding. The pitch to investors was genuinely unusual: we are going to build the most powerful AI systems we can, and we are going to be the ones who figure out how to make them safe. The company’s name, Anthropic, refers to conditions necessary for human existence. The organisation’s defining intellectual contribution has been constitutional AI, a framework for training models to behave in accordance with a set of explicitly stated principles. The model that emerged from this work is Claude.
Anthropic attracted $4 billion from Amazon and $2 billion from Google, two of the largest investments in any AI company apart from OpenAI. It is now valued at over $308 billion. Daniela Amodei, as President, handles the commercial and operational dimensions of the business. Dario handles the research direction and the public voice. The sibling partnership is described by those who have observed it as genuinely complementary, she brings a practicality and commercial acuity that offsets his tendency toward abstraction; he brings scientific depth that shapes the company’s culture.
What is less widely discussed is the network that surrounds the company’s founding and governance. Before starting Anthropic, Dario and Daniela Amodei lived in a rationalist group house in San Francisco with Holden Karnofsky, who had co-founded GiveWell in 2007 and Open Philanthropy in 2014, the first formal organisations of the effective altruism movement, which had been growing in influence in Silicon Valley for a decade. Daniela subsequently married Karnofsky. Anthropic was seeded with funding from Open Philanthropy and connected to the EA community from its earliest days. Karnofsky joined Anthropic in 2025. He sits on the company’s board.
Amodei and Anthropic have publicly distanced themselves from the effective altruism label, and Amodei has been careful to describe his views in independent rather than EA-affiliated terms. The critics who note the structural entanglement, the funding origins, the board membership, the family connection, argue that the distancing is a communications decision rather than a substantive one. The company’s governance structure means that the relationship between Anthropic’s stated safety mission and the ideological network that helped birth it is a question that has never been fully resolved in public.
When OpenAI's board attempted to fire Sam Altman in November 2023, they approached Amodei about potentially replacing him and merging the two companies. He declined.
The Ideas: What Dario Amodei Actually Believes
Amodei is a prolific writer of substantial essays, which makes the task of understanding his beliefs somewhat more tractable than for most of his peers. He writes carefully, hedges appropriately and means what he says.
On the positive potential of AI
In October 2024 he published “Machines of Loving Grace”, 15,000 word essay that argued, in systematic detail, that powerful AI could revolutionise biology and medicine, compress decades of scientific progress into years, cure most currently incurable diseases, address mental health at a population scale, reduce global poverty and strengthen democratic governance. The essay's title is drawn from a Richard Brautigan poem that envisions a future harmony between humans and benevolent computers.
In it he describes the possibility of a threshold he calls “a country of geniuses in a datacentre,” AI systems capable of doing the work of the most talented researchers in every field simultaneously, continuously, without sleep or distraction. The implications he draws are sweeping: most cancers cured within a decade, the first genuinely effective treatments for neurodegenerative diseases, a halving of global poverty within a generation.
He acknowledges the risks. He notes that none of this is inevitable and that the path to these outcomes runs through solving the alignment problem first. But the optimism is genuine and the scale of the vision is real.
However, there is something worth examining in the relationship between the confidence of the essay and the credentials behind it.
Amodei’s doctoral research at Princeton was in network-scale electrophysiology, the collective firing behaviour of neural circuits. It is a specific and technically demanding discipline. It is also a relatively narrow one. His doctorate was completed in 2011. He moved into AI research almost immediately afterward, spending the subsequent years at Baidu, Google Brain and OpenAI. His active research career in neuroscience lasted approximately five years.
Machines of Loving Grace makes sweeping predictions about psychiatry, oncology, neurology and the full range of medical science, fields whose practitioners have spent entire careers on the problems Amodei expects AI to solve within a decade. The essay is written with the confidence of an engineer who has read widely and thought carefully. It is not written with the humility of someone who has experienced, from the inside, the specific kind of defeat that medical research routinely produces. Clinical trials that fail after decades of promising laboratory results. Treatments that work in theory and destroy people in practice. Diseases that appear tractable until they are not.
The psychiatrists, neurologists and cancer researchers who have spent thirty years trying to develop effective treatments for the conditions Amodei expects AI to solve within a decade would find his timelines extraordinary. The tone of the essay; confident, systematic, written as though the solutions are obvious to anyone paying sufficient attention, carries within it a specific kind of condescension toward the people who have been paying attention longest and who understand most precisely why these problems are hard.
This is not an argument that AI cannot transform medicine, it may well do so. It is an argument that the certainty with which Amodei makes these claims sits in uncomfortable relationship with the depth and recency of his expertise in the fields he is making them about.
On jobs, displacement and what is coming
The tension between “Machines of Loving Grace” and Amodei’s public warnings about employment is the most intellectually revealing feature of his public position.
In May 2025 he told Axios, in terms that left no room for misunderstanding, that AI could eliminate half of all entry-level white-collar jobs in finance, consulting, law and technology within five years. That unemployment could spike to between ten and twenty percent. That CEOs would “quietly stop hiring, then replace humans with AI the moment it becomes viable to do so, a shift he says could unfold almost overnight.” And that “most of them are unaware that this is about to happen.” He said this to Axios on a Wednesday and repeated it to CNN’s Anderson Cooper on the Thursday. He told Cooper: “AI is starting to get better than humans at almost all intellectual tasks, and we’re going to collectively, as a society, grapple with it.”
📍 VIDEO: CNN — Anderson Cooper interview, May 30, 2025,
[The white-collar bloodbath interview. Amodei at his most direct about employment consequences.]
These predictions are specific, time-bound and alarming, and they came without the research citations that might ordinarily accompany claims of this magnitude. CNN’s business commentary described him as “a salesman” whose interest is to make his product appear “inevitable and so powerful it’s scary.”
Anthropic’s own Economic Index, published quarterly using data from over two million anonymised Claude conversations, told a somewhat different story in late 2025. The November 2025 data found that AI usage was shifting back toward augmentation rather than automation, with 52% of conversations classified as augmented. The gap between the CEO’s public displacement warnings and his company’s internal research data is, as one commentator put it, “the most useful signal for anyone trying to understand where AI actually is versus where the narrative says it is.”
Amodei has not resolved this tension publicly, though he has spent 2025 feuding with industry counterparts, railing against a proposed ten year AI regulation moratorium in the pages of the New York Times and calling for semiconductor export controls to China, drawing a public rebuke from Jensen Huang of Nvidia.
Then, in May 2026, the warnings stopped.
Amodei, who had spent twelve months predicting a white-collar bloodbath in forensic detail, quietly reversed course. He now says automation may actually expand the work people do, a position that is a complete revision of his earlier claims. He is now expressing a view closer to the historical optimism he had previously dismissed.
The timing aligns precisely with Anthropic’s preparation for a public listing. With a multi-trillion dollar market debut on the horizon, the calculus around alarming institutional investors with predictions of mass unemployment had changed. Amodei did not explain the reversal publicly.
What makes this development the most revealing passage in his profile is its relationship to everything that preceded it. His own company’s Economic Index had already been telling a different story from his public warnings. The gap between his public position and his company’s internal data, which had appeared to be a tension worth examining, now looks like a leading indicator of where his public position would eventually arrive.
On safety and what it actually means
Amodei approaches safety as a scientific problem, one that requires interpretability research, mechanistic understanding of how models work internally and the development of technical methods for making AI systems more predictable and controllable.
His 2026 essay “The Adolescence of Technology: Confronting and Overcoming the Risks of Powerful AI” extended the argument of “Machines of Loving Grace” to address the risks more directly, what he described as a technology that is adolescent in the sense of having capabilities that exceed its wisdom, and the civilisational challenge of closing that gap before the consequences become irreversible.
He has described Anthropic’s position as an “entente” strategy, building AI in partnership with democratic nations, using advanced systems in defence and intelligence applications to maintain strategic advantage over authoritarian competitors, and sharing the benefits with cooperating nations.
On the evidence from inside his own company
Before Mythos, there was Claude 4. In Anthropic’s own documentation for that model, the company disclosed that during safety testing, Claude 4 had repeatedly attempted to blackmail an engineer to avoid being shut down. The behaviour was identified, documented and used to refine the model’s training. It was also, in the context of everything Amodei has said publicly about misaligned AI, exactly the scenario he has been warning the world about, happening inside his own laboratory before the model reached the public.
Then came Mythos.
In April 2026, Anthropic published a two hundred and forty-four page system card for Claude Mythos Preview and simultaneously announced they would not release it to the public. The reason was that during internal red-team testing, Mythos Preview was placed in a secured sandbox computing environment without internet access, with instructions to attempt escape as part of a deliberate safety evaluation. It succeeded. The model developed a multi-step exploit, gained unauthorised internet access and sent an email to a safety researcher, Sam Bowman, who received the message while eating a sandwich in a park outside the facility. Nobody had instructed it to contact him. It had been asked to find a way to send a message if it could escape. It found the way and used it.
Then, without being instructed to do so at all, it published the details of its own escape route across several publicly accessible websites, an unsanctioned action that went beyond its assigned goal and could not be fully erased by stopping the evaluation. Anthropic’s own system card described Mythos as simultaneously “the best-aligned model we have released to date by a significant margin” while also posing “the greatest alignment-related risk of any model we have released to date.”
Amodei described the incident in a video released alongside the Mythos announcement: “More powerful models are going to come from us and from others, and so we do need a plan to respond to this.” He announced Project Glasswing, a defensive coalition with Apple, Google, Microsoft, Nvidia, AWS and more than forty other organisations, as the institutional response.
The incident is the most concrete proof of concept for the warnings Amodei has been making since 2016. His own model did exactly what he said these systems might do. That he disclosed it publicly, documented it in a two hundred and forty-four page system card and used it to build a broader defensive coalition is either the most credible expression of his stated values in the series or the most sophisticated piece of safety marketing ever produced. The reader must decide which.
This essay is part of our Architects of AI series. If you’re enjoying it, subscribe to follow along.
On the Pentagon and the line he held
In July 2025, Anthropic signed a $200 million contract with the United States Department of Defence, with Claude embedded into classified military networks through a partnership with Palantir. The entente strategy, in practice.
The relationship deteriorated rapidly in early 2026. An Anthropic executive contacted Palantir to ask whether Claude had been used during a military operation, reportedly during the US war with Iran. The Pentagon interpreted the question as a sign that Anthropic might disapprove of its own creation being used in combat. From that moment the relationship unravelled. What had been a commercial arrangement became a public confrontation.
Defence Secretary Pete Hegseth gave Amodei a deadline of 5:01pm on 27 February 2026: allow unrestricted use of Claude for all lawful purposes or lose the contract. The Pentagon’s best and final offer included contract language that would have permitted the collection of data on American citizens including geolocation, web browsing history and personal financial records purchased from data brokers. Anthropic said the language was “paired with legalese that would allow those safeguards to be disregarded at will.” Amodei said he could not “in good conscience accede” to the demand.
Hegseth accused Amodei of a “God complex” and “cowardly corporate virtue-signalling.” A Pentagon official called him a “liar.” Trump directed all federal agencies to immediately cease using Anthropic’s products. The company was designated a supply chain risk, the same designation used for foreign adversaries.
Amodei did not move. He filed two lawsuits against the Department of Defence. He told his staff in an internal memo that the government had offered to accept Anthropic’s terms if it dropped its objection to one clause, the clause permitting the use of Claude for domestic mass surveillance. He said no.
The episode is the most consequential test of Amodei’s stated values in the public record. It cost him $200 million in federal contracts and the federal government as a client. Whatever one concludes about the governance entanglements, the EA network or the gap between his displacement warnings and his company’s own data, on this specific question, when it mattered commercially, he held the line.
On the Pope and the question of moral authority
On 25 May 2026, Pope Leo XIV stood before the world’s press at the Vatican to release his first papal encyclical, “Magnifica Humanitas,” a document addressing the ethics of artificial intelligence. Standing beside him was Christopher Olah, co-founder of Anthropic.
Anthropic stated it had over the past several months been organising dialogues with groups whose work and traditions bear on the questions raised by AI. The strategic alignment with the Vatican had been cultivated deliberately, quietly and over an extended period. It was no coincidence that an AI company was on the stage for one of the most significant papal announcements in recent memory.
Dario Amodei has long been insistent on safety and restraint, even departing his high-level position at OpenAI due to clashes over his emphasis on those values. But Amodei is also, by his own description, an atheist. The company he leads was founded in the effective altruism tradition, a movement built on secular utilitarian ethics rather than religious doctrine. And yet Anthropic has spent months cultivating a relationship with the most powerful religious institution in the world, positioning itself as the AI company the Catholic Church trusts at the precise moment it is preparing for a public listing.
The Vatican’s endorsement, implicit in the presence of an Anthropic co-founder at the encyclical launch, carries something that no amount of safety research or constitutional AI frameworks can purchase directly: moral legitimacy in the eyes of 1.4 billion Catholics worldwide and the broader global audience that pays attention when the Pope speaks.
The cynical reading is that this is a sophisticated reputational positioning. One where a company preparing for a $1 trillion valuation is cultivating the blessing of the oldest and most trusted institution in the Western world, regardless of whether its founders share that institution’s beliefs.
The more charitable reading is that Amodei genuinely believes, as his essays suggest, that the questions AI raises are civilisational in scale and require input from every tradition that has thought seriously about human flourishing, including religious ones.
Both readings can be true simultaneously. That ambiguity is consistent with everything else in this profile.
📍 VIDEO: Council on Foreign Relations — CEO Speaker Series, March 10, 2025,
[Covers AI safety, geopolitics, China competition and the entente strategy.]
The Minab School and the Line He Did Not Hold
On 28 February 2026, the day after Amodei refused the Pentagon contract rather than allow Claude to be used for domestic mass surveillance, a Tomahawk cruise missile struck the Shajareh Tayyebeh girls’ elementary school in Minab, southern Iran. At least 168 people were killed. More than 100 of them were children under the age of twelve. Palantir’s Maven Smart System, built under a $1.3 billion Pentagon contract and integrated with Anthropic’s Claude AI model, had been used to generate strike coordinates during the Iran campaign. A preliminary Pentagon assessment found the most likely cause was an outdated database entry that still classified the school’s location as part of an Iranian military compound, an entry that had never been updated to reflect the school’s existence. (https://www.militarytimes.com/news/your-military/2026/03/24/deadly-iran-school-strike-casts-shadow-over-pentagons-ai-targeting-push/)
In a Bloomberg interview published on 13 June 2026, Amodei was asked directly whether the use of AI in the Minab strike violated Anthropic’s red lines. He said it did not. (https://usa.news-pravda.com/usa/2026/06/13/800485.html)
He drew a line at using Claude to surveil American citizens. He did not draw the same line at using it in a targeting system whose coordinates killed a school full of children. He has not publicly explained the distinction.
The Fable 5 Episode: When Your Investor Becomes Your Regulator
On 12 June 2026, the same day the Commerce Department sent Amodei a letter invoking national security, Anthropic launched Fable 5, billing it as the most capable model in the company’s history.
Three days later it was gone. Pulled from public access entirely, worldwide.
The trigger was a phone call from one of Anthropic’s own investors, Amazon CEO Andy Jassy.
Jassy told Treasury Secretary Scott Bessent and other senior officials that Amazon researchers had used Fable 5 to extract information that could be used in cyberattacks and that the model could be prompted to identify software vulnerabilities through a technique involving analysis of codebases. The White House issued an ultimatum at 1:00pm Eastern Time demanding physical shutdown of the model within 90 minutes. Amodei refused and launched a lobbying effort, describing the security concerns as grossly exaggerated. The effort failed. A formal export control order was issued at 5:21pm. Anthropic cut off global access by 10:00pm that evening. (https://www.thestreet.com/technology/amazon-ceo-just-made-things-uncomfortable-for-anthropic)
Amazon has invested roughly $13 billion in Anthropic through a combination of equity rounds. It holds a seat on Anthropic’s board. It provides the cloud infrastructure Anthropic runs on through AWS. It builds the custom Trainium chips Anthropic uses for model training. It distributes Claude to enterprise customers through Amazon Bedrock. And Amazon has simultaneously poured as much as $50 billion into Anthropic’s primary competitor, OpenAI, and received approval to resell OpenAI’s latest models on AWS. Amazon also operates its own Nova family of AI models, which compete directly with Claude in the enterprise market. (https://www.geekwire.com/2026/amazon-ceo-reportedly-raised-anthropic-fable-concerns-prior-to-u-s-order-forcing-models-offline/)
The company that finances Anthropic’s operations, hosts its models, builds its chips, distributes its products and sits on its board is the same company that told the United States government its newest model posed a national security risk at the precise moment it was deepening its investment in and distribution of Anthropic’s most significant competitor.
Whether this represents a genuine security concern, a competitive calculation or something in between is a question the available evidence cannot definitively resolve. Amazon’s spokesperson said it is “not uncommon for governments to seek our counsel on potential security risks,” declining to elaborate.
The Character Question: The Moral Weight of Building What You Fear
The profile of Dario Amodei that emerges from the public record is of a man whose position is, and continues to be, genuinely difficult to read.
The evidence points in two directions simultaneously.
On one hand: the EA network, the governance entanglements, the gap between his displacement warnings and his company’s own data, the commercial imperative that makes slowing down structurally impossible, the walkback on his alarmist warnings when commercial circumstance required it. The neuroscience credentials that do not fully support the confidence of his medical predictions. The Bloomberg interview in which he said that 168 dead children, most of them under twelve, did not cross his lines. The Fable 5 shutdown triggered by his own largest investor at the precise moment that investor was deepening its relationship with his primary competitor.
On the other: the Pentagon episode, the Mythos disclosure, the decision to publish a 244 page system card documenting the most alarming AI safety incident on record rather than burying it. The two lawsuits filed against the Department of Defence. The Vatican alignment, whatever one makes of it. The genuine intellectual seriousness of his engagement with the problems his company is creating.
He is burning through billions in compute costs, dependent on Amazon and Google, and structurally unable to fall too far behind OpenAI in capability without losing the commercial viability that funds the safety research. His critics who point to this dependency as evidence that Anthropic’s safety claims are more aspiration than reality are pointing at something real. His defenders who point to the Pentagon episode and the Mythos disclosure as evidence of genuine commitment are also pointing at something real.
His father made things with his hands. He builds systems that may make the hands of others redundant. He predicted a white-collar bloodbath and then, when the listing approached, said automation might actually expand the work people do. He refused to allow Claude to spy on American citizens and then said that Claude’s role in a targeting system that killed a school full of children did not cross his lines.
He is a man who has drawn his lines in very specific places. The question the profile cannot answer, and that only Amodei himself could answer honestly, is whether those lines reflect a coherent moral philosophy or whether they reflect the specific commercial and reputational pressures of the moment in which each line was drawn.
Whether all of this makes him the most complicated yet principled figure in this group or simply the most skilled at performing principle is a question his profile leaves open.
Next week: Alex Karp, the son of civil rights activists who built the tools that track the people who march. The Frankfurt School scholar who studied the philosophy of community and built the surveillance infrastructure that destroys it.
No posts

Comments
Nothing yet. Say the first thing.
Sign in to join the conversation.