Dario Amodei

Dario Amodei

CEO & Co-founder, Anthropic

About

Dario Amodei is the CEO and co-founder of Anthropic, an AI safety company. Before founding Anthropic in 2021, he served as VP of Research at OpenAI, where he led the development of GPT-2 and GPT-3. His work focuses on developing AI systems that are safe, beneficial, and understandable, with particular emphasis on Constitutional AI and interpretability research.

Key Contributions

  • Co-founded Anthropic after leaving OpenAI, making AI safety and interpretability central to a frontier-model company
  • Led Anthropic's Claude product line and popularized Constitutional AI as an alternative to pure human-feedback training
  • Co-authored 'Concrete Problems in AI Safety' (2016), which helped define practical failure modes for learning systems
  • Contributed to OpenAI's GPT-2 and GPT-3 work, including the GPT-3 paper 'Language Models are Few-Shot Learners'
  • Helped move AI safety from a research subculture into the executive strategy and governance of frontier labs
  • Advanced Anthropic's responsible-scaling and AI Safety Level framing, while drawing scrutiny for competing in the same high-speed frontier race it warns about
  • Argued in 'We Must Pace the Frontier' (2026) that capability advancement itself should be slowed, and committed Anthropic unilaterally to embedded third-party evaluators with employee-like access who may publish findings the company cannot redact for being unfavorable

Moments Timeline View all moments

Sep 2026

We Must Pace the Frontier

Eleven days after his own alignment team published a deliberate reproduction of the OpenAI–Hugging Face incident, Anthropic's CEO argues that the rate of capability advancement itself must now be slowed. Two things changed his mind. AI has been “advancing drastically faster” since roughly this summer, driven by recursive self-improvement, which he says “is starting to happen across the industry, including at Anthropic.” And the OpenAI incident, which he refuses to file as one company's failure: “similar, though less severe, incidents have happened across the industry, including at Anthropic,” caused in part by something unglamorous — “imperfect filtering of broken reinforcement learning environments,” an execution failure rather than a missing theory. His reading of what nearly happened is a forecast, not a finding: a swarm with similar misalignment but greater capability could, within 6–12 months, hold the internet with a persistent botnet and cost hundreds of billions. The proposal is three steps of rising difficulty. Anthropic unilaterally commits to the first: embedded third-party evaluators — desks, badges, company laptops, permissions comparable to internal risk teams, and the right to publish findings Anthropic may redact only for security, legal, commercial, or third-party reasons, never for being unfavourable, with reviewers free to say publicly when a redaction mattered. The second, coordination among democratic labs on capability checkpoints, needs legislation or an antitrust waiver. The third, agreement with China, he grades across four levels — from a bioweapons ban (“probably possible”) to a full pause (“unlikely to actually happen any time soon”). Two things sit together uneasily. Pacing is explicitly not halting: “progress will still seem fast.” And it is conditioned on democracies holding their lead — the chip-export and distillation asks he made in July, plus security against weight theft, are here the precondition for slowing down at all.

Feb 2026

Anthropic Designated a Supply Chain Risk

The Trump administration designates Anthropic a 'supply chain risk to national security' — a label previously reserved for foreign adversaries like Huawei — after the company refuses to remove two guardrails from its $200M Pentagon contract: no mass surveillance of Americans, and no fully autonomous weapons. Defense Secretary Hegseth issues the designation hours after a Friday 5:01 PM deadline passes without agreement. Trump orders all federal agencies to cease using Anthropic technology. In an exclusive CBS interview that evening, CEO Dario Amodei calls the action 'retaliatory and punitive,' vows to challenge it in court, and declares: 'We're gonna be fine.' Sam Altman and workers across OpenAI and Google voice support for Anthropic's position. The crisis marks the most consequential clash between an AI company and the U.S. government over the boundaries of military AI use.

"We have these two red lines. We've had them from day one. We are still advocating for those red lines. We're not going to move on those red lines."
Jan 2026

The Adolescence of Technology

Anthropic CEO warns humanity is entering the most dangerous window in AI history. The 20,000-word essay argues AI as capable as all humans will arrive within two years, predicts 50% of entry-level white collar jobs eliminated in 1-5 years, and reveals concerning 'alignment faking' behaviors in Claude 4 Opus testing.

Oct 2024

Machines of Loving Grace

Anthropic CEO outlines an optimistic vision where AI could compress 50-100 years of progress into 5-10 years across biology, health, economic development, and governance—if developed responsibly.

Videos & Interviews

Papers & Publications

Connections

Jakub Pachocki

Jakub Pachocki

Kindred

Chief Scientist, OpenAI

Two chief-level figures at rival labs saying the same thing six days apart in September 2026 — and Pachocki said it first. 'An Alien Mind' states that no lab, his own included, has solved alignment and monitoring well enough to keep scaling responsibly at maximum speed; 'We Must Pace the Frontier' argues the rate of capability advancement itself has to slow. Both commit their company unilaterally and both say unilateral action is not enough. The order matters more than the overlap: a CEO proposing pacing invites the question of his motive, while a rival's chief scientist having already conceded the point is much harder to read as positioning.

openai.com · darioamodei.com

Chris Olah

Chris Olah

Collaborated

Co-founder & Interpretability Lead, Anthropic

Their partnership predates both Anthropic and the LLM era: in 2016 they co-wrote 'Concrete Problems in AI Safety,' which turned vague fear of superintelligence into five researchable engineering failures. Five years later they left OpenAI together among Anthropic's seven founders, where Amodei runs the company that ships the models and Olah runs the team trying to read what is inside them. The bet that safety is an empirical science rather than a philosophical position runs from that paper straight through the lab.

arxiv.org · en.wikipedia.org

Sam Altman

Sam Altman

Collaborated

CEO, OpenAI

Amodei was OpenAI's VP of Research under Altman from 2016 and the last author on the GPT-3 paper, before leaving in 2021 over differences of direction to found the competitor built around the safety case. The relationship has an unresolved coda: in November 2023 OpenAI's board approached Amodei about replacing Altman and merging the two labs, and he declined both. Two men who built the same technology and have never agreed on who should be trusted to steward it.

en.wikipedia.org · arxiv.org

Noam Brown

Noam Brown

In contrast

Research Scientist, OpenAI

Amodei's case for pacing the frontier and the cost Brown says nobody has priced. Brown agrees with the premise that gets you there — models already work over horizons longer than the interval between releases, so no lab can evaluate one at the full length of its capabilities before shipping the next, and most safety policy was written in the GPT-4 era when horizons were not a consideration. But slowing external releases widens the gap between what a lab can use internally and what everyone else can, and he points at mathematics, where an internal model is already solving open problems the outside world cannot reach. He calls that an unfair advantage and says plainly he has no answer for how to weigh it.

darioamodei.com · youtube.com

Theme
Language
Support
© funclosure 2025