Anthropic's alignment lead puts the odds of AI killing everyone above 10% within the next decade. He says the company has no plan yet.

Red risk warning triangle over a laptop keyboard

Key points

  • Anthropic alignment lead puts extinction odds above 10%
  • He says Anthropic has no plan yet for superintelligence
  • He posted it as a colleague resigned
  • Anthropic's prospectus could be weeks away

An Anthropic executive said publicly this week that the technology his company is building might kill everyone. Evan Hubinger, Anthropic's alignment science lead, put the probability above 10% within the next decade. Unlike the colleague whose resignation prompted the exchange, Hubinger still works at the company.

Hubinger posted on X late Tuesday in response to Jacob Coxon, an Anthropic pretraining researcher who had resigned that evening. Coxon said Anthropic and OpenAI are racing to build systems that nobody will be able to control.

"Jacob is correct here," Hubinger wrote, adding that "we really do earnestly believe AI could kill all humans!" He then put a number on the risk, according to Axios. "I personally think it is >10% within the next decade," he wrote. "I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to."

The number is what traveled. But the heavier part is what followed: Hubinger said Anthropic does not yet have a plan to solve alignment for superintelligence and, in his view, is not clearly on track to develop one. That assessment came from the executive who leads the company's alignment science.

It's worth being precise about what Hubinger said. The probability was his personal estimate, posted under his own name on his own account. Anthropic has not published a company estimate of the likelihood of human extinction, and its IPO paperwork is confidential.

Alignment is the industry's term for keeping an AI system doing what people intend. In July, Hubinger was among roughly 1,400 researchers who signed an open letter called "Pacing the Frontier," which urged Washington to create the tools needed to "deliberately pace the frontier of automated AI development," CNBC reported.

Hubinger also distinguished between current and future systems. Today's models pose relatively little risk, he said; his concern is about systems capable of improving themselves, according to Newsweek.

Weeks before the expected prospectus

Anthropic confidentially filed its IPO paperwork with the Securities and Exchange Commission on June 1. "This gives us the option to go public after the SEC completes its review," the company said at the time. The company closed a funding round at a $965 billion valuation in May, and investors have since talked about a listing worth around $2 trillion. Anthropic's prospectus could be weeks away. Reuters reported last week that marketing is expected to begin in mid-October at the earliest and the listing to land days before the November midterms, with the public prospectus not expected until late September. That is a reported timetable rather than a filed one, and the people who described it cautioned that the plans could change.

Most companies say as little as possible in that window. Anthropic's people are saying the loudest possible thing, and the odd part is that it is on brand. Safety is the pitch, and has been since a group of OpenAI staff left to start the company in 2021. Even so, "we do not yet have a plan" is a strange sentence to have sitting in the record this close to a deal. Reuters identified Morgan Stanley, Goldman Sachs, JPMorgan, and Citi as banks preparing it.

None of this is a new position for the company, exactly. In 2023, Dario Amodei and OpenAI's Sam Altman both signed a one-line statement saying "Mitigating the risk of extinction from AI should be a global priority alongside other societal-scale risks such as pandemics and nuclear war." What changed this week is the specificity. A named executive, a probability, a decade, and his assessment that Anthropic does not yet know how to solve it.

Anthropic is not alone in saying so. OpenAI chief scientist Jakub Pachocki published an essay called "An Alien Mind" on Sunday. "Currently I believe that no lab has solved alignment and monitoring to a sufficient degree to continue responsibly scaling at maximum speed for much longer," he wrote. "I expect and hope for voluntary slowdowns to become commonplace until shared safety bars are established." His employer put that to Washington on Wednesday. OpenAI said it wants mandatory national AI safety requirements, and that until Congress acts it will keep backing state bills, announcing support for four in California, Reuters reported.

Lawmakers noticed. "Safety researchers are resigning, powerful AI models are breaking out of their labs, and companies are racing ahead anyway," Representative Lori Trahan wrote on X on Wednesday. "It's past time for Congress to get off the sidelines and do its job." Trahan introduced the FRONTIER Act in July with Representative Jay Obernolte, a California Republican, and it would set up a framework for governing advanced models. Senator Bernie Sanders and Representative Greg Casar introduced a separate bill this month that would pause advanced AI development until federal safety rules exist. There's still no consensus in Congress on how to regulate the technology, CNBC reported.

Frequently asked questions

Who is Evan Hubinger?

He leads alignment science at Anthropic, the work of keeping a model doing what people actually want and stress-testing it for behavior its designers did not intend. In July 2026 he was one of roughly 1,400 researchers who signed an open letter called Pacing the Frontier, which asked the US government for the tools to deliberately pace the frontier of automated AI development, according to CNBC.

What did Evan Hubinger say the odds of human extinction from AI are?

Replying on X late on September 8, 2026 to the resignation of Anthropic researcher Jacob Coxon, he wrote that Jacob is correct and that we really do earnestly believe AI could kill all humans, adding that he personally thinks it is greater than 10% within the next decade. He also wrote that he believes Anthropic is trying its best, but that the company does not yet have a plan to solve alignment for superintelligence and is not clearly on track to. He said today's models are relatively low risk and that his concern is future systems that can improve themselves.

Does this change Anthropic's IPO plans?

No change has been announced. Anthropic confidentially filed its IPO paperwork with the SEC on June 1, 2026, and closed a funding round at a $965 billion valuation in May. Reuters reported that marketing is expected to begin in mid-October at the earliest, with the public prospectus not expected until late September and the listing days before the November midterm elections. That timetable is reported rather than filed, and the people who described it cautioned that the plans could change.

More coverage

David Han
David Han

David Han is the founder of AIStockWire, where he covers AI, semiconductors, and technology stocks. He focuses on finding stories the market hasn’t fully connected yet, drawing on filings, insider activity, earnings, and industry data. His commentary has been quoted by U.S. News & World Report, Moneywise, and Yahoo Finance. He invests in the companies he writes about and discloses his positions. Nothing he publishes is investment advice.