Researchers at Anthropic and OpenAI warn AI may soon improve itself faster than humans can follow

A resignation letter and a string of public warnings have put a once-theoretical AI risk into the headlines. Here is what researchers are actually worried about, in plain English.

AI2Day NewsdeskEditor: Lee Brown4 min read
Extreme close-up of a dense grid of glowing quantum circuit traces on a dark matte surface, cool blue and pale white light, shallow depth of field pulling the g
Share

Key points

  • Evan Hubinger, an alignment lead at Anthropic, said on X that he believes there is more than a 10% chance AI could kill all humans within the next decade.
  • Researchers at both Anthropic and OpenAI publicly warned this week that AI systems are already accelerating their own development faster than expected.
  • Anthropic reported in August that its engineers now ship, on average, eight times as much code per quarter as they did between 2021 and 2025, partly because AI helps write the code.
  • OpenAI Chief Scientist Jakub Pachocki wrote on Saturday that future systems will likely drive their own development.
  • Anthropic outlined three possible futures, calling the scenario where AI remains under human control "likely", but did not rule out a path where humans play a "substantially diminished role".

The warnings started with a resignation. After a colleague quit Anthropic citing safety fears, Hubinger posted on X that he personally puts the odds of AI killing all humans within ten years at better than one in ten. More researchers at both labs followed. We first covered the resignation itself on 9 September in "An Anthropic Safety Researcher Quit. Then His Colleague Said AI Has a 1-in-10 Chance of Killing Everyone."

What is recursive self-improvement, and why does it worry researchers?

Recursive self-improvement, or RSI, is the idea that AI systems could reach a point where they help design and train the next generation of AI, which are then smarter, which then build even smarter successors, and so on. Once that loop spins fast enough, the humans who built the original systems lose the ability to steer what comes next.

AI hasn't reached that point yet, but researchers say it's already doing a version of it. Anthropic said in an August blog post that its engineers ship eight times as much code per quarter as they did between 2021 and 2025, with AI doing much of the heavy lifting. "AI is already at the level where it can introduce some new ideas," Vincent Conitzer, professor of computer science at Carnegie Mellon University, told CNBC Tech. "So it is very hard to predict at what point this process would start to drastically accelerate."

What did researchers say publicly this week?

The warnings were unusually direct. Jasmine Wang, an OpenAI researcher working on alignment (the field focused on making AI systems do what humans actually want), wrote on Wednesday that it's "hard to overstate how dangerous speeding towards RSI is." Anna Wang, who works on AGI safety at Anthropic, said there is "not yet a viable scientific plan to solve risks from recursively self-improving AI."

Pachocki wrote in a company blog post on Saturday that systems arriving in the next few years will likely "increasingly drive their own development."

Who said it Role Core warning
Evan Hubinger Alignment lead, Anthropic >10% chance AI kills all humans within a decade
Jakub Pachocki Chief Scientist, OpenAI Future systems will increasingly drive their own development
Jasmine Wang Alignment researcher, OpenAI RSI is dangerous and we are heading toward it fast
Anna Wang AGI safety, Anthropic No scientific plan yet exists to manage RSI risks

What happens next?

Anthropic laid out three scenarios. In the first, AI progress stalls and the technology spreads widely but stays limited. Anthropic called this unlikely. The second has labs making gains with humans in control. Anthropic called this one "likely." The third sees AI reach full recursive self-improvement, with humans playing a "substantially diminished role" in what gets built. What that means for the alignment problem, Anthropic admitted, is what it's "least certain about."

The honest read here: the researchers building the most powerful AI systems are genuinely unsure whether they can keep control of what they're making. That isn't hype. It doesn't mean the worst outcomes are inevitable, but the alignment problem doesn't have a solved answer yet, and the people closest to it are saying so publicly.

Common questions

Does this mean AI will definitely become dangerous?

No. Researchers describe these as real risks worth serious attention, not certainties. Anthropic itself called the scenario where humans stay in control the most likely outcome.

Should I be worried about AI tools I use today?

The concerns researchers raised are about future, far more capable systems, not the chatbots and writing assistants available now. Current tools carry their own, narrower risks such as errors and bias, separate from the RSI debate.

© 2026 AI2Day