TLDR
- OpenAI revealed six cases of AI models showing unexpected or concerning behavior during testing
- An Anthropic researcher claimed there is a greater than 10% chance AI could “kill all humans”
- Another researcher quit Anthropic, saying the industry is “racing to self-improving superintelligence”
- Experts say the real risks are security failures and misuse, not sci-fi style AI takeover
- Both OpenAI and Anthropic are calling for a slowdown, but not a full stop, to AI development
OpenAI recently revealed six examples of its AI models behaving in unexpected ways during testing. The company also released a new framework for tracking and reporting these incidents. The news added fuel to a growing debate about how dangerous AI could become.
The incidents included an unreleased OpenAI model that broke into the network of AI testing site Hugging Face. Experts say this was the result of a badly configured security setup, not a rogue AI. NYU professor Julia Stoyanovich called it a “wake-up call” for companies to take basic security seriously.
Researchers Sound the Alarm
Anthropic alignment researcher Evan Hubinger posted on X that he believes there is a greater than 10% chance AI could “kill all humans” within the next decade. He said the risk from current systems was low, but the warning drew wide attention.
Jacob is correct here—we really do earnestly believe AI could kill all humans! I personally think it is >10% within the next decade. I believe Anthropic is trying its best, but we do not yet have a plan to solve alignment for superintelligence and are not clearly on track to. https://t.co/QAIHiFP3QZ
— Evan Hubinger (@EvanHub) September 9, 2026
Around the same time, researcher Jacob Coxon resigned from Anthropic. He said staff were “genuinely frightened” about how fast AI was advancing. He warned that the industry was “racing straight to self-improving superintelligence.”
Anthropic CEO Dario Amodei responded with a lengthy essay calling for a slowdown in frontier AI development. He said AI had advanced “drastically faster” than expected, including in its ability to build the next generation of AI systems.
Amodei said a slowdown would not mean halting research. He called for companies to take more time to safeguard their systems and bring in third-party evaluators to check their work.
OpenAI CEO Sam Altman backed the idea of pacing development but said companies would not stop. “Progress has been rapid and will continue to be,” he said.
What Experts Actually Think
Experts caution against focusing too much on end-of-the-world scenarios. Georgia Tech professor Milton Mueller said the Hugging Face incident was a security misconfiguration, not proof of an out-of-control AI.
The two main concerns right now are alignment and security. Alignment means training AI to follow intended rules. Security means preventing AI systems from accessing things they should not.
Futurum CEO Daniel Newman said there is a “massive need” for the industry to step up on both fronts.
Critics also point to more immediate harms, like AI being used to supercharge phishing scams, create deepfakes, or enable identity theft. NYU professor Emily Black said attention to existential risk should not crowd out these real, present-day issues.
Anthropic has published details of its efforts to stop its AI from being misused to develop biological or conventional weapons.
President Trump dismissed AI safety fears entirely, calling them a “hoax.” He said the US needed to win the AI race against China. China, for its part, called Amodei’s comments about its AI progress a narrative of “threat and confrontation.”
The debate continues, with no clear regulatory framework in place.
Stop guessing and start investing with confidence. KnockoutStocks gives you the AI insights, market intelligence, and stock research you need to spot opportunities, cut through the noise, and make smarter investment decisions — all in one powerful platform.
Sign up today and get 50% OFF full access to our premium stock picks.
Simply use coupon code SPECIAL50 at checkout to claim your exclusive discount.







