AI Almost Started a U.S.–China War — and No One Seems to Care
The Intercept · L · trust 50/100

The tech world is too busy dreaming up imaginary doomsday scenarios to focus on a very real one that flared up between nuclear powers.
Share Copy link Share on Facebook Share on Bluesky Share on X Share on LinkedIn Share on WhatsApp Protesters spin the “Wheel of the Probability of Doom” during a demonstration by the group Pause AI outside Downing Street in London on Sept. 16, 2026. Photo: Vuk Valcic/SIPA Images/Sipa USA via AP Images The global tech industry, politicians at home and abroad, and international media coverage are transfixed by one question: Will rogue AI end humanity? Few seem to be asking this question of the Department of Defense.
The artificial intelligence sector has rocked itself in recent weeks following a string of disclosures from the frontier AI labs and some of their personnel. OpenAI and Anthropic have both disclosed incidents in which their software, during routine internal testing, unexpectedly broke into the networks of other corporations . The tests were akin to a disastrous demonstration of a guided missile: After humans hit launch, the autonomous technology veered off course in deeply alarming ways.
Observers and industry figures quickly interpreted the incidents not simply as an indicator of the software’s power. Instead, they argued that it presaged an age of computers possessing something no computer ever has: bad intentions.
If computer systems could behave in ways that are unexpected — or even shocking — without being directly steered by humans, it seems only a short step until computers might have desires, ambitions, and even malice. If OpenAI’s tools could unexpectedly crack the servers of a rival corporation to accomplish a task posed to it by engineers, couldn’t it also surprise us with actions that result in harm to people, or even deaths?
Frontier lab employees and executives have answered with an emphatic yes. On September 8, Anthropic researcher Jacob Coxon announced his resignation on X, writing , “The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt.” Coxon’s former Anthropic colleague Evan Hubinger replied, casually agreed: “Jacob is correct here — we really do earnestly believe AI could kill all humans!” Hubinger noted he personally puts the odds of the species’ extermination by AI, somehow, at over 10 percent.
This has all resulted, unsurprisingly, in a global panic, with a sudden flurry of calls for regulatory intervention, from legislation that would require an emergency “kill switch” for AI systems, to a proposal by Sen. Bernie Sanders to halt AI development altogether.
Throughout this period of alarm, how exactly a large language model could literally end humanity has remained vague. Some have speculated that this technology could foster the creation of some sort of novel bioweapon; others worry many thousands of AI agents could somehow hack all vital infrastructure simultaneously. In both those supposed doomsdays, the details are still fuzzy.
It should have come as quite the shock, then, when CNN reported on September 18 of a recent incident in which the use of artificial intelligence could have genuinely led to the extinction of the species.
It was under direct human supervision that a large language model almost misled the U.S. into instigating a war with his country.
This past spring, according to the news outlet, a U.S. Special Operations Command analyst used a large language model to generate an intelligence report that indicated “a Chinese ship in the Middle East was transporting components of a nuclear weapons program.” The finding prompted the U.S. military to make quick preparations to intercept the vessel by force. But this AI-generated intelligence, according to one source who spoke to CNN, was “entirely false.” Yet, the source said, it “almost started a war.”
A shooting war between the U.S. and China could play out in an incalculably wide variety of ways. But one entirely plausible path would be an exchange between the world’s first and third largest nuclear weapons arsenals — an event that would transcend warfare into global cataclysm.
CNN’s reporting garnered attention, but was not followed by nearly the degree of sustained, grave concern of the AI safety news cycles that preceded it. It didn’t seem to prompt company scientists to question their careers, nor did it spur calls for regulatory intervention or self-imposed limits by the companies who furnish the Pentagon with this technology. After weeks of discussion of how AI could hypothetically kill everyone, the public learned of a concrete way in which AI really could have started a war that might have killed everyone, and the world quickly lost interest.
“This [China] incident and the lack of response is an attention problem that the nuclear field has been grappling with for decades.”
At a summit this week with President Donald Trump, Chinese President Xi Jinpeng argued that nations must “ensure that the development of AI is always under human control,” again underscoring fears of AI breaking free of human oversight. But it was under direct human supervision that a large language model almost misled the U.S. into instigating a war with his country.
The discrepancy illustrates a gap in concern by both the general public and policymakers over AI risks: There is great alarm over an AI system “going rogue,” disobeying its commands and the interests of its creators and operators, and pursuing a malign agenda of its own. Much of the AI safety discourse revolves around the hypothetical threat of “ superintelligence ,” defined broadly as a computer system that can outthink even the smartest humans across domains. There’s also the belief that software systems will both achieve sentience and bear a grudge against their creators and act to undermine them, if not destroy them outright.
Yet there is considerably less public alarm over AI performing exactly the actions that the Pentagon asks of…
Read the original at The Intercept →
Open in TruthVane →