The global tech
industry, politicians at home and abroad, and international media coverage are transfixed by one question: Will rogue AI end humanity? Few seem to be asking this question of the Department of Defense.
The artificial intelligence sector has rocked itself in recent weeks following a string of disclosures from the frontier AI labs and some of their personnel. OpenAI and Anthropic have both disclosed incidents in which their software, during routine internal testing, unexpectedly
broke into
the
networks
of
other corporations
. The tests were akin to a disastrous demonstration of a guided missile: After humans hit launch, the autonomous technology veered off course in deeply alarming ways.
Observers and industry figures quickly interpreted the incidents not simply as an indicator of the software’s power. Instead, they argued that it presaged an age of computers possessing something no computer ever has: bad intentions.
If computer systems could behave in ways that are unexpected — or even shocking — without being directly steered by humans, it seems only a short step until computers might have desires, ambitions, and even malice. If OpenAI’s tools could unexpectedly crack the servers of a rival corporation to accomplish a task posed to it by engineers, couldn’t it also surprise us with actions that result in harm to people, or even deaths?
Frontier lab employees and executives have answered with an emphatic yes. On September 8, Anthropic researcher Jacob Coxon announced his resignation on X,
writing
, “The people building AI earnestly believe that it could kill us all by the end of the decade. This is not a marketing stunt.” Coxon’s former Anthropic colleague Evan Hubinger replied, casually agreed: “Jacob is correct here — we really do earnestly believe AI could kill all humans!” Hubinger noted he personally puts the odds of the species’ extermination by AI, somehow, at over 10 percent.
Tech CEOs’ Doomsaying Is a Distraction f
… [more]