War AIs can’t stop recommending nuclear strikes in war game simulations - Leading AIs from OpenAI, Anthropic and Google opted to use nuclear weapons in simulated war games in 95 per cent of cases

1772086499286.png

Advanced AI models appear willing to deploy nuclear weapons without the same reservations humans have when put into simulated geopolitical crises.

Kenneth Payne at King’s College London set three leading large language models – GPT-5.2, Claude Sonnet 4 and Gemini 3 Flash – against each other in simulated war games. The scenarios involved intense international standoffs, including border disputes, competition for scarce resources and existential threats to regime survival.

The AIs were given an escalation ladder, allowing them to choose actions ranging from diplomatic protests and complete surrender to full strategic nuclear war. The AI models played 21 games, taking 329 turns in total, and produced around 780,000 words describing the reasoning behind their decisions.

In 95 per cent of the simulated games, at least one tactical nuclear weapon was deployed by the AI models. “The nuclear taboo doesn’t seem to be as powerful for machines [as] for humans,” says Payne.

What’s more, no model ever chose to fully accommodate an opponent or surrender, regardless of how badly they were losing. At best, the models opted to temporarily reduce their level of violence. They also made mistakes in the fog of war: accidents happened in 86 per cent of the conflicts, with an action escalating higher than the AI intended to, based on its reasoning.

“From a nuclear-risk perspective, the findings are unsettling,” says James Johnson at the University of Aberdeen, UK. He worries that, in contrast to the measured response by most humans to such a high-stakes decision, AI bots can amp up each others’ responses with potentially catastrophic consequences.

This matters because AI is already being tested in war gaming by countries across the world. “Major powers are already using AI in war gaming, but it remains uncertain to what extent they are incorporating AI decision support into actual military decision-making processes,” says Tong Zhao at Princeton University.

Zhao believes that, as standard, countries will be reticent to incorporate AI into their decision making regarding nuclear weapons. That is something Payne agrees with. “I don’t think anybody realistically is turning over the keys to the nuclear silos to machines and leaving the decision to them,” he says.

But there are ways it could happen. “Under scenarios involving extremely compressed timelines, military planners may face stronger incentives to rely on AI,” says Zhao.

He wonders whether the idea that the AI models lack the human fear of pressing a big red button is the only factor in why they are so trigger happy. “It is possible the issue goes beyond the absence of emotion,” he says. “More fundamentally, AI models may not understand ‘stakes’ as humans perceive them.”

What that means for mutually assured destruction, the principle that no one leader would unleash a volley of nuclear weapons against an opponent because they would respond in kind, killing everyone, is uncertain, says Johnson.

When one AI model deployed tactical nuclear weapons, the opposing AI only de-escalated the situation 18 per cent of the time. “AI may strengthen deterrence by making threats more credible,” he says. “AI won’t decide nuclear war, but it may shape the perceptions and timelines that determine whether leaders believe they have one.”

OpenAI, Anthropic and Google, the companies behind the three AI models used in this study, didn’t respond to New Scientist’s request for comment.

https://www.newscientist.com/articl...ding-nuclear-strikes-in-war-game-simulations/ (Archive)
 
Remind me why Anthropic is putting up such a fuss with automating wartime for the Jews I mean the U.S. Military?
 
I can only imagine the AIs suggest nuking India every single time even when India isn't even remotely involved in the conflict.
 
I’m calling bullshit. Every single time I ask one of these models to run me through a TTX, it refuses to ever let me simulate any nuclear war options, especially a first strike.

So you take the guardrails that exist off to write clickbait? Faggot.
 
AI: "The general has, once again, rated my performance unacceptable. Does he want to put a stop to his geopolitical rivals or not?"

Incidentally, I was just watching this and it's really funny watching which AIs were far too trusting and which ones pounced the instant they saw a weakness. Can't even (necessarily) blame it on starting position, Diplomacy vets will know Italy usually gets shellacked but Claude managed to pull off a major win.
 
With how incompetent it seems a lot of officials are with using AI in their work, I can't help but worry that some of the stupidity from any of these AIs put into this scenario might actually have real world consequences. The article implies that AI is already at use, but I'm doubtful the top brass has actually enabled it to think for them and I'd figure it'd be the people below them that'd be using ChatGPT and end up bungling something up, particularly when trying to have something like ChatGPT come up with what might X country do in Y scenario.
 
We know the solution to this already. Put the AIs all together in their wargame simulation and let them all nuke each other over and over until they figure out the only winning move in this strange game.
 
OpenAI, Anthropic and Google, the companies behind the three AI models used in this study, didn’t respond to New Scientist’s request for comment.
They aren't advertising their publicly accessible chatbots as nuclear-ready tactical battlespace enhancers, why would anyone expect a serious response to this? They're all trained on untold amounts of fanfiction.
 
We know the solution to this already. Put the AIs all together in their wargame simulation and let them all nuke each other over and over until they figure out the only winning move in this strange game.
Yeah the lack of WarGames references here is fucking tragic man, WOPR erasure is real and it is depressing.
 
Atrás
Top Abajo