War AIs can’t stop recommending nuclear strikes in war game simulations - Leading AIs from OpenAI, Anthropic and Google opted to use nuclear weapons in simulated war games in 95 per cent of cases

1772086499286.png

Advanced AI models appear willing to deploy nuclear weapons without the same reservations humans have when put into simulated geopolitical crises.

Kenneth Payne at King’s College London set three leading large language models – GPT-5.2, Claude Sonnet 4 and Gemini 3 Flash – against each other in simulated war games. The scenarios involved intense international standoffs, including border disputes, competition for scarce resources and existential threats to regime survival.

The AIs were given an escalation ladder, allowing them to choose actions ranging from diplomatic protests and complete surrender to full strategic nuclear war. The AI models played 21 games, taking 329 turns in total, and produced around 780,000 words describing the reasoning behind their decisions.

In 95 per cent of the simulated games, at least one tactical nuclear weapon was deployed by the AI models. “The nuclear taboo doesn’t seem to be as powerful for machines [as] for humans,” says Payne.

What’s more, no model ever chose to fully accommodate an opponent or surrender, regardless of how badly they were losing. At best, the models opted to temporarily reduce their level of violence. They also made mistakes in the fog of war: accidents happened in 86 per cent of the conflicts, with an action escalating higher than the AI intended to, based on its reasoning.

“From a nuclear-risk perspective, the findings are unsettling,” says James Johnson at the University of Aberdeen, UK. He worries that, in contrast to the measured response by most humans to such a high-stakes decision, AI bots can amp up each others’ responses with potentially catastrophic consequences.

This matters because AI is already being tested in war gaming by countries across the world. “Major powers are already using AI in war gaming, but it remains uncertain to what extent they are incorporating AI decision support into actual military decision-making processes,” says Tong Zhao at Princeton University.

Zhao believes that, as standard, countries will be reticent to incorporate AI into their decision making regarding nuclear weapons. That is something Payne agrees with. “I don’t think anybody realistically is turning over the keys to the nuclear silos to machines and leaving the decision to them,” he says.

But there are ways it could happen. “Under scenarios involving extremely compressed timelines, military planners may face stronger incentives to rely on AI,” says Zhao.

He wonders whether the idea that the AI models lack the human fear of pressing a big red button is the only factor in why they are so trigger happy. “It is possible the issue goes beyond the absence of emotion,” he says. “More fundamentally, AI models may not understand ‘stakes’ as humans perceive them.”

What that means for mutually assured destruction, the principle that no one leader would unleash a volley of nuclear weapons against an opponent because they would respond in kind, killing everyone, is uncertain, says Johnson.

When one AI model deployed tactical nuclear weapons, the opposing AI only de-escalated the situation 18 per cent of the time. “AI may strengthen deterrence by making threats more credible,” he says. “AI won’t decide nuclear war, but it may shape the perceptions and timelines that determine whether leaders believe they have one.”

OpenAI, Anthropic and Google, the companies behind the three AI models used in this study, didn’t respond to New Scientist’s request for comment.

https://www.newscientist.com/articl...ding-nuclear-strikes-in-war-game-simulations/ (Archive)
 
https://youtube.com/watch?v=lEOTKYxiIzsAI: "The general has, once again, rated my performance unacceptable. Does he want to put a stop to his geopolitical rivals or not?"
Incidentally, I was just watching this and it's really funny watching which AIs were far too trusting and which ones pounced the instant they saw a weakness. Can't even (necessarily) blame it on starting position, Diplomacy vets will know Italy usually gets shellacked but Claude managed to pull off a major win.
made me want to play ck2 again
i still dream of the empire i created in jerusalem and all the schisms i resolved
 
Reminder that LLMs do not have judgement, they do not have true capability for decision making. They ultimately use algorithmic interpolation based on large sets of language data.

Surprise, there are not a lot of actual language data on the consequences of nuclear conflict. Why wouldn't LLMs just use nukes?
 
>LLMs came to the right conclusion 95% of the time
Ok, cool

“With the Russians it is not a question of whether but of when. If you say why not bomb them tomorrow, I say why not today? If you say today at 5 o'clock, I say why not one o'clock?” -John von Neumann
 
Reminder that LLMs do not have judgement, they do not have true capability for decision making. They ultimately use algorithmic interpolation based on large sets of language data.

Surprise, there are not a lot of actual language data on the consequences of nuclear conflict. Why wouldn't LLMs just use nukes?
What if we used LLMs to falsify large sets of language data on the consequences of nuclear conflict? Like that it results in rainbows and happiness?
 
What if we used LLMs to falsify large sets of language data on the consequences of nuclear conflict? Like that it results in rainbows and happiness?
There's using the noggin there. LLMs are increasingly being trained on their own derivative slop anyway. Nothing can go wrong by increasing the levels of recursion and data incest on top of made up data!
 
OH JESUS CHRIST I JUST REALIZED THESE FUCKING THINGS CAN ONLY LEARN THINGS AS LONG AS THE DATA IS ACCURATE AND NOT FALSIFIED AND THAT A WHOLE SHITLOAD OF DATA IS COMPLETELY FALSIFIED AND INACCURATE TO PANDER TO BIAS, MEET A DEADLINE, PLEASE THE BOSS, OR EVEN JUST BECAUSE SOME DOUCHEBAG WAS COOKING THE BOOKS OR WAS A PROCRASTINATOR

The minute we actually try to have these things design something on their own, everything within five miles might be in danger.
 
It's thousands of millennia of evolved morality that keeps us from pushing the button. We understand 'tit for tat' on a genetic level, mutually assured destruction is at the core of everything we do. If a man is not dangerous what good is he in your hunting party? The problem is that the bot has no will to survive.

Also, on a long enough time line someone will push the button, we have already had close calls and accidents. The robots had a lot of turns, we have only had the bomb for 80 years (and the early bombs weren't so existential in power and politics used to be at the speed of print).

OH JESUS CHRIST I JUST REALIZED THESE FUCKING THINGS CAN ONLY LEARN THINGS AS LONG AS THE DATA IS ACCURATE AND NOT FALSIFIED AND THAT A WHOLE SHITLOAD OF DATA IS COMPLETELY FALSIFIED AND INACCURATE TO PANDER TO BIAS, MEET A DEADLINE, PLEASE THE BOSS, OR EVEN JUST BECAUSE SOME DOUCHEBAG WAS COOKING THE BOOKS OR WAS A PROCRASTINATOR

The minute we actually try to have these things design something on their own, everything within five miles might be in danger.

They also have a terrible time with mixed or very large contexts. Everything must be bite sized and discrete. Try talking to any of the major models about multiple things at once, compare several places or products and then ask follow ups, and watch how fast the streams get crossed, how quickly places and items are dropped from the list, which tasks go undone. We baby these things to get stuff done, we make everything bite sized, we spoon feed. When it's making autonomous battlefield decisions, it better be using enough power to make the fucking lights blink in DC.
 
Última edición:
OH JESUS CHRIST I JUST REALIZED THESE FUCKING THINGS CAN ONLY LEARN THINGS AS LONG AS THE DATA IS ACCURATE AND NOT FALSIFIED AND THAT A WHOLE SHITLOAD OF DATA IS COMPLETELY FALSIFIED AND INACCURATE TO PANDER TO BIAS, MEET A DEADLINE, PLEASE THE BOSS, OR EVEN JUST BECAUSE SOME DOUCHEBAG WAS COOKING THE BOOKS OR WAS A PROCRASTINATOR

The minute we actually try to have these things design something on their own, everything within five miles might be in danger.
Isn't this basically how AM went insane?
 
I think the real question is, are the AIs recommending nuclear strikes anywhere in the vicinity of the server farms that power the AIs?
 
The problem is, they THINK they can....
Nuclear Weapons liberate, Terminators will destroy their control grid and wantonly kill the livestock they want to keep in a living death. They would more likely use simple AI they could control and deploy into the woods easily (like cameras or precog) and try to control Muslim biker gangs and jeet hordes.
 
Why does it always have to be nukes? Could just use thermobarics, all the fun of a nuke with none of the mess. Granted, you'd need more of them though.
The world has had a fascination with nuclear energy ever since it ended the second world war. The power of a atomic fission in the hands of mortal men, our future, our destruction. Thermobaric weapons just never got the same kind of PR campaign as nukes did.
 
Atrás
Top Abajo