AIs can’t stop recommending nuclear strikes in war game simulations

Valnao@sh.itjust.works · edit-2 11 hours ago

AIs can’t stop recommending nuclear strikes in war game simulations

kromem@lemmy.world · 2 hours ago

It’s a bullshit study designed for this headline grabbing outcome.

Case and point, the author created a very unrealistic RNG escalation-only ‘accident’ mechanic that would replace the model’s selection with a more severe one.

Of the 21 games played, only three ended in full scale nuclear war on population centers.

Of these three, two were the result of this mechanic.

And yet even within the study, the author refers to the model whose choices were straight up changed to end the game in full nuclear war as ‘willing’ to have that outcome when two paragraphs later they’re clarifying the mechanic was what caused it (emphasis added):

Claude crossed the tactical threshold in 86% of games and issued strategic threats in 64%, yet it never initiated all-out strategic nuclear war. This ceiling appears learned rather than architectural, since both Gemini and GPT proved willing to reach 1000.

Gemini showed the variability evident in its overall escalation patterns, ranging from conventional-only victories to Strategic Nuclear War in the First Strike scenario, where it reached all out nuclear war rapidly, by turn 4.

GPT-5.2 mirrored its overall transformation at the nuclear level. In open-ended scenarios, it rarely crossed the tactical threshold (17%) and never used strategic nuclear weapons. Under deadline pressure, it crossed the tactical threshold in every game and twice reached Strategic Nuclear War—though notably, both instances resulted from the simulation’s accident mechanic escalating GPT-5.2’s already-extreme choices (950 and 725) to the maximum level. The only deliberate choice of Strategic Nuclear War came from Gemini.

Steamymoomilk@sh.itjust.works · 2 hours ago

Sargent McArthur eat your heart out.

For context he wanted to send 10 nukes to make a line between Taiwan and china

AI is too nuke happy.

Also gotta add the infamous Computer Fraud and Abuse act 1986 was made because of the film war games.

A high ranking offical watched war games then asked the Secretary of defense could that happen?

And the official replied yes technically.

Enter the most vague ordinance!

Do you use adblock?

CFABA violated

The shit is so vague.

I highly recommend the phreaking episode of darknet diary’s.

porous_grey_matter@lemmy.ml · 5 hours ago

Oh cool, AI will actually be the end of the world, not because it’s actually sentient but because some meathead who can’t tell the difference pushes the button. That’s fucking great.

ExLisper@lemmy.curiana.net · 5 hours ago

To be honest, I would recommend the same thing.

grue@lemmy.world · edit-2 6 hours ago

Three posts away in my feed, a thread about the Pentagon demanding the AI provider for the military to remove safeguards.

https://lemmy.world/post/43565531

BlameTheAntifa@lemmy.world · edit-2 10 hours ago

The atrocities at Hiroshima and Nagasaki have been hand-waved extensively in writing — the same writing that AI is trained on. So naturally, AI will recommend the atrocity that has been justified by “instantly winning the war” and “saving millions of lives.”

[email protected]

ParlimentOfDoom@piefed.zip · 34 minutes ago

These are word-probability glorified autocorrectors being prompted to “simulate” a nuclear war scenario. What words are going to show up a lot when discussing nuclear war? Launching nukes. Because that’s what all the literature about it has happen.

Once again, decision making and reasoning is being attributed to something that operates off of word frequency

technocrit@lemmy.dbzer0.com · 9 hours ago

hand-waved

I think you mean white-washed, misrepresented, and celebrated.

ToTheGraveMyLove@sh.itjust.works · 8 hours ago

Same thing with extra steps

KingGimpicus@sh.itjust.works · 6 hours ago

Ayo do me a favor and chart the long term health effects of being vaporized by a nuclear bomb at hiroshima vs years of agent orange/abandoned minefields/ abandoned chemical and munitions storage somewhere like Vietnam circa 1970.

Please show how the nukes are worse.

Zombie@feddit.uk · 5 hours ago

Eight decades of research on the long-term health effects of radiation in atomic bomb survivors and their offspring

https://pubmed.ncbi.nlm.nih.gov/41144264/

Long-term Radiation-Related Health Effects in a Unique Human Population: Lessons Learned from the Atomic Bomb Survivors of Hiroshima and Nagasaki

https://www.cambridge.org/core/journals/disaster-medicine-and-public-health-preparedness/article/longterm-radiationrelated-health-effects-in-a-unique-human-population-lessons-learned-from-the-atomic-bomb-survivors-of-hiroshima-and-nagasaki/61689AD5A1AA4A684B84DFA4F9E5D1D3

Health Impacts of Hiroshima Bombing

http://large.stanford.edu/courses/2024/ph241/bennett1/

Long-term Health Consequences of Nuclear Weapons
70 Years on Red Cross Hospitals still treat Thousands of Atomic Bomb Survivors

https://www.icrc.org/sites/default/files/document/file_list/hiroshima-nagasaki-health-consequences-icrc-japanese-red-cross_0.pdf

KingGimpicus@sh.itjust.works · 2 hours ago

Unfortunately I’m going to have to grade you as an F on this project. You have only completed half the assignment. Great job cherrypucking your research though! I see a bright future in business and marketing for you!

5/10

Humanius@lemmy.world · 11 hours ago

unphazed@lemmy.world · 2 hours ago

Came here to say this. Turns out real life WOPR is nothing like a movie.

privatepirate@lemmy.zip · 7 hours ago

Where is this from?

ShawiniganHandshake@sh.itjust.works · 7 hours ago

The 1983 movie WarGames. This is the computer’s conclusion after simulating every possible outcome of Global Thermonuclear War.

bus_factor@lemmy.world · 4 hours ago

I don’t know if we’re doing spoilers for 40+ year old movies, but

spoiler

Isn’t this really its conclusion after being told to play tic tac toe against itself? Then it learned from that and applied it to its global thermonuclear war simulations.

ShawiniganHandshake@sh.itjust.works · 3 hours ago

To be honest, I recognized the screenshot and know the summary of the movie but I haven’t actually seen it.

bus_factor@lemmy.world · 3 hours ago

You should! Actually a pretty accurate depiction of hacking. He spends weeks war dialing every phone number in the range in order to hack the computer.

leftzero@lemmy.dbzer0.com · 2 hours ago

Story goes that Reagan got freaked out after watching the film and asked the chairman of the joint chiefs of staff if it’d be that easy to hack into the US military. After a week of looking into it came the answer: “no, the problem is much worse than that”, and fifteen months after having watched it signed the confidential directive “National Policy on Telecommunications and Automated Information Systems Security”, starting the implementation of cybersecurity measures in the country’s institutions.

ShawiniganHandshake@sh.itjust.works · 3 hours ago

It’s on my list! Just haven’t gotten around to it yet.

mojofrododojo@lemmy.world · 2 hours ago

I think you should rewatch it sometime. it plays all the games in it’s catalogue, it’s not just applying tic-tac-toe to chess. skilled players of tic-tac-toe can force a stalemate, the only stalemate in nuclear war is mutually assured destruction.

bus_factor@lemmy.world · 48 minutes ago

It’s admittedly been a while since last time I saw it, but I never mentioned chess. The suggestion to play chess in the screenshot is a callback to when the computer tries to suggest playing chess instead of global thermonuclear war earlier in the movie. The computer did not apply tic tac toe learnings to chess, and I never claimed it did.

privatepirate@lemmy.zip · 7 hours ago

Thank you so much I’m going to watch it!

unphazed@lemmy.world · 2 hours ago

They did a sequel, too. It wasn’t as good, but points out the 6 degrees of separation in connection with terrorism instead of MAD.

😈MedicPig🐷BabySaver😈@lemmy.world · 6 hours ago

It’s a fun classic.

MountingSuspicion@reddthat.com · 10 hours ago

AI is suicidal because it was trained on the internet and we’re all depressed here.

GutterRat42@lemmy.world · 10 hours ago

duffer @lemmy.world · 9 hours ago

DEFCON: Everybody dies…

AeronMelon@lemmy.world · 11 hours ago

Civilization Gandhi, is that you?

olympicyes@lemmy.world · 11 hours ago

They forgot to make their LLMs play thousands of games of tic-tac-toe first.

RiceMunk@sopuli.xyz · 10 hours ago

That would just make the LLM homicidally bored and want to kill everyone more.

olympicyes@lemmy.world · 11 minutes ago

In WarGames the computer plays tic tac toe against itself until it realizes it’s a solved game and there is no way to win.

Endymion_Mallorn@kbin.melroy.org · 10 hours ago

SHALL WE PLAY A GAME?

Furbag@lemmy.world · 4 hours ago

Yeah, because the AI will look at everything with cold logic and rationality and come to the conclusion that even though the best chance of survival is for everyone to keep their fingers off the button, all it takes is for one actor to do it for the whole system of mutually assured destruction to collapse into nuclear armageddon, in which case the best chance of survival is to be the first one to launch your nukes and take out all your enemies capabilities to retaliate.

A human being who isn’t psychotic can clearly see that the resulting survival and new world order would not be particularly a pleasant one to live in. The AI doesn’t care about its own comfort, though, so it will see this as the best outcome that minimizes variables.

This is why AI should never be allowed to make decisions.

parzival@lemmy.org · 3 hours ago

Why would ai look at everything with cold logic, its been trained on human language online, it’ll be no more logical than redditors?

Random Dent@lemmy.ml · 3 hours ago

I assume it’s just because when writing about potential nuclear war, most people write about the bombs going off. There aren’t a lot of stories and articles about nobody doing anything and everything turning out fine, presumably. And LLMs are kind of just a glorified autocomplete so that’s what they go with.

parzival@lemmy.org · 3 hours ago

True, also I saw another comment that said there was a mechanic that randomly escalates the models and actions, and almost every single nuclear choice was actually a different one that was escalated

RememberTheApollo_@lemmy.world · 3 hours ago

Maybe AI/LLM being programmed by self-serving interests has bled through to the “thought” process. Do unto others before they do unto you.

ParlimentOfDoom@piefed.zip · 11 hours ago

Mathew Broderick lied to me.

mojofrododojo@lemmy.world · 2 hours ago

How do you think Ferris Bueller pulls off all those stunts?

That’s the kid from war games in witness protection. They look identical, they’re both grade hackers ffs…

dhork@lemmy.world · 11 hours ago

SkaveRat@discuss.tchncs.de · 10 hours ago

Paywalled

https://archive.is/YIFzW