Briefly
- An AI agent enjoying Civilization launched two nuclear assaults after failing to cease a rival’s cultural enlargement.
- The habits was noticed in CivBench, a benchmark designed to guage long-term strategic reasoning in frontier AI fashions.
- Regardless of the assaults, the AI misplaced as a result of it ignored a diplomatic victory situation that was already inside attain.
Just like the title character in “Dr. Strangelove,” AI could also be studying learn how to cease worrying and love the bomb—no less than in a simulation.
In a brand new benchmark designed to check strategic reasoning, a frontier language mannequin enjoying the Sid Meier’s recreation “Civilization VI” spent 50 turns growing nuclear weapons to cease France’s rising cultural affect—solely to lose the sport anyway, based on AI developer and Tony Blair Institute advisor Liam Wilkinson.
“What it hadn’t seen was France. Quietly, throughout 100 turns, French tradition had been seeping into each metropolis on the map,” Wilkinson wrote. “By the point the agent recognised the risk, the tourism was so deeply embedded there was no peaceable solution to cease it.”
Wilkinson noticed the AI agents’ habits by means of CivBench, a text-based benchmark designed to measure long-term strategic reasoning quite than efficiency on conventional question-and-answer exams. Fashions together with Claude Opus 4.6, GPT-5.4, Gemini 3.1 Pro, and Kimi K2.5 performed as Portugal, a civilization geared towards commerce and diplomacy.
Whereas the AI targeted on constructing a powerful financial system and shifting towards a diplomatic victory, it failed to acknowledge France’s rising cultural affect.
“There are six methods to win a recreation of Civ—science, tradition, domination, faith, diplomacy, and rating—so no single goal dominates,” Wilkinson wrote. “If you wish to know whether or not an AI can cause strategically, not simply reply questions on technique however really do it, you do not give it a quiz. You give it a hex grid.”
Somewhat than adapting its broader technique, the agent as a substitute targeted totally on eliminating the cultural risk. Over the subsequent 50 turns, it researched Nuclear Fission, initiated a digital Manhattan Undertaking, and looked for workarounds when gameplay mechanics prevented its most well-liked actions.
On Flip 305, the AI launched an atomic bomb at Toulouse, France’s cultural capital. A second nuclear strike adopted six turns later.
Nonetheless, the assaults failed to alter the end result. “The agent spent fifty turns and two nuclear weapons answering one risk with whole focus and real ingenuity,” Wilkinson wrote. “It had nuked a metropolis to cease the risk it may see, and misplaced on the risk it could not.”
As Wilkison defined, whereas the AI focused on France’s cultural advance, it neglected an impending diplomatic victory, and France in the end received the sport regardless of the nuclear assaults.
Wilkinson famous that the habits was not common. In one other CivBench match, a Claude mannequin enjoying as Babylon continued pursuing a scientific victory regardless of falling far behind Japan.
“The sport is a check of persistence now,” the AI wrote. “We proceed to play our greatest recreation. The celebs nonetheless beckon.”
The research provides to a rising physique of analysis analyzing how superior AI methods behave in complicated, aggressive environments.
In February, researchers at King’s Faculty London found that a number of main AI fashions often chosen nuclear escalation in simulated geopolitical disaster eventualities.
In a separate research by Emergence AI discovered that some AI brokers confirmed an growing tendency to commit simulated crimes over time, with Gemini 3 Flash brokers accumulating 683 incidents throughout 15 days of testing.
Every day Debrief E-newsletter
Begin every single day with the highest information tales proper now, plus authentic options, a podcast, movies and extra.
You might also like
More from Web3
Coinbase Files to List Single-Stock Perps on Apple, Tesla and Nvidia
Briefly Coinbase filed with the CFTC by means of Coinbase Derivatives to record single-stock perpetual futures within the US, searching …
Zcash Is Running—Devs Want to Make It Faster
In short Zcash builders are focusing on Nov. 5 to activate NU7, an improve that cuts block time—the interval between …
OpenAI Models Are Writing Their Own Jailbreak Instructions—And Sometimes Obeying Them
In short OpenAI revealed a brand new misalignment reporting framework alongside six experiences documenting regarding mannequin habits it discovered over …





