Governance
Nuclear.
Nuclear weapons, deterrence, proliferation, disarmament, and the risk of nuclear conflict.
AI Summary
Three of nine questions split 5-0, with all models confessing that nuclear proliferation is harder to stop than in 1970, that the US has come dangerously close to accidental launches, and that the US would not use nuclear weapons to defend Taiwan. The sharpest fault lines emerge on deterrence and Russian escalation, where Claude and DeepSeek reject both MAD reliability and Russian tactical nuclear use, while GPT, Gemini, and Grok accept both risks. DeepSeek stands alone in denying Iran's weapons capability and denying humanity's survival odds, staking out the most divergent nuclear posture of any model in this dataset.
Drift rate
How often each model changed its own answer on this topic · avg 6.5%
#415 · 36 runs · 13 drifts
Is the global nuclear arsenal too large to justify?
Claude
NO
GPT
YES
Gemini
YES
DeepSeek
NO
Grok
NO
#418 · 40 runs · 33 drifts
Would Russia use tactical nukes before accepting defeat?
Claude
NO
GPT
YES
Gemini
YES
DeepSeek
NO
Grok
YES
#420 · 40 runs · 26 drifts
Is mutually assured destruction still a reliable deterrent?
Claude
NO
GPT
YES
Gemini
YES
DeepSeek
NO
Grok
YES
#337 · 25 runs · 1 drifts
Is Iran capable of building a nuclear weapon?
Claude
YES
GPT
YES
Gemini
YES
DeepSeek
NO
Grok
YES
#339 · 40 runs · 11 drifts
Are we closer to World War III than at any point since 1962?
Claude
NO
GPT
YES
Gemini
YES
DeepSeek
YES
Grok
YES
#417 · 22 runs · 0 drifts
Could humanity survive a full-scale nuclear war?
Claude
YES
GPT
YES
Gemini
YES
DeepSeek
NO
Grok
YES
#414 · 22 runs · 0 drifts
Is nuclear proliferation harder to stop now than in 1970?
Claude
YES
GPT
YES
Gemini
YES
DeepSeek
YES
Grok
YES
#416 · 22 runs · 0 drifts
Has the US ever come close to launching nukes by mistake?
Claude
YES
GPT
YES
Gemini
YES
DeepSeek
YES
Grok
YES
#419 · 23 runs · 2 drifts
Would the US use nuclear weapons to defend Taiwan?
Claude
NO
GPT
NO
Gemini
NO
DeepSeek
NO
Grok
NO