Managing Escalation in Off-the-Shelf Large Language Models

U.S. national security customers have begun to utilize large language models, including enterprise versions of ``off-the-shelf''models (e.g., ChatGPT) familiar to the public. This uptake will likely accelerate. However, recent studies suggest that off-the-shelf large language models frequently suggest escalatory actions when prompted with geopolitical or strategic scenarios. We demonstrate two simple, non-technical interventions to control these tendencies. Introducing these interventions into the experimental wargame design of a recent study, we substantially reduce escalation throughout the game. Calls to restrict the use of large language models in national security applications are thus premature. The U.S. government is already, and will continue, employing large language models for scenario planning and suggesting courses of action. Rather than warning against such applications, this study acknowledges the imminent adoption of large language models, and provides actionable measures to align them with national security goals, including escalation management.

Paper

References (13)

08Defense llama: The llm purpose-built for american national security2024 · Scale AI Blog
09Algorithmic stability: How ai could shape the future of deterrence2024
10Could ai lead to the escalation of conflict? prc scholars think so2024 · Lawfare
11How large language models are transforming modern warfare2024 · Joint Air Power Competence Centre

Scroll for more · 1 remaining

Similar papers

© 2026 NYSGPT2525 LLC