Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing

Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing

Author

Listed:

Nunzio Lor`e
Babak Heydari

Abstract

This paper investigates the strategic decision-making capabilities of three Large Language Models (LLMs): GPT-3.5, GPT-4, and LLaMa-2, within the framework of game theory. Utilizing four canonical two-player games -- Prisoner's Dilemma, Stag Hunt, Snowdrift, and Prisoner's Delight -- we explore how these models navigate social dilemmas, situations where players can either cooperate for a collective benefit or defect for individual gain. Crucially, we extend our analysis to examine the role of contextual framing, such as diplomatic relations or casual friendships, in shaping the models' decisions. Our findings reveal a complex landscape: while GPT-3.5 is highly sensitive to contextual framing, it shows limited ability to engage in abstract strategic reasoning. Both GPT-4 and LLaMa-2 adjust their strategies based on game structure and context, but LLaMa-2 exhibits a more nuanced understanding of the games' underlying mechanics. These results highlight the current limitations and varied proficiencies of LLMs in strategic decision-making, cautioning against their unqualified use in tasks requiring complex strategic reasoning.

Suggested Citation

Nunzio Lor`e & Babak Heydari, 2023. "Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing," Papers 2309.05898, arXiv.org.

Handle: RePEc:arx:papers:2309.05898

Download full text from publisher

References listed on IDEAS

Yiting Chen & Tracy Xiao Liu & You Shan & Songfa Zhong, 2023. "The emergence of economic rationality of GPT," Proceedings of the National Academy of Sciences, Proceedings of the National Academy of Sciences, vol. 120(51), pages 2316205120-, December.
- Yiting Chen & Tracy Xiao Liu & You Shan & Songfa Zhong, 2023. "The Emergence of Economic Rationality of GPT," Papers 2305.12763, arXiv.org, revised Nov 2023.
Fulin Guo, 2023. "GPT in Game Theory Experiments," Papers 2305.05516, arXiv.org, revised Dec 2023.
John J. Horton & Apostolos Filippas & Benjamin S. Manning, 2023. "Large Language Models as Simulated Economic Agents: What Can We Learn from Homo Silicus?," Papers 2301.07543, arXiv.org, revised Feb 2026.
Joseph N. Luchman, 2021. "Determining relative importance in Stata using dominance analysis: domin and domme," Stata Journal, StataCorp LLC, vol. 21(2), pages 510-538, June.
Steve Phelps & Yvan I. Russell, 2023. "The Machine Psychology of Cooperation: Can GPT models operationalise prompts for altruism, cooperation, competitiveness and selfishness in economic games?," Papers 2305.07970, arXiv.org, revised Jun 2024.

Full references (including those not matched with items on IDEAS)

Citations

Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.

Cited by:

Cagno, Daniela Di & Lin, Lihui, 2025. "How do individuals interact with an AI advisor in strategic reasoning? An experimental study in beauty contest," Journal of Economic Behavior & Organization, Elsevier, vol. 237(C).
Ayato Kitadai & Sinndy Dayana Rico Lugo & Yudai Tsurusaki & Yusuke Fukasawa & Nariaki Nishino, 2024. "Can AI with High Reasoning Ability Replicate Human-like Decision Making in Economic Experiments?," Papers 2406.11426, arXiv.org.
Chuanhao Li & Runhan Yang & Tiankai Li & Milad Bafarassat & Kourosh Sharifi & Dirk Bergemann & Zhuoran Yang, 2024. "STRIDE: A Tool-Assisted LLM Agent Framework for Strategic and Interactive Decision-Making," Cowles Foundation Discussion Papers 2393, Cowles Foundation for Research in Economics, Yale University.
Zarifhonarvar, Ali, 2026. "Generating inflation expectations with large language models," Journal of Monetary Economics, Elsevier, vol. 157(C).
Eléonore Dodivers & Ismaël Rafaï, 2025. "Uncovering the Fairness of AI: Exploring Focal Point, Inequality Aversion, and Altruism in ChatGPT's Dictator Game Decisions," GREDEG Working Papers 2025-09, Groupe de REcherche en Droit, Economie, Gestion (GREDEG CNRS), Université Côte d'Azur, France.
Jin Han & Balaraju Battu & Ivan Romić & Talal Rahwan & Petter Holme, 2025. "Static network structure cannot stabilize cooperation among large language model agents," PLOS ONE, Public Library of Science, vol. 20(5), pages 1-16, May.

Most related items

These are the items that most often cite the same works as this one and are cited by the same works as this one.

Kirshner, Samuel N., 2024. "GPT and CLT: The impact of ChatGPT's level of abstraction on consumer recommendations," Journal of Retailing and Consumer Services, Elsevier, vol. 76(C).
Christoph Engel & Max R. P. Grossmann & Axel Ockenfels, 2023. "Integrating machine behavior into human subject experiments: A user-friendly toolkit and illustrations," Discussion Paper Series of the Max Planck Institute for Behavioral Economics 2024_01, Max Planck Institute for Behavioral Economics.
- Christoph Engel & Max R. P. Grossmann & Axel Ockenfels, 2024. "Integrating Machine Behavior into Human Subject Experiments: A User-Friendly Toolkit and Illustrations," ECONtribute Discussion Papers Series 302, University of Bonn and University of Cologne, Germany.
Bauer, Kevin & Liebich, Lena & Hinz, Oliver & Kosfeld, Michael, 2023. "Decoding GPT's hidden "rationality" of cooperation," SAFE Working Paper Series 401, Leibniz Institute for Financial Research SAFE.
Philip Brookins & Jason DeBacker, 2024. "Playing games with GPT: What can we learn about a large language model from canonical strategic games?," Economics Bulletin, AccessEcon, vol. 44(1), pages 25-37.
Jiafu An & Difang Huang & Chen Lin & Mingzhu Tai, 2024. "Measuring Gender and Racial Biases in Large Language Models," Papers 2403.15281, arXiv.org.
Jingru Jia & Zehua Yuan & Junhao Pan & Paul E. McNamara & Deming Chen, 2024. "Decision-Making Behavior Evaluation Framework for LLMs under Uncertain Context," Papers 2406.05972, arXiv.org, revised Nov 2024.
Thomas R. Cook & Sophia Kazinnik & Zach Modig & Nathan M. Palmer, 2025. "What Do LLMs Want?," Research Working Paper RWP 25-19, Federal Reserve Bank of Kansas City.
- Thomas R. Cook & Sophia Kazinnik & Zach Modig & Nathan M. Palmer, 2026. "What Do LLMs Want?," Finance and Economics Discussion Series 2026-006, Board of Governors of the Federal Reserve System (U.S.).
Cagno, Daniela Di & Lin, Lihui, 2025. "How do individuals interact with an AI advisor in strategic reasoning? An experimental study in beauty contest," Journal of Economic Behavior & Organization, Elsevier, vol. 237(C).
Ennio Bilancini & Leonardo Boncinelli & Eugenio Vicario, 2024. "AI-powered Chatbots: Effective Communication Styles for Sustainable Development Goals," Papers 2407.01057, arXiv.org.
Yingnan Yan & Tianming Liu & Yafeng Yin, 2025. "Valuing Time in Silicon: Can Large Language Models Replicate Human Value of Travel Time," Papers 2507.22244, arXiv.org, revised Dec 2025.
Ali Goli & Amandeep Singh, 2024. "Frontiers: Can Large Language Models Capture Human Preferences?," Marketing Science, INFORMS, vol. 43(4), pages 709-722, July.
Iñaki Aldasoro & Ajit Desai, 2025. "Money Talks: AI Agents for Cash Management in Payment Systems," Staff Working Papers 25-35, Bank of Canada.
Vítor Castro & Rodrigo Martins, 2024. "Lockdowns, vaccines, and the economy: How economic perceptions were shaped during the COVID‐19 pandemic†," Scottish Journal of Political Economy, Scottish Economic Society, vol. 71(3), pages 439-456, July.
Kevin Leyton-Brown & Paul Milgrom & Neil Newman & Ilya Segal, 2024. "Artificial Intelligence and Market Design: Lessons Learned from Radio Spectrum Reallocation," NBER Chapters, in: New Directions in Market Design, pages 119-151, National Bureau of Economic Research, Inc.
Pawe{l} Niszczota & Tomasz Grzegorczyk & Alexander Pastukhov, 2025. "People Are Highly Cooperative with Large Language Models, Especially When Communication Is Possible or Following Human Interaction," Papers 2507.18639, arXiv.org.
C. Monica Capra & Thomas J. Kniesner, 2025. "Daniel Kahneman’s underappreciated last published paper: Empirical implications for benefit-cost analysis and a chat session discussion with bots," Journal of Risk and Uncertainty, Springer, vol. 71(1), pages 29-51, August.
- Capra, C. Monica & Kniesner, Thomas J., 2025. "Daniel Kahneman’s Underappreciated Last Published Paper: Empirical Implications for Benefit-Cost Analysis and a Chat Session Discussion with Bots," IZA Discussion Papers 17841, IZA Network @ LISER.
Jolene Tan, 2023. "Perceptions towards pronatalist policies in Singapore," Journal of Population Research, Springer, vol. 40(3), pages 1-27, September.
Turan G. Bali & Luca Del Viva & Menatalla El Hefnawy & Lenos Trigeorgis, 2024. "Value Uncertainty," Management Science, INFORMS, vol. 70(7), pages 4548-4563, July.
Jean-Victor Alipour, 2026. "Performance Pay in the Hybrid Work Economy," CESifo Working Paper Series 12680, CESifo.
Hongshen Sun & Juanjuan Zhang, 2025. "From Model Choice to Model Belief: Establishing a New Measure for LLM-Based Research," Papers 2512.23184, arXiv.org.

More about this item

NEP fields

This paper has been announced in the following NEP Reports:

NEP-AIN-2023-10-16 (Artificial Intelligence)
NEP-EXP-2023-10-16 (Experimental Economics)
NEP-GER-2023-10-16 (German Papers)
NEP-GTH-2023-10-16 (Game Theory)

Statistics

Access and download statistics

Corrections

All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:arx:papers:2309.05898. See general information about how to correct material in RePEc.

If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: arXiv administrators (email available below). General contact details of provider: https://arxiv.org/ .

Please note that corrections may take a couple of weeks to filter through the various RePEc services.

IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.

Browse Econ Literature

More features

Strategic Behavior of Large Language Models: Game Structure vs. Contextual Framing

Author

Abstract

Suggested Citation

Download full text from publisher

References listed on IDEAS

Citations

Most related items

More about this item

NEP fields

Statistics

Corrections

More services and features

MyIDEAS

Author registration

Rankings

RePEc Genealogy

RePEc Biblio

MPRA

New papers by email

EconAcademics

Plagiarism

About RePEc

RePEc home

Blog

Help/FAQ

RePEc team

Participating archives

Privacy statement

Help us

Corrections

Volunteers

Get papers listed

Open a RePEc archive

Get RePEc data