IDEAS home Printed from https://ideas.repec.org/p/arx/papers/2606.27845.html

LLM Agents as Static Level-k Players in Behavioural Games

Author

Listed:
  • Po Han Teo

Abstract

Large Language Models (LLMs) are increasingly used as stand-ins in behavioural games. These stand-ins rely on the assumption that the LLM's distribution of choices meaningfully matches how humans play the same game. This study tests that assumption through two games. The first is a p-beauty contest, and the second one is a public goods game. The study first investigates five local-model settings within the same model family. These settings are varied together in a 360-cell factorial, which balances temperature, scale (0.5-32B), quantisation, instruct vs base, and framing. Each cell's distribution is then compared against whole choice distributions in published human data. Each deployment setting, except for quantisation, governs a different aspect of fidelity. Mechanically, while the dispersion of human players can be somewhat recovered through deployment settings, the strategic process behind it cannot. Through the lens of the level-k cognitive theory, we find that LLMs act as static, category-retrieved level-k players, where k is set by the model scale. The models also do not run within-game belief-updating or backward induction throughout multiple-round horizon settings. While human contributions decayed in the public goods game, LLMs stayed flat or rose at every scale. When the horizon test was administered, LLMs were more cooperative under an indefinite horizon compared to a finite one. However, LLMs ignore their relative round position, so no last-round defection was displayed. This implies that LLMs retrieved levels relative to the horizon category rather than working out iteratively from the specific game setting.

Suggested Citation

  • Po Han Teo, 2026. "LLM Agents as Static Level-k Players in Behavioural Games," Papers 2606.27845, arXiv.org.
  • Handle: RePEc:arx:papers:2606.27845
    as

    Download full text from publisher

    File URL: https://arxiv.org/pdf/2606.27845
    File Function: Latest version
    Download Restriction: no
    ---><---

    References listed on IDEAS

    as
    1. Colin F. Camerer & Teck-Hua Ho & Juin-Kuan Chong, 2004. "A Cognitive Hierarchy Model of Games," The Quarterly Journal of Economics, President and Fellows of Harvard College, vol. 119(3), pages 861-898.
    2. Antoni Bosch-Domènech & José G. Montalvo & Rosemarie Nagel & Albert Satorra, 2002. "One, Two, (Three), Infinity, ...: Newspaper and Lab Beauty-Contest Experiments," American Economic Review, American Economic Association, vol. 92(5), pages 1687-1701, December.
    3. Jennifer Zelmer, 2003. "Linear Public Goods Experiments: A Meta-Analysis," Experimental Economics, Springer;Economic Science Association, vol. 6(3), pages 299-310, November.
    4. McKelvey Richard D. & Palfrey Thomas R., 1995. "Quantal Response Equilibria for Normal Form Games," Games and Economic Behavior, Elsevier, vol. 10(1), pages 6-38, July.
    5. Alekseenko, Iuliia & Dagaev, Dmitry & Paklina, Sofiia & Parshakov, Petr, 2025. "Strategizing with AI: Insights from a beauty contest experiment," Journal of Economic Behavior & Organization, Elsevier, vol. 240(C).
    6. Ho, Teck-Hua & Camerer, Colin & Weigelt, Keith, 1998. "Iterated Dominance and Iterated Best Response in Experimental "p-Beauty Contests."," American Economic Review, American Economic Association, vol. 88(4), pages 947-969, September.
    7. Gary Charness & Matthew Rabin, 2002. "Understanding Social Preferences with Simple Tests," The Quarterly Journal of Economics, President and Fellows of Harvard College, vol. 117(3), pages 817-869.
    8. Daniel Kahneman & Jack L. Knetsch & Richard H. Thaler, 1991. "Anomalies: The Endowment Effect, Loss Aversion, and Status Quo Bias," Journal of Economic Perspectives, American Economic Association, vol. 5(1), pages 193-206, Winter.
    9. Nagel, Rosemarie, 1995. "Unraveling in Guessing Games: An Experimental Study," American Economic Review, American Economic Association, vol. 85(5), pages 1313-1326, December.
    10. Ernst Fehr & Simon Gächter, 2002. "Altruistic punishment in humans," Nature, Nature, vol. 415(6868), pages 137-140, January.
    11. John J. Horton & Apostolos Filippas & Benjamin S. Manning, 2023. "Large Language Models as Simulated Economic Agents: What Can We Learn from Homo Silicus?," NBER Working Papers 31122, National Bureau of Economic Research, Inc.
    12. Elif Akata & Lion Schulz & Julian Coda-Forno & Seong Joon Oh & Matthias Bethge & Eric Schulz, 2025. "Playing repeated games with large language models," Nature Human Behaviour, Nature, vol. 9(7), pages 1380-1390, July.
    13. Qiaozhu Mei & Yutong Xie & Walter Yuan & Matthew O. Jackson, 2024. "A Turing test of whether AI chatbots are behaviorally similar to humans," Proceedings of the National Academy of Sciences, Proceedings of the National Academy of Sciences, vol. 121(9), pages 2313925121-, February.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Hanaki, Nobuyuki & Koriyama, Yukio & Sutan, Angela & Willinger, Marc, 2019. "The strategic environment effect in beauty contest games," Games and Economic Behavior, Elsevier, vol. 113(C), pages 587-610.
    2. Heller, Yuval, 2015. "Three steps ahead," Theoretical Economics, Econometric Society, vol. 10(1), January.
    3. Breitmoser, Yves, 2010. "Hierarchical Reasoning versus Iterated Reasoning in p-Beauty Contest Guessing Games," MPRA Paper 19893, University Library of Munich, Germany.
    4. Burchardi, Konrad B. & Penczynski, Stefan P., 2014. "Out of your mind: Eliciting individual reasoning in one shot games," Games and Economic Behavior, Elsevier, vol. 84(C), pages 39-57.
    5. Dimitris Batzilis & Sonia Jaffe & Steven Levitt & John A. List & Jeffrey Picel, 2019. "Behavior in Strategic Settings: Evidence from a Million Rock-Paper-Scissors Games," Games, MDPI, vol. 10(2), pages 1-34, April.
    6. Despoina Alempaki & Andrew M. Colman & Felix Kölle & Graham Loomes & Briony D. Pulford, 2022. "Investigating the failure to best respond in experimental games," Experimental Economics, Springer;Economic Science Association, vol. 25(2), pages 656-679, April.
    7. Breitmoser, Yves, 2012. "Strategic reasoning in p-beauty contests," Games and Economic Behavior, Elsevier, vol. 75(2), pages 555-569.
    8. Georganas, Sotiris & Healy, Paul J. & Weber, Roberto A., 2015. "On the persistence of strategic sophistication," Journal of Economic Theory, Elsevier, vol. 159(PA), pages 369-400.
    9. Nagel, Rosemarie & Bühren, Christoph & Frank, Björn, 2017. "Inspired and inspiring: Hervé Moulin and the discovery of the beauty contest game," Mathematical Social Sciences, Elsevier, vol. 90(C), pages 191-207.
    10. Benito Arruñada & Marco Casari & Francesca Pancotto, 2012. "Are self-regarding subjects more rational?," Economics Working Papers 1306, Department of Economics and Business, Universitat Pompeu Fabra.
    11. Vincent P. Crawford & Nagore Iriberri, 2004. "Fatal Attraction: Focality, Naivete, and Sophistication in Experimental Hide-and-Seek Games," Levine's Bibliography 122247000000000316, UCLA Department of Economics.
    12. Haruvy, Ernan & Stahl, Dale O., 2007. "Equilibrium selection and bounded rationality in symmetric normal-form games," Journal of Economic Behavior & Organization, Elsevier, vol. 62(1), pages 98-119, January.
    13. Volker Benndorf & Dorothea Kübler & Hans-Theo Normann, 2017. "Depth of Reasoning and Information Revelation: An Experiment on the Distribution of k-Levels," International Game Theory Review (IGTR), World Scientific Publishing Co. Pte. Ltd., vol. 19(04), pages 1-18, December.
    14. Allain, Marie-Laure & Chambolle, Claire & Rey, Patrick & Teyssier, Sabrina, 2021. "Vertical integration as a source of hold-up: An experiment," European Economic Review, Elsevier, vol. 137(C).
    15. Shapiro, Dmitry & Shi, Xianwen & Zillante, Artie, 2014. "Level-k reasoning in a generalized beauty contest," Games and Economic Behavior, Elsevier, vol. 86(C), pages 308-329.
    16. Miguel A Costa-Gomes & Vincent P Crawford & Nagore Iriberri, 2008. "Comparing Models of Strategic Thinking in Van Huyck, Battalio, and Beil’s Coordination Games," Levine's Working Paper Archive 122247000000002346, David K. Levine.
    17. Bayer, Ralph C. & Renou, Ludovic, 2016. "Logical omniscience at the laboratory," Journal of Behavioral and Experimental Economics (formerly The Journal of Socio-Economics), Elsevier, vol. 64(C), pages 41-49.
    18. Seel, Christian & Tsakas, Elias, 2017. "Rationalizability and Nash equilibria in guessing games," Games and Economic Behavior, Elsevier, vol. 106(C), pages 75-88.
    19. Vincent P. Crawford & Nagore Iriberri, 2007. "Level-k Auctions: Can a Nonequilibrium Model of Strategic Thinking Explain the Winner's Curse and Overbidding in Private-Value Auctions?," Econometrica, Econometric Society, vol. 75(6), pages 1721-1770, November.
    20. Benndorf, Volker & Kübler, Dorothea & Normann, Hans-Theo, 2015. "Privacy concerns, voluntary disclosure of information, and unraveling: An experiment," European Economic Review, Elsevier, vol. 75(C), pages 43-59.

    More about this item

    NEP fields

    This paper has been announced in the following NEP Reports:

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:arx:papers:2606.27845. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: arXiv administrators (email available below). General contact details of provider: https://arxiv.org/ .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.