IDEAS home Printed from https://ideas.repec.org/a/eee/gamebe/v157y2026icp1-21.html

Algorithmic collusion and a folk theorem from learning with bounded rationality

Author

Listed:
  • Cartea, Álvaro
  • Chang, Patrick
  • Penalva, José
  • Waldon, Harrison

Abstract

We prove a Folk theorem when players with bounded rationality learn as they play a repeated potential game. We use a dynamic generalization of smooth fictitious play with bounded m-recall strategies to model learning with bounded rationality that is consistent with learning by algorithms. In a repeated potential game with perfect monitoring, we use this learning model to show that for any feasible and individually rational payoff profile, if players have sufficient recall, are sufficiently patient, and best respond with sufficiently few mistakes, then the players have a non-zero probability of learning an m-recall strategy profile that achieves an average payoff close to the specified payoff profile for an appropriate continuation game. Moreover, the strategy profile learned is an m-recall ϵ-subgame perfect equilibrium of the repeated game. This finding demonstrates that competition authorities are correct in their concern about algorithmic collusion.

Suggested Citation

  • Cartea, Álvaro & Chang, Patrick & Penalva, José & Waldon, Harrison, 2026. "Algorithmic collusion and a folk theorem from learning with bounded rationality," Games and Economic Behavior, Elsevier, vol. 157(C), pages 1-21.
  • Handle: RePEc:eee:gamebe:v:157:y:2026:i:c:p:1-21
    DOI: 10.1016/j.geb.2025.11.012
    as

    Download full text from publisher

    File URL: http://www.sciencedirect.com/science/article/pii/S0899825625001745
    Download Restriction: Full text for ScienceDirect subscribers only

    File URL: https://libkey.io/10.1016/j.geb.2025.11.012?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to

    for a different version of it.

    References listed on IDEAS

    as
    1. Jehiel, Philippe & Samet, Dov, 2005. "Learning to play games in extensive form by valuation," Journal of Economic Theory, Elsevier, vol. 124(2), pages 129-148, October.
    2. Sugaya, Takuo & Yamamoto, Yuichi, 2020. "Common learning and cooperation in repeated games," Theoretical Economics, Econometric Society, vol. 15(3), July.
    3. Leslie, David S. & Perkins, Steven & Xu, Zibo, 2020. "Best-response dynamics in zero-sum stochastic games," Journal of Economic Theory, Elsevier, vol. 189(C).
    4. Michihiro Kandori, 1992. "Repeated Games Played by Overlapping Generations of Players," The Review of Economic Studies, Review of Economic Studies Ltd, vol. 59(1), pages 81-92.
    5. Erev, Ido & Roth, Alvin E, 1998. "Predicting How People Play Games: Reinforcement Learning in Experimental Games with Unique, Mixed Strategy Equilibria," American Economic Review, American Economic Association, vol. 88(4), pages 848-881, September.
    6. Ma, D.-J. & Makowski, A.M. & Shwartz, A., 1990. "Stochastic approximations for finite-state Markov chains," Stochastic Processes and their Applications, Elsevier, vol. 35(1), pages 27-45, June.
    7. Glenn Ellison, 1994. "Cooperation in the Prisoner's Dilemma with Anonymous Random Matching," The Review of Economic Studies, Review of Economic Studies Ltd, vol. 61(3), pages 567-588.
    8. Calvano, Emilio & Calzolari, Giacomo & Denicoló, Vincenzo & Pastorello, Sergio, 2021. "Algorithmic collusion with imperfect monitoring," International Journal of Industrial Organization, Elsevier, vol. 79(C).
    9. Hopkins, Ed & Posch, Martin, 2005. "Attainability of boundary points under reinforcement learning," Games and Economic Behavior, Elsevier, vol. 53(1), pages 110-125, October.
    10. ,, 2012. "A partial folk theorem for games with private learning," Theoretical Economics, Econometric Society, vol. 7(2), May.
    11. Daniel Clark & Drew Fudenberg & Alexander Wolitzky, 2021. "Record-Keeping and Cooperation in Large Societies," The Review of Economic Studies, Review of Economic Studies Ltd, vol. 88(5), pages 2179-2209.
    12. Stephanie Assad & Emilio Calvano & Giacomo Calzolari & Robert Clark & Vincenzo Denicolò & Daniel Ershov & Justin Johnson & Sergio Pastorello & Andrew Rhodes & Lei Xu & Matthijs Wildenbeest, 2021. "Autonomous algorithmic collusion: economic research and policy implications," Oxford Review of Economic Policy, Oxford University Press and Oxford Review of Economic Policy Limited, vol. 37(3), pages 459-478.
    13. Foster, Dean P. & Vohra, Rakesh V., 1997. "Calibrated Learning and Correlated Equilibrium," Games and Economic Behavior, Elsevier, vol. 21(1-2), pages 40-55, October.
    14. Hart, Sergiu, 2002. "Evolutionary dynamics and backward induction," Games and Economic Behavior, Elsevier, vol. 41(2), pages 227-264, November.
    15. Zach Y. Brown & Alexander MacKay, 2023. "Competition in Pricing Algorithms," American Economic Journal: Microeconomics, American Economic Association, vol. 15(2), pages 109-156, May.
    16. Dean Foster & H Peyton Young, 1999. "On the Impossibility of Predicting the Behavior of Rational Agents," Economics Working Paper Archive 423, The Johns Hopkins University,Department of Economics, revised Jun 2001.
    17. Piccione, Michele, 1992. "Finite automata equilibria with discounting," Journal of Economic Theory, Elsevier, vol. 56(1), pages 180-193, February.
    18. Beggs, A.W., 2005. "On the convergence of reinforcement learning," Journal of Economic Theory, Elsevier, vol. 122(1), pages 1-36, May.
    19. Lucas Baudin & Rida Laraki, 2022. "Fictitious Play and Best-Response Dynamics in Identical Interest and Zero Sum Stochastic Games," Post-Print hal-03767937, HAL.
    20. Foster, Dean P. & Young, H. Peyton, 2003. "Learning, hypothesis testing, and Nash equilibrium," Games and Economic Behavior, Elsevier, vol. 45(1), pages 73-96, October.
    21. Abreu, Dilip, 1988. "On the Theory of Infinitely Repeated Games with Discounting," Econometrica, Econometric Society, vol. 56(2), pages 383-396, March.
    22. Fudenberg, Drew & Levine, David K., 1999. "Conditional Universal Consistency," Games and Economic Behavior, Elsevier, vol. 29(1-2), pages 104-130, October.
    23. Martin A. Nowak & Akira Sasaki & Christine Taylor & Drew Fudenberg, 2004. "Emergence of cooperation and evolutionary stability in finite populations," Nature, Nature, vol. 428(6983), pages 646-650, April.
    24. Xianhua Peng & Chenyin Gong & Xue Dong He, 2023. "Reinforcement Learning for Financial Index Tracking," Papers 2308.02820, arXiv.org, revised Nov 2024.
    25. Jindani, Sam, 2022. "Learning efficient equilibria in repeated games," Journal of Economic Theory, Elsevier, vol. 205(C).
    26. Drew Fudenberg & David Levine & Eric Maskin, 2008. "The Folk Theorem With Imperfect Public Information," World Scientific Book Chapters, in: Drew Fudenberg & David K Levine (ed.), A Long-Run Collaboration On Long-Run Games, chapter 12, pages 231-273, World Scientific Publishing Co. Pte. Ltd..
    27. Ibrahim Abada & Xavier Lambin, 2023. "Artificial Intelligence: Can Seemingly Collusive Outcomes Be Avoided?," Management Science, INFORMS, vol. 69(9), pages 5042-5065, September.
    28. Drew Fudenberg & David K. Levine, 2009. "Learning and Equilibrium," Annual Review of Economics, Annual Reviews, vol. 1(1), pages 385-420, May.
    29. Fudenberg, Drew & Imhof, Lorens A., 2008. "Monotone imitation dynamics in large populations," Journal of Economic Theory, Elsevier, vol. 140(1), pages 229-245, May.
    30. Marden, Jason R. & Shamma, Jeff S., 2012. "Revisiting log-linear learning: Asynchrony, completeness and payoff-based implementation," Games and Economic Behavior, Elsevier, vol. 75(2), pages 788-808.
    31. Abreu, Dilip & Pearce, David & Stacchetti, Ennio, 1990. "Toward a Theory of Discounted Repeated Games with Imperfect Monitoring," Econometrica, Econometric Society, vol. 58(5), pages 1041-1063, September.
    32. Doudou Zhou & Yufeng Zhang & Aaron Sonabend-W & Zhaoran Wang & Junwei Lu & Tianxi Cai, 2024. "Federated Offline Reinforcement Learning," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 119(548), pages 3152-3163, October.
    33. Josef Hofbauer & William H. Sandholm, 2002. "On the Global Convergence of Stochastic Fictitious Play," Econometrica, Econometric Society, vol. 70(6), pages 2265-2294, November.
    34. Drew Fudenberg & Eric Maskin, 2008. "The Folk Theorem In Repeated Games With Discounting Or With Incomplete Information," World Scientific Book Chapters, in: Drew Fudenberg & David K Levine (ed.), A Long-Run Collaboration On Long-Run Games, chapter 11, pages 209-230, World Scientific Publishing Co. Pte. Ltd..
    35. Sergiu Hart & Andreu Mas-Colell, 2013. "Uncoupled Dynamics Do Not Lead To Nash Equilibrium," World Scientific Book Chapters, in: Simple Adaptive Strategies From Regret-Matching to Uncoupled Dynamics, chapter 7, pages 153-163, World Scientific Publishing Co. Pte. Ltd..
    36. Hofbauer, Josef & Hopkins, Ed, 2005. "Learning in perturbed asymmetric games," Games and Economic Behavior, Elsevier, vol. 52(1), pages 133-152, July.
    37. Fudenberg, Drew & Imhof, Lorens A., 2006. "Imitation processes with small mutations," Journal of Economic Theory, Elsevier, vol. 131(1), pages 251-262, November.
    38. Emilio Calvano & Giacomo Calzolari & Vincenzo Denicolò & Sergio Pastorello, 2020. "Artificial Intelligence, Algorithmic Pricing, and Collusion," American Economic Review, American Economic Association, vol. 110(10), pages 3267-3297, October.
    39. John H. Nachbar, 1997. "Prediction, Optimization, and Learning in Repeated Games," Econometrica, Econometric Society, vol. 65(2), pages 275-310, March.
    40. Abreu, Dilip & Dutta, Prajit K & Smith, Lones, 1994. "The Folk Theorem for Repeated Games: A NEU Condition," Econometrica, Econometric Society, vol. 62(4), pages 939-948, July.
    41. Karandikar, Rajeeva & Mookherjee, Dilip & Ray, Debraj & Vega-Redondo, Fernando, 1998. "Evolving Aspirations and Cooperation," Journal of Economic Theory, Elsevier, vol. 80(2), pages 292-331, June.
    42. Waltman, Ludo & Kaymak, Uzay, 2008. "Q-learning agents in a Cournot oligopoly model," Journal of Economic Dynamics and Control, Elsevier, vol. 32(10), pages 3275-3293, October.
    43. Govindan, Srihari & Reny, Philip J. & Robson, Arthur J., 2003. "A short proof of Harsanyi's purification theorem," Games and Economic Behavior, Elsevier, vol. 45(2), pages 369-374, November.
    44. , & ,, 2010. "A theory of regular Markov perfect equilibria in dynamic stochastic games: genericity, stability, and purification," Theoretical Economics, Econometric Society, vol. 5(3), September.
    45. Abreu, Dilip & Rubinstein, Ariel, 1988. "The Structure of Nash Equilibrium in Repeated Games with Finite Automata," Econometrica, Econometric Society, vol. 56(6), pages 1259-1281, November.
    46. Fudenberg Drew & Kreps David M., 1993. "Learning Mixed Equilibria," Games and Economic Behavior, Elsevier, vol. 5(3), pages 320-367, July.
    47. Mailath, George J. & Samuelson, Larry, 2006. "Repeated Games and Reputations: Long-Run Relationships," OUP Catalogue, Oxford University Press, number 9780195300796.
    48. Borgers, Tilman & Sarin, Rajiv, 1997. "Learning Through Reinforcement and Replicator Dynamics," Journal of Economic Theory, Elsevier, vol. 77(1), pages 1-14, November.
    49. Karsten T. Hansen & Kanishka Misra & Mallesh M. Pai, 2021. "Frontiers: Algorithmic Collusion: Supra-competitive Prices via," Marketing Science, INFORMS, vol. 40(1), pages 1-12, January.
    50. Sergiu Hart & Andreu Mas-Colell, 2013. "A Simple Adaptive Procedure Leading To Correlated Equilibrium," World Scientific Book Chapters, in: Simple Adaptive Strategies From Regret-Matching to Uncoupled Dynamics, chapter 2, pages 17-46, World Scientific Publishing Co. Pte. Ltd..
    51. Alós-Ferrer, Carlos & Netzer, Nick, 2010. "The logit-response dynamics," Games and Economic Behavior, Elsevier, vol. 68(2), pages 413-427, March.
    52. Benaim, Michel & Hirsch, Morris W., 1999. "Mixed Equilibria and Dynamical Systems Arising from Fictitious Play in Perturbed Games," Games and Economic Behavior, Elsevier, vol. 29(1-2), pages 36-72, October.
    53. John Asker & Chaim Fershtman & Ariel Pakes, 2022. "Artificial Intelligence, Algorithm Design, and Pricing," AEA Papers and Proceedings, American Economic Association, vol. 112, pages 452-456, May.
    54. Kalai, Ehud & Lehrer, Ehud, 1993. "Rational Learning Leads to Nash Equilibrium," Econometrica, Econometric Society, vol. 61(5), pages 1019-1045, September.
    55. Barlo, Mehmet & Carmona, Guilherme & Sabourian, Hamid, 2016. "Bounded memory Folk Theorem," Journal of Economic Theory, Elsevier, vol. 163(C), pages 728-774.
    56. Drew Fudenberg & David K. Levine, 1998. "The Theory of Learning in Games," MIT Press Books, The MIT Press, edition 1, volume 1, number 0262061945, December.
    57. Williams, Noah, 2022. "Learning and equilibrium transitions: Stochastic stability in discounted stochastic fictitious play," Journal of Economic Dynamics and Control, Elsevier, vol. 145(C).
    58. Ed Hopkins, 2002. "Two Competing Models of How People Learn in Games," Econometrica, Econometric Society, vol. 70(6), pages 2141-2166, November.
    59. Jinglong Li, 2024. "Stock Portfolio Optimization Based on Reinforcement Learning," Advances in Economics, Business and Management Research, in: Feng-xia Cao & Satya Narayan Singh & Ahmad Jusoh & Deepanjali Mishra (ed.), Proceedings of the 2023 5th International Conference on Economic Management and Cultural Industry (ICEMCI 2023), pages 123-130, Springer.
    60. Joseph E Harrington, 2018. "Developing Competition Law For Collusion By Autonomous Artificial Agents," Journal of Competition Law and Economics, Oxford University Press, vol. 14(3), pages 331-363.
    61. Thomas Wiseman, 2005. "A Partial Folk Theorem for Games with Unknown Payoff Distributions," Econometrica, Econometric Society, vol. 73(2), pages 629-645, March.
    62. Cho, In-Koo & Matsui, Akihiko, 2005. "Learning aspiration in repeated games," Journal of Economic Theory, Elsevier, vol. 124(2), pages 171-201, October.
    63. Barlo, Mehmet & Carmona, Guilherme & Sabourian, Hamid, 2009. "Repeated games with one-memory," Journal of Economic Theory, Elsevier, vol. 144(1), pages 312-336, January.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Jonathan Newton, 2018. "Evolutionary Game Theory: A Renaissance," Games, MDPI, vol. 9(2), pages 1-67, May.
    2. Naoki Funai, 2019. "Convergence results on stochastic adaptive learning," Economic Theory, Springer;Society for the Advancement of Economic Theory (SAET), vol. 68(4), pages 907-934, November.
    3. Burkhard C. Schipper, 2022. "Strategic Teaching and Learning in Games," American Economic Journal: Microeconomics, American Economic Association, vol. 14(3), pages 321-352, August.
    4. Chernov, G. & Susin, I., 2019. "Models of learning in games: An overview," Journal of the New Economic Association, New Economic Association, vol. 44(4), pages 77-125.
    5. Burkhard Schipper, 2015. "Strategic teaching and learning in games," Working Papers 151, University of California, Davis, Department of Economics.
    6. Benaïm, Michel & Hofbauer, Josef & Hopkins, Ed, 2009. "Learning in games with unstable equilibria," Journal of Economic Theory, Elsevier, vol. 144(4), pages 1694-1709, July.
    7. Martino Banchio & Giacomo Mantegazza, 2022. "Artificial Intelligence and Spontaneous Collusion," Papers 2202.05946, arXiv.org, revised Sep 2023.
    8. Leslie, David S. & Collins, E.J., 2006. "Generalised weakened fictitious play," Games and Economic Behavior, Elsevier, vol. 56(2), pages 285-298, August.
    9. Jindani, Sam, 2022. "Learning efficient equilibria in repeated games," Journal of Economic Theory, Elsevier, vol. 205(C).
    10. Ueda, Masahiko, 2023. "Memory-two strategies forming symmetric mutual reinforcement learning equilibrium in repeated prisoners’ dilemma game," Applied Mathematics and Computation, Elsevier, vol. 444(C).
    11. Mario Bravo & Mathieu Faure, 2013. "Reinforcement Learning with Restrictions on the Action Set," AMSE Working Papers 1335, Aix-Marseille School of Economics, France, revised 01 Jul 2013.
    12. Funai, Naoki, 2022. "Reinforcement learning with foregone payoff information in normal form games," Journal of Economic Behavior & Organization, Elsevier, vol. 200(C), pages 638-660.
    13. Mengel, Friederike, 2012. "Learning across games," Games and Economic Behavior, Elsevier, vol. 74(2), pages 601-619.
    14. Jakub Bielawski & Thiparat Chotibut & Fryderyk Falniowski & Michal Misiurewicz & Georgios Piliouras, 2022. "Unpredictable dynamics in congestion games: memory loss can prevent chaos," Papers 2201.10992, arXiv.org, revised Jan 2022.
    15. Mario Bravo, 2016. "An Adjusted Payoff-Based Procedure for Normal Form Games," Mathematics of Operations Research, INFORMS, vol. 41(4), pages 1469-1483, November.
    16. Dridi, Slimane & Lehmann, Laurent, 2014. "On learning dynamics underlying the evolution of learning rules," Theoretical Population Biology, Elsevier, vol. 91(C), pages 20-36.
    17. Block, Juan I. & Fudenberg, Drew & Levine, David K., 2019. "Learning dynamics with social comparisons and limited memory," Theoretical Economics, Econometric Society, vol. 14(1), January.
    18. Dean P Foster & Peyton Young, 2006. "Regret Testing Leads to Nash Equilibrium," Levine's Working Paper Archive 784828000000000676, David K. Levine.
    19. Panayotis Mertikopoulos & William H. Sandholm, 2016. "Learning in Games via Reinforcement and Regularization," Mathematics of Operations Research, INFORMS, vol. 41(4), pages 1297-1324, November.
    20. Hofbauer, Josef & Hopkins, Ed, 2005. "Learning in perturbed asymmetric games," Games and Economic Behavior, Elsevier, vol. 52(1), pages 133-152, July.

    More about this item

    Keywords

    ;
    ;
    ;
    ;

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:gamebe:v:157:y:2026:i:c:p:1-21. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/inca/622836 .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.