Policy Learning with Competing Agents

My bibliography Save this paper

Policy Learning with Competing Agents

Author

Listed:

Roshni Sahoo
Stefan Wager

Registered:

Abstract

Decision makers often aim to learn a treatment assignment policy under a capacity constraint on the number of agents that they can treat. When agents can respond strategically to such policies, competition arises, complicating estimation of the optimal policy. In this paper, we study capacity-constrained treatment assignment in the presence of such interference. We consider a dynamic model where the decision maker allocates treatments at each time step and heterogeneous agents myopically best respond to the previous treatment assignment policy. When the number of agents is large but finite, we show that the threshold for receiving treatment under a given policy converges to the policy's mean-field equilibrium threshold. Based on this result, we develop a consistent estimator for the policy gradient. In a semi-synthetic experiment with data from the National Education Longitudinal Study of 1988, we demonstrate that this estimator can be used for learning capacity-constrained policies in the presence of strategic behavior.

Suggested Citation

Roshni Sahoo & Stefan Wager, 2022. "Policy Learning with Competing Agents," Papers 2204.01884, arXiv.org, revised Mar 2025.

Handle: RePEc:arx:papers:2204.01884

Download full text from publisher

References listed on IDEAS

Heckman, James J & Lochner, Lance & Taber, Christopher, 1998. "General-Equilibrium Treatment Effects: A Study of Tuition Policy," American Economic Review, American Economic Association, vol. 88(2), pages 381-386, May.
- James J. Heckman & Lance Lochner & Christopher Taber, 1998. "General Equilibrium Treatment Effects: A Study of Tuition Policy," NBER Working Papers 6426, National Bureau of Economic Research, Inc.
John Bound & Brad Hershbein & Bridget Terry Long, 2009. "Playing the Admissions Game: Student Reactions to Increasing College Competition," Journal of Economic Perspectives, American Economic Association, vol. 23(4), pages 119-146, Fall.
- John Bound & Brad Hershbein & Bridget Terry Long, 2009. "Playing the Admissions Game: Student Reactions to Increasing College Competition," NBER Working Papers 15272, National Bureau of Economic Research, Inc.
Toru Kitagawa & Aleksey Tetenov, 2018. "Who Should Be Treated? Empirical Welfare Maximization Methods for Treatment Choice," Econometrica, Econometric Society, vol. 86(2), pages 591-616, March.
- Toru Kitagawa & Aleksey Tetenov, 2015. "Who should be treated? Empirical welfare maximization methods for treatment choice," CeMMAP working papers CWP10/15, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
- Toru Kitagawa & Aleksey Tetenov, 2017. "Who should be treated? Empirical welfare maximization methods for treatment choice," CeMMAP working papers CWP24/17, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
- Toru Kitagawa & Aleksey Tetenov, 2015. "Who should be Treated? Empirical Welfare Maximization Methods for Treatment Choice," Carlo Alberto Notebooks 402, Collegio Carlo Alberto.
Evan Munro & Xu Kuang & Stefan Wager, 2021. "Treatment Effects in Market Equilibrium," Papers 2109.11647, arXiv.org, revised Feb 2025.
Kandori, Michihiro & Mailath, George J & Rob, Rafael, 1993. "Learning, Mutation, and Long Run Equilibria in Games," Econometrica, Econometric Society, vol. 61(1), pages 29-56, January.
- Kandori, M. & Mailath, G.J., 1991. "Learning, Mutation, And Long Run Equilibria In Games," Papers 71, Princeton, Woodrow Wilson School - John M. Olin Program.
- M. Kandori & G. Mailath & R. Rob, 1999. "Learning, Mutation and Long Run Equilibria in Games," Levine's Working Paper Archive 500, David K. Levine.
Bhattacharya, Debopam & Dupas, Pascaline, 2012. "Inferring welfare maximizing treatment assignment under budget constraints," Journal of Econometrics, Elsevier, vol. 167(1), pages 168-196.
- Debopam Bhattacharya & Pascaline Dupas, 2008. "Inferring Welfare Maximizing Treatment Assignment under Budget Constraints," NBER Working Papers 14447, National Bureau of Economic Research, Inc.
Daron Acemoglu & Martin Kaae Jensen, 2015. "Robust Comparative Statics in Large Dynamic Economies," Journal of Political Economy, University of Chicago Press, vol. 123(3), pages 587-640.
- Daron Acemoglu & Martin Kaae Jensen, 2012. "Robust Comparative Statics in Large Dynamic Economies," NBER Working Papers 18178, National Bureau of Economic Research, Inc.
- Daron Acemoglu & Martin Kaae Jensen, 2012. "Robust Comparative Statics in Large Dynamic Economies," Levine's Working Paper Archive 786969000000000507, David K. Levine.
Charles F. Manski, 2004. "Statistical Treatment Rules for Heterogeneous Populations," Econometrica, Econometric Society, vol. 72(4), pages 1221-1246, July.
- Charles F. Manski, 2003. "Statistical treatment rules for heterogeneous populations," CeMMAP working papers 03/03, Institute for Fiscal Studies.
- Charles F. Manski, 2003. "Statistical treatment rules for heterogeneous populations," CeMMAP working papers CWP03/03, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
Jon Kleinberg & Himabindu Lakkaraju & Jure Leskovec & Jens Ludwig & Sendhil Mullainathan, 2018. "Human Decisions and Machine Predictions," The Quarterly Journal of Economics, President and Fellows of Harvard College, vol. 133(1), pages 237-293.
- Jon Kleinberg & Himabindu Lakkaraju & Jure Leskovec & Jens Ludwig & Sendhil Mullainathan, 2017. "Human Decisions and Machine Predictions," NBER Working Papers 23180, National Bureau of Economic Research, Inc.
Raj Chetty, 2009. "Sufficient Statistics for Welfare Analysis: A Bridge Between Structural and Reduced-Form Methods," Annual Review of Economics, Annual Reviews, vol. 1(1), pages 451-488, May.
- Raj Chetty, 2008. "Sufficient Statistics for Welfare Analysis: A Bridge Between Structural and Reduced-Form Methods," NBER Working Papers 14399, National Bureau of Economic Research, Inc.
- Chetty, Nadarajan, 2009. "Sufficient Statistics for Welfare Analysis: A Bridge Between Structural and Reduced-Form Methods," Scholarly Articles 9748528, Harvard University Department of Economics.
Daniel Bjorkegren & Joshua E. Blumenstock & Samsun Knight, 2020. "Manipulation-Proof Machine Learning," Papers 2004.03865, arXiv.org.
Meena Jagadeesan & Celestine Mendler-Dunner & Moritz Hardt, 2021. "Alternative Microfoundations for Strategic Classification," Papers 2106.12705, arXiv.org.
Monderer, Dov & Sela, Aner, 1996. "A2 x 2Game without the Fictitious Play Property," Games and Economic Behavior, Elsevier, vol. 14(1), pages 144-148, May.
Monderer, Dov & Shapley, Lloyd S., 1996. "Fictitious Play Property for Games with Identical Interests," Journal of Economic Theory, Elsevier, vol. 68(1), pages 258-265, January.

Full references (including those not matched with items on IDEAS)

Citations

Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.

Cited by:

Luofeng Liao & Christian Kroer, 2023. "Statistical Inference and A/B Testing for First-Price Pacing Equilibria," Papers 2301.02276, arXiv.org, revised Jun 2023.
Luofeng Liao & Christian Kroer, 2024. "Statistical Inference and A/B Testing in Fisher Markets and Paced Auctions," Papers 2406.15522, arXiv.org, revised Mar 2025.
Daido Kido, 2023. "Incorporating Preferences Into Treatment Assignment Problems," Papers 2311.08963, arXiv.org.
Luofeng Liao & Yuan Gao & Christian Kroer, 2022. "Statistical Inference for Fisher Market Equilibrium," Papers 2209.15422, arXiv.org, revised Feb 2025.

Most related items

These are the items that most often cite the same works as this one and are cited by the same works as this one.

Garbero, Alessandra & Sakos, Grayson & Cerulli, Giovanni, 2023. "Towards data-driven project design: Providing optimal treatment rules for development projects," Socio-Economic Planning Sciences, Elsevier, vol. 89(C).
- Garbero, Alessandra & Sakos, Grayson & Cerulli, Giovanni, 2021. "Towards Data-driven Project design: Providing Optimal Treatment Rules for Development Projects," 2021 Annual Meeting, August 1-3, Austin, Texas 314016, Agricultural and Applied Economics Association.
Susan Athey & Stefan Wager, 2021. "Policy Learning With Observational Data," Econometrica, Econometric Society, vol. 89(1), pages 133-161, January.
- Susan Athey & Stefan Wager, 2017. "Policy Learning with Observational Data," Papers 1702.02896, arXiv.org, revised Sep 2020.
Juan Carlos Perdomo, 2023. "The Relative Value of Prediction in Algorithmic Decision Making," Papers 2312.08511, arXiv.org, revised May 2024.
Firpo, Sergio & Galvao, Antonio F. & Kobus, Martyna & Parker, Thomas & Rosa-Dias, Pedro, 2025. "Loss aversion and the welfare ranking of policy interventions," Journal of Econometrics, Elsevier, vol. 252(PB).
Eric Mbakop & Max Tabord‐Meehan, 2021. "Model Selection for Treatment Choice: Penalized Welfare Maximization," Econometrica, Econometric Society, vol. 89(2), pages 825-848, March.
- Eric Mbakop & Max Tabord-Meehan, 2016. "Model Selection for Treatment Choice: Penalized Welfare Maximization," Papers 1609.03167, arXiv.org, revised Dec 2020.
Anders Bredahl Kock & Martin Thyrsgaard, 2017. "Optimal sequential treatment allocation," Papers 1705.09952, arXiv.org, revised Aug 2018.
Firpo, Sergio & Galvao, Antonio F. & Kobus, Martyna & Parker, Thomas & Rosa-Dias, Pedro, 2025. "Loss aversion and the welfare ranking of policy interventions," Journal of Econometrics, Elsevier, vol. 252(PB).
- Firpo, Sergio & Galvao, Antonio F. & Kobus, Martyna & Parker, Thomas & Rosa-Dias, Pedro, 2020. "Loss Aversion and the Welfare Ranking of Policy Interventions," IZA Discussion Papers 13176, Institute of Labor Economics (IZA).
- Sergio Firpo & Antonio F. Galvao & Martyna Kobus & Thomas Parker & Pedro Rosa-Dias, 2020. "Loss aversion and the welfare ranking of policy interventions," Papers 2004.08468, arXiv.org, revised Sep 2023.
Giovanni Cerulli, 2020. "Optimal Policy Learning: From Theory to Practice," Papers 2011.04993, arXiv.org.
Kock, Anders Bredahl & Preinerstorfer, David & Veliyev, Bezirgen, 2023. "Treatment recommendation with distributional targets," Journal of Econometrics, Elsevier, vol. 234(2), pages 624-646.
- Anders Bredahl Kock & David Preinerstorfer & Bezirgen Veliyev, 2020. "Treatment recommendation with distributional targets," Papers 2005.09717, arXiv.org, revised Apr 2022.
Nan Liu & Yanbo Liu & Yuya Sasaki & Yuanyuan Wan, 2025. "Nonparametric Uniform Inference in Binary Classification and Policy Values," Papers 2511.14700, arXiv.org, revised Dec 2025.
Shosei Sakaguchi, 2021. "Estimation of Optimal Dynamic Treatment Assignment Rules under Policy Constraints," Papers 2106.05031, arXiv.org, revised Aug 2024.
Nygaard, Vegard M. & Sørensen, Bent E. & Wang, Fan, 2022. "Optimal allocations to heterogeneous agents with an application to stimulus checks," Journal of Economic Dynamics and Control, Elsevier, vol. 138(C).
- SÃ¸rensen, Bent E & Nygaard, Vegard M. & Wang, Fan, 2020. "Optimal allocations to heterogeneous agents with an application to stimulus checks," CEPR Discussion Papers 15283, C.E.P.R. Discussion Papers.
- Vegard M. Nygaard & Bent E. S{o}rensen & Fan Wang, 2022. "Optimal allocations to heterogeneous agents with an application to stimulus checks," Papers 2204.03799, arXiv.org.
Toru Kitagawa & Weining Wang & Mengshan Xu, 2022. "Policy Choice in Time Series by Empirical Welfare Maximization," Papers 2205.03970, arXiv.org, revised Nov 2025.
Yu-Chang Chen & Haitian Xie, 2022. "Personalized Subsidy Rules," Papers 2202.13545, arXiv.org, revised Mar 2022.
Johannes Haushofer & Paul Niehaus & Carlos Paramo & Edward Miguel & Michael Walker, 2025. "Targeting Impact versus Deprivation," American Economic Review, American Economic Association, vol. 115(6), pages 1936-1974, June.
- Johannes Haushofer & Paul Niehaus & Carlos Paramo & Edward Miguel & Michael W. Walker, 2022. "Targeting Impact versus Deprivation," NBER Working Papers 30138, National Bureau of Economic Research, Inc.
- Haushofer, Johannes & Niehaus, Paul & Paramo, Carlos & Miguel, Edward & Walker, Michael W, 2022. "Targeting Impact Versus Deprivation," Department of Economics, Working Paper Series qt07j8n9vz, Department of Economics, Institute for Business and Economic Research, UC Berkeley.
Huber, Martin, 2019. "An introduction to flexible methods for policy evaluation," FSES Working Papers 504, Faculty of Economics and Social Sciences, University of Freiburg/Fribourg Switzerland.
- Martin Huber, 2019. "An introduction to flexible methods for policy evaluation," Papers 1910.00641, arXiv.org.
Juliano Assunção & Robert McMillan & Joshua Murphy & Eduardo Souza-Rodrigues, 2019. "Optimal Environmental Targeting in the Amazon Rainforest," NBER Working Papers 25636, National Bureau of Economic Research, Inc.
Toru Kitagawa & Weining Wang & Mengshan Xu, 2024. "Policy choice in time series by empirical welfare maximization," CeMMAP working papers 27/24, Institute for Fiscal Studies.
Yuya Sasaki & Takuya Ura, 2020. "Welfare Analysis via Marginal Treatment Effects," Papers 2012.07624, arXiv.org.
Karun Adusumilli & Friedrich Geiecke & Claudio Schilter, 2019. "Dynamically Optimal Treatment Allocation," Papers 1904.01047, arXiv.org, revised Nov 2024.

More about this item

NEP fields

This paper has been announced in the following NEP Reports:

NEP-GTH-2022-05-16 (Game Theory)
NEP-MIC-2022-05-16 (Microeconomics)

Statistics

Access and download statistics

Corrections

All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:arx:papers:2204.01884. See general information about how to correct material in RePEc.

If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: arXiv administrators (email available below). General contact details of provider: http://arxiv.org/ .

Please note that corrections may take a couple of weeks to filter through the various RePEc services.

IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.

Browse Econ Literature

More features

Policy Learning with Competing Agents

Author

Abstract

Suggested Citation

Download full text from publisher

References listed on IDEAS

Citations

Most related items

More about this item

NEP fields

Statistics

Corrections

More services and features

MyIDEAS

Author registration

Rankings

RePEc Genealogy

RePEc Biblio

MPRA

New papers by email

EconAcademics

Plagiarism

About RePEc

RePEc home

Blog

Help/FAQ

RePEc team

Participating archives

Privacy statement

Help us

Corrections

Volunteers

Get papers listed

Open a RePEc archive

Get RePEc data