Optimal learning and experimentation in bandit problems
Author
Abstract
Suggested Citation
Download full text from publisher
As the access to this document is restricted, you may want to
for a different version of it.References listed on IDEAS
- Monica Brezzi & Tze Leung Lai, 2000. "Incomplete Learning from Endogenous Data in Dynamic Allocation," Econometrica, Econometric Society, vol. 68(6), pages 1511-1516, November.
- Rothschild, Michael, 1974. "A two-armed bandit theory of market pricing," Journal of Economic Theory, Elsevier, vol. 9(2), pages 185-202, October.
- McLennan, Andrew, 1984. "Price dispersion and incomplete learning in the long run," Journal of Economic Dynamics and Control, Elsevier, vol. 7(3), pages 331-347, September.
- Banks, Jeffrey S & Sundaram, Rangarajan K, 1994. "Switching Costs and the Gittins Index," Econometrica, Econometric Society, vol. 62(3), pages 687-694, May.
Most related items
These are the items that most often cite the same works as this one and are cited by the same works as this one.- Arthur Charpentier & Romuald Élie & Carl Remlinger, 2023. "Reinforcement Learning in Economics and Finance," Computational Economics, Springer;Society for Computational Economics, vol. 62(1), pages 425-462, June.
- Vives, Xavier, 1997. "Learning from Others: A Welfare Analysis," Games and Economic Behavior, Elsevier, vol. 20(2), pages 177-200, August.
- Bergemann, Dirk & Valimaki, Juuso, 2002.
"Entry and Vertical Differentiation,"
Journal of Economic Theory, Elsevier, vol. 106(1), pages 91-125, September.
- Dirk Bergemann & Juuso Valimaki, 2000. "Entry and Vertical Differentiation," Cowles Foundation Discussion Papers 1277, Cowles Foundation for Research in Economics, Yale University.
- Dirk Bergemann & Valimaki Juuso, 2001. "Entry and Vertical Differentiation," Cowles Foundation Discussion Papers 1302, Cowles Foundation for Research in Economics, Yale University.
- Esponda, Ignacio & Pouzo, Demian, 2021.
"Equilibrium in misspecified Markov decision processes,"
Theoretical Economics, Econometric Society, vol. 16(2), May.
- Ignacio Esponda & Demian Pouzo, 2015. "Equilibrium in Misspecified Markov Decision Processes," Papers 1502.06901, arXiv.org, revised May 2016.
- Fishman, Arthur & Gandal, Neil, 1994.
"Experimentation and learning with networks effects,"
Economics Letters, Elsevier, vol. 44(1-2), pages 103-108.
- Arthur Fishman & Neil Gandal, 1993. "Experimentation and Learning with Network Effects," Industrial Organization 9309001, University Library of Munich, Germany.
- Kuhle, Wolfgang, 2021. "Equilibrium with computationally constrained agents," Mathematical Social Sciences, Elsevier, vol. 109(C), pages 77-92.
- Omar Besbes & Assaf Zeevi, 2015. "On the (Surprising) Sufficiency of Linear Models for Dynamic Pricing with Demand Learning," Management Science, INFORMS, vol. 61(4), pages 723-739, April.
- Camargo, Braz, 2014.
"Learning in society,"
Games and Economic Behavior, Elsevier, vol. 87(C), pages 381-396.
- Braz Camargo, 2006. "Learning in Society," 2006 Meeting Papers 435, Society for Economic Dynamics.
- J. Michael Harrison & N. Bora Keskin & Assaf Zeevi, 2012. "Bayesian Dynamic Pricing Policies: Learning and Earning Under a Binary Prior Distribution," Management Science, INFORMS, vol. 58(3), pages 570-586, March.
- Blume, Andreas & Heidhues, Paul, 2006.
"Private monitoring in auctions,"
Journal of Economic Theory, Elsevier, vol. 131(1), pages 179-211, November.
- Andreas Blume & Paul Heidhues, 2003. "Private Monitoring in Auctions," CIG Working Papers SP II 2003-14, Wissenschaftszentrum Berlin (WZB), Research Unit: Competition and Innovation (CIG).
- Bolton, P. & Harris, C., 1996. "Strategic Experimentation : A Revision," Other publications TiSEM 2cd2755d-6931-488f-948e-5, Tilburg University, School of Economics and Management.
- Bolton, P. & Harris, C., 1996. "Strategic Experimentation : A Revision," Discussion Paper 1996-27, Tilburg University, Center for Economic Research.
- Spagat, M., 1995. "Leaving some stones unturned: A reassessment of iterative planning theory," Journal of Public Economics, Elsevier, vol. 58(1), pages 85-105, September.
- Smith, L. & Sorensen, P., 1997.
"Informational Herding and Optimal Experimentation,"
Economics Papers
139, Economics Group, Nuffield College, University of Oxford.
- Lones Smith & Peter Norman Sorensen, 2006. "Informational Herding and Optimal Experimentation," Cowles Foundation Discussion Papers 1552, Cowles Foundation for Research in Economics, Yale University.
- Lones Smith & Peter Norman Sørensen, 2005. "Informational Herding and Optimal Experimentation," Discussion Papers 05-13, University of Copenhagen. Department of Economics.
- Smith, L. & Sorensen, P., 1997. "Informational Herding and Optimal Experientation," Working papers 97-22, Massachusetts Institute of Technology (MIT), Department of Economics.
- Goyal, Sanjeev, 2003. "Learning in Networks: a survey," Economics Discussion Papers 9983, University of Essex, Department of Economics.
- Urtzi Ayesta & M Erausquin & E Ferreira & P Jacko, 2016. "Optimal Dynamic Resource Allocation to Prevent Defaults," Post-Print hal-01300681, HAL.
- Keller, Godfrey & Oldale, Alison, 2003. "Branching bandits: a sequential search process with correlated pay-offs," Journal of Economic Theory, Elsevier, vol. 113(2), pages 302-315, December.
- Hao Zhang, 2022. "Analytical Solution to a Discrete-Time Model for Dynamic Learning and Decision Making," Management Science, INFORMS, vol. 68(8), pages 5924-5957, August.
- Mason, Robin & Välimäki, Juuso, 2011. "Learning about the arrival of sales," Journal of Economic Theory, Elsevier, vol. 146(4), pages 1699-1711, July.
- Konon, Alexander, 2016. "Career choice under uncertainty," VfS Annual Conference 2016 (Augsburg): Demographic Change 145583, Verein für Socialpolitik / German Economic Association.
Corrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:dyncon:v:27:y:2002:i:1:p:87-108. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/jedc .
Please note that corrections may take a couple of weeks to filter through the various RePEc services.
Printed from https://ideas.repec.org/a/eee/dyncon/v27y2002i1p87-108.html