Optimal learning and experimentation in bandit problems
Author
Abstract
Suggested Citation
Download full text from publisher
As the access to this document is restricted, you may want to search for a different version of it.
References listed on IDEAS
- Rothschild, Michael, 1974. "A two-armed bandit theory of market pricing," Journal of Economic Theory, Elsevier, vol. 9(2), pages 185-202, October.
- Banks, Jeffrey S & Sundaram, Rangarajan K, 1994. "Switching Costs and the Gittins Index," Econometrica, Econometric Society, vol. 62(3), pages 687-694, May.
- Monica Brezzi & Tze Leung Lai, 2000. "Incomplete Learning from Endogenous Data in Dynamic Allocation," Econometrica, Econometric Society, vol. 68(6), pages 1511-1516, November.
- McLennan, Andrew, 1984. "Price dispersion and incomplete learning in the long run," Journal of Economic Dynamics and Control, Elsevier, vol. 7(3), pages 331-347, September.
Citations
Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
Cited by:
- Noah Gans & George Knox & Rachel Croson, 2007. "Simple Models of Discrete Choice and Their Performance in Bandit Experiments," Manufacturing & Service Operations Management, INFORMS, vol. 9(4), pages 383-408, December.
- Eric M. Schwartz & Eric T. Bradlow & Peter S. Fader, 2017. "Customer Acquisition via Display Advertising Using Multi-Armed Bandit Experiments," Marketing Science, INFORMS, vol. 36(4), pages 500-522, July.
- Brenner, Thomas & Vriend, Nicolaas J., 2006.
"On the behavior of proposers in ultimatum games,"
Journal of Economic Behavior & Organization, Elsevier, vol. 61(4), pages 617-631, December.
- Thomas Brenner & Nicolaas J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," CEEL Working Papers 0304, Cognitive and Experimental Economics Laboratory, Department of Economics, University of Trento, Italia.
- Thomas Brenner & Nicolaas J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," Working Papers 502, Queen Mary University of London, School of Economics and Finance.
- T. Brenner & N.J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," Papers on Economics and Evolution 2003-08, Philipps University Marburg, Department of Geography.
- Stephen Chick & Martin Forster & Paolo Pertile, 2017.
"A Bayesian decision theoretic model of sequential experimentation with delayed response,"
Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 79(5), pages 1439-1462, November.
- Stephen Chick & Martin Forster & Paolo Pertile, 2015. "A Bayesian Decision-Theoretic Model of Sequential Experimentation with Delayed Response," Discussion Papers 15/09, Department of Economics, University of York.
- Janet M. Currie & W. Bentley MacLeod, 2018.
"Understanding Doctor Decision Making: The Case of Depression,"
NBER Working Papers
24955, National Bureau of Economic Research, Inc.
- Janet M. Currie & W. Bentley MacLeod, 2020. "Understanding Doctor Decision Making: The Case of Depression," Working Papers 2020-77, Princeton University. Economics Department..
- Michael Jong Kim, 2020. "Variance Regularization in Sequential Bayesian Optimization," Mathematics of Operations Research, INFORMS, vol. 45(3), pages 966-992, August.
- Konon, Alexander, 2016. "Career choice under uncertainty," VfS Annual Conference 2016 (Augsburg): Demographic Change 145583, Verein für Socialpolitik / German Economic Association.
- Ilya O. Ryzhov & Warren B. Powell & Peter I. Frazier, 2012. "The Knowledge Gradient Algorithm for a General Class of Online Learning Problems," Operations Research, INFORMS, vol. 60(1), pages 180-195, February.
- Hart E. Posen & Daniel A. Levinthal, 2012. "Chasing a Moving Target: Exploitation and Exploration in Dynamic Environments," Management Science, INFORMS, vol. 58(3), pages 587-601, March.
- Mingyu Joo & Michael L. Thompson & Greg M. Allenby6, 2019. "Optimal Product Design by Sequential Experiments in High Dimensions," Management Science, INFORMS, vol. 65(7), pages 3235-3254, July.
- Stephen E. Chick & Peter Frazier, 2012. "Sequential Sampling with Economics of Selection Procedures," Management Science, INFORMS, vol. 58(3), pages 550-569, March.
- Morozov, Sergei & Mathur, Sudhanshu, 2009. "Massively parallel computation using graphics processors with application to optimal experimentation in dynamic control," MPRA Paper 30298, University Library of Munich, Germany, revised 04 Apr 2011.
- Felipe Caro & Jérémie Gallien, 2007. "Dynamic Assortment with Demand Learning for Seasonal Consumer Goods," Management Science, INFORMS, vol. 53(2), pages 276-292, February.
- Kanishka Misra & Eric M. Schwartz & Jacob Abernethy, 2019. "Dynamic Online Pricing with Incomplete Information Using Multiarmed Bandit Experiments," Marketing Science, INFORMS, vol. 38(2), pages 226-252, March.
- Samuel N. Cohen & Tanut Treetanthiploet, 2019. "Gittins' theorem under uncertainty," Papers 1907.05689, arXiv.org, revised Jun 2021.
- Raluca M. Ursu & Qingliang Wang & Pradeep K. Chintagunta, 2020. "Search Duration," Marketing Science, INFORMS, vol. 39(5), pages 849-871, September.
- Philipp Afèche & Barış Ata, 2013. "Bayesian Dynamic Pricing in Queueing Systems with Unknown Delay Cost Characteristics," Manufacturing & Service Operations Management, INFORMS, vol. 15(2), pages 292-304, May.
- Sergei Morozov & Sudhanshu Mathur, 2012. "Massively Parallel Computation Using Graphics Processors with Application to Optimal Experimentation in Dynamic Control," Computational Economics, Springer;Society for Computational Economics, vol. 40(2), pages 151-182, August.
- Kevin Glazebrook & Joern Meissner & Jochen Schurr, 2012. "How big should my store be? On the interplay between shelf-space, demand learning and assortment decisions," Working Papers MRG/0021, Department of Management Science, Lancaster University, revised Dec 2012.
- Victor F. Araman & René A. Caldentey, 2022. "Diffusion Approximations for a Class of Sequential Experimentation Problems," Management Science, INFORMS, vol. 68(8), pages 5958-5979, August.
- Mathur, Sudhanshu & Morozov, Sergei, 2009. "Massively Parallel Computation Using Graphics Processors with Application to Optimal Experimentation in Dynamic Control," MPRA Paper 16721, University Library of Munich, Germany.
- Pai, Mallesh & Hansen, Karsten, 2020. "Algorithmic Collusion: Supra-competitive Prices via Independent Algorithms," CEPR Discussion Papers 14372, C.E.P.R. Discussion Papers.
- Stephen E. Chick & Noah Gans, 2009. "Economic Analysis of Simulation Selection Problems," Management Science, INFORMS, vol. 55(3), pages 421-437, March.
- Brenner, Thomas & Vriend, Nicolaas J., 2006.
"On the behavior of proposers in ultimatum games,"
Journal of Economic Behavior & Organization, Elsevier, vol. 61(4), pages 617-631, December.
- Thomas Brenner & Nicolaas J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," Working Papers 502, Queen Mary University of London, School of Economics and Finance.
- Thomas Brenner & Nicolaas J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," Working Papers 502, Queen Mary University of London, School of Economics and Finance.
- T. Brenner & N.J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," Papers on Economics and Evolution 2003-08, Philipps University Marburg, Department of Geography.
- Thomas Brenner & Nicolaas J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," CEEL Working Papers 0304, Cognitive and Experimental Economics Laboratory, Department of Economics, University of Trento, Italia.
- Karsten T. Hansen & Kanishka Misra & Mallesh M. Pai, 2021. "Frontiers: Algorithmic Collusion: Supra-competitive Prices via," Marketing Science, INFORMS, vol. 40(1), pages 1-12, January.
Most related items
These are the items that most often cite the same works as this one and are cited by the same works as this one.- Konon, Alexander, 2016. "Career choice under uncertainty," VfS Annual Conference 2016 (Augsburg): Demographic Change 145583, Verein für Socialpolitik / German Economic Association.
- Jason Delaney & Sarah Jacobson & Thorsten Moenig, 2020.
"Preference discovery,"
Experimental Economics, Springer;Economic Science Association, vol. 23(3), pages 694-715, September.
- Jason Delaney & Sarah Jacobson & Thorsten Moenig, 2017. "Preference Discovery," Department of Economics Working Papers 2017-02, Department of Economics, Williams College, revised Dec 2018.
- Jason Delaney & Sarah Jacobson & Thorsten Moenig, 2019. "Preference Discovery," Department of Economics Working Papers 2019-08, Department of Economics, Williams College, revised Jul 2019.
- Arthur Charpentier & Romuald Élie & Carl Remlinger, 2023. "Reinforcement Learning in Economics and Finance," Computational Economics, Springer;Society for Computational Economics, vol. 62(1), pages 425-462, June.
- Klimenko, Mikhail M., 2004. "Industrial targeting, experimentation and long-run specialization," Journal of Development Economics, Elsevier, vol. 73(1), pages 75-105, February.
- Bergemann, Dirk & Valimaki, Juuso, 1996.
"Learning and Strategic Pricing,"
Econometrica, Econometric Society, vol. 64(5), pages 1125-1149, September.
- Dirk Bergemann & Juuso Valimaki, 1996. "Learning and Strategic Pricing," Cowles Foundation Discussion Papers 1113, Cowles Foundation for Research in Economics, Yale University.
- Sorensen, Morten, 2007. "Learning by Investing: Evidence from Venture Capital," SIFR Research Report Series 53, Institute for Financial Research.
- Vives, Xavier, 1997. "Learning from Others: A Welfare Analysis," Games and Economic Behavior, Elsevier, vol. 20(2), pages 177-200, August.
- John Robst & Kimmarie McGOLDRICK, 1999. "The Measurement of Firm Information About Product Demand," Review of Industrial Organization, Springer;The Industrial Organization Society, vol. 15(2), pages 149-163, September.
- Bergemann, Dirk & Valimaki, Juuso, 2002.
"Entry and Vertical Differentiation,"
Journal of Economic Theory, Elsevier, vol. 106(1), pages 91-125, September.
- Dirk Bergemann & Juuso Valimaki, 2000. "Entry and Vertical Differentiation," Cowles Foundation Discussion Papers 1277, Cowles Foundation for Research in Economics, Yale University.
- Dirk Bergemann & Valimaki Juuso, 2001. "Entry and Vertical Differentiation," Cowles Foundation Discussion Papers 1302, Cowles Foundation for Research in Economics, Yale University.
- Umberto Garfagnini & Bruno Strulovici, 2012. "Social Learning and Innovation Cycles (revision of DP#1516, The Dynamics of Innovation)," Discussion Papers 1546, Northwestern University, Center for Mathematical Studies in Economics and Management Science.
- Smith, L. & Sorensen, P., 1997.
"Informational Herding and Optimal Experientation,"
Working papers
97-22, Massachusetts Institute of Technology (MIT), Department of Economics.
- Lones Smith & Peter Norman Sørensen, 2005. "Informational Herding and Optimal Experimentation," Discussion Papers 05-13, University of Copenhagen. Department of Economics.
- Smith, L. & Sorensen, P., 1997. "Informational Herding and Optimal Experimentation," Economics Papers 139, Economics Group, Nuffield College, University of Oxford.
- Lones Smith & Peter Norman Sorensen, 2006. "Informational Herding and Optimal Experimentation," Cowles Foundation Discussion Papers 1552, Cowles Foundation for Research in Economics, Yale University.
- Elena Pastorino, 2004. "Optimal Job Design and Career Dynamics in the Presence of Uncertainty," Econometric Society 2004 North American Summer Meetings 292, Econometric Society.
- Ignacio Esponda & Demian Pouzo, 2015. "Equilibrium in Misspecified Markov Decision Processes," Papers 1502.06901, arXiv.org, revised May 2016.
- Fishman, Arthur & Gandal, Neil, 1994.
"Experimentation and learning with networks effects,"
Economics Letters, Elsevier, vol. 44(1-2), pages 103-108.
- Arthur Fishman & Neil Gandal, 1993. "Experimentation and Learning with Network Effects," Industrial Organization 9309001, University Library of Munich, Germany.
- Fishman, Arthur & Rob, Rafael, 1998.
"Experimentation and Competition,"
Journal of Economic Theory, Elsevier, vol. 78(2), pages 299-320, February.
- Arthur Fishman & Rafael Rob, "undated". ""Experimentation and Competition''," CARESS Working Papres 97-12, University of Pennsylvania Center for Analytic Research and Economics in the Social Sciences.
- Arthur Fishman & Rafael Rob, "undated". "Experimentation and Competition," Penn CARESS Working Papers b530e9a0ad08e45aeff62efaf, Penn Economics Department.
- Urtzi Ayesta & M Erausquin & E Ferreira & P Jacko, 2016. "Optimal Dynamic Resource Allocation to Prevent Defaults," Working Papers hal-01300681, HAL.
- Wieland, Volker, 2000.
"Learning by doing and the value of optimal experimentation,"
Journal of Economic Dynamics and Control, Elsevier, vol. 24(4), pages 501-534, April.
- Volker W. Wieland, 1996. "Learning by doing and the value of optimal experimentation," Finance and Economics Discussion Series 96-5, Board of Governors of the Federal Reserve System (U.S.).
- Kuhle, Wolfgang, 2021. "Equilibrium with computationally constrained agents," Mathematical Social Sciences, Elsevier, vol. 109(C), pages 77-92.
- Omar Besbes & Assaf Zeevi, 2015. "On the (Surprising) Sufficiency of Linear Models for Dynamic Pricing with Demand Learning," Management Science, INFORMS, vol. 61(4), pages 723-739, April.
- Camargo, Braz, 2014.
"Learning in society,"
Games and Economic Behavior, Elsevier, vol. 87(C), pages 381-396.
- Braz Camargo, 2006. "Learning in Society," 2006 Meeting Papers 435, Society for Economic Dynamics.
Corrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:dyncon:v:27:y:2002:i:1:p:87-108. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/jedc .
Please note that corrections may take a couple of weeks to filter through the various RePEc services.