Simple fixes that accommodate switching costs in multi-armed bandits
Author
Abstract
Suggested Citation
DOI: 10.1016/j.ejor.2024.09.017
Download full text from publisher
As the access to this document is restricted, you may want to search for a different version of it.
References listed on IDEAS
- Banks, Jeffrey S & Sundaram, Rangarajan K, 1994. "Switching Costs and the Gittins Index," Econometrica, Econometric Society, vol. 62(3), pages 687-694, May.
- Ali Yekkehkhany & Ebrahim Arian & Rakesh Nagi & Ilan Shomorony, 2021. "A cost–based analysis for risk–averse explore–then–commit finite–time bandits," IISE Transactions, Taylor & Francis Journals, vol. 53(10), pages 1094-1108, October.
- Brezzi, Monica & Lai, Tze Leung, 2002. "Optimal learning and experimentation in bandit problems," Journal of Economic Dynamics and Control, Elsevier, vol. 27(1), pages 87-108, November.
- Malekipirbazari, Milad & Çavuş, Özlem, 2024. "Index policy for multiarmed bandit problem with dynamic risk measures," European Journal of Operational Research, Elsevier, vol. 312(2), pages 627-640.
- Xu, Jianyu & Chen, Lujie & Tang, Ou, 2021. "An online algorithm for the risk-aware restless bandit," European Journal of Operational Research, Elsevier, vol. 290(2), pages 622-639.
- Daniel Russo & Benjamin Van Roy, 2014. "Learning to Optimize via Posterior Sampling," Mathematics of Operations Research, INFORMS, vol. 39(4), pages 1221-1243, November.
Most related items
These are the items that most often cite the same works as this one and are cited by the same works as this one.- Malekipirbazari, Milad, 2025. "Optimizing sequential decision-making under risk: Strategic allocation with switching penalties," European Journal of Operational Research, Elsevier, vol. 321(1), pages 160-176.
- Michael Jong Kim, 2020. "Variance Regularization in Sequential Bayesian Optimization," Mathematics of Operations Research, INFORMS, vol. 45(3), pages 966-992, August.
- José Niño-Mora, 2023. "Markovian Restless Bandits and Index Policies: A Review," Mathematics, MDPI, vol. 11(7), pages 1-27, March.
- Eric M. Schwartz & Eric T. Bradlow & Peter S. Fader, 2017. "Customer Acquisition via Display Advertising Using Multi-Armed Bandit Experiments," Marketing Science, INFORMS, vol. 36(4), pages 500-522, July.
- David Simchi-Levi & Rui Sun & Huanan Zhang, 2022. "Online Learning and Optimization for Revenue Management Problems with Add-on Discounts," Management Science, INFORMS, vol. 68(10), pages 7402-7421, October.
- Rong Jin & David Simchi-Levi & Li Wang & Xinshang Wang & Sen Yang, 2021. "Shrinking the Upper Confidence Bound: A Dynamic Product Selection Problem for Urban Warehouses," Management Science, INFORMS, vol. 67(8), pages 4756-4771, August.
- Theodore Papageorgiou, 2022.
"Occupational Matching and Cities,"
American Economic Journal: Macroeconomics, American Economic Association, vol. 14(3), pages 82-132, July.
- Theodore Papageorgiou, 2020. "Occupational Matching and Cities," Boston College Working Papers in Economics 1011, Boston College Department of Economics.
- José Niño-Mora, 2020. "Fast Two-Stage Computation of an Index Policy for Multi-Armed Bandits with Setup Delays," Mathematics, MDPI, vol. 9(1), pages 1-36, December.
- Raluca M. Ursu & Qingliang Wang & Pradeep K. Chintagunta, 2020. "Search Duration," Marketing Science, INFORMS, vol. 39(5), pages 849-871, September.
- Victor F. Araman & René A. Caldentey, 2022. "Diffusion Approximations for a Class of Sequential Experimentation Problems," Management Science, INFORMS, vol. 68(8), pages 5958-5979, August.
- John R. Hauser & Guilherme (Gui) Liberali & Glen L. Urban, 2014. "Website Morphing 2.0: Switching Costs, Partial Exposure, Random Exit, and When to Morph," Management Science, INFORMS, vol. 60(6), pages 1594-1616, June.
- Mengying Zhu & Xiaolin Zheng & Yan Wang & Yuyuan Li & Qianqiao Liang, 2019. "Adaptive Portfolio by Solving Multi-armed Bandit via Thompson Sampling," Papers 1911.05309, arXiv.org, revised Nov 2019.
- Brenner, Thomas & Vriend, Nicolaas J., 2006.
"On the behavior of proposers in ultimatum games,"
Journal of Economic Behavior & Organization, Elsevier, vol. 61(4), pages 617-631, December.
- Thomas Brenner & Nicolaas J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," Working Papers 502, Queen Mary University of London, School of Economics and Finance.
- Thomas Brenner & Nicolaas J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," Working Papers 502, Queen Mary University of London, School of Economics and Finance.
- T. Brenner & N.J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," Papers on Economics and Evolution 2003-08, Philipps University Marburg, Department of Geography.
- Thomas Brenner & Nicolaas J. Vriend, 2003. "On the Behavior of Proposers in Ultimatum Games," CEEL Working Papers 0304, Cognitive and Experimental Economics Laboratory, Department of Economics, University of Trento, Italia.
- Song Lin & Juanjuan Zhang & John R. Hauser, 2015. "Learning from Experience, Simply," Marketing Science, INFORMS, vol. 34(1), pages 1-19, January.
- Bergemann, Dirk & Valimaki, Juuso, 2001.
"Stationary multi-choice bandit problems,"
Journal of Economic Dynamics and Control, Elsevier, vol. 25(10), pages 1585-1594, October.
- Dirk Bergemann & Juuso Vaimaki, 1999. "Stationary Multi Choice Bandit Problems," Cowles Foundation Discussion Papers 1240, Cowles Foundation for Research in Economics, Yale University.
- John Kennan & James R. Walker, 2011.
"The Effect of Expected Income on Individual Migration Decisions,"
Econometrica, Econometric Society, vol. 79(1), pages 211-251, January.
- Kennan,J. & Walker,J.R., 2003. "The effect of expected income on individual migration decisions," Working papers 7, Wisconsin Madison - Social Systems.
- John Kennan & James R. Walker, 2003. "The Effect of Expected Income on Individual Migration Decisions," NBER Working Papers 9585, National Bureau of Economic Research, Inc.
- Pengjie Zhou & Haoyu Wei & Huiming Zhang, 2024. "Selective Reviews of Bandit Problems in AI via a Statistical View," Papers 2412.02251, arXiv.org, revised Feb 2025.
- Stephen Chick & Martin Forster & Paolo Pertile, 2017.
"A Bayesian decision theoretic model of sequential experimentation with delayed response,"
Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 79(5), pages 1439-1462, November.
- Stephen Chick & Martin Forster & Paolo Pertile, 2015. "A Bayesian Decision-Theoretic Model of Sequential Experimentation with Delayed Response," Discussion Papers 15/09, Department of Economics, University of York.
- Mingyu Joo & Michael L. Thompson & Greg M. Allenby6, 2019. "Optimal Product Design by Sequential Experiments in High Dimensions," Management Science, INFORMS, vol. 65(7), pages 3235-3254, July.
- Anand Kalvit & Aleksandrs Slivkins & Yonatan Gur, 2024. "Incentivized Exploration via Filtered Posterior Sampling," Papers 2402.13338, arXiv.org.
More about this item
Keywords
Multi-armed bandit; Switching cost; Ambiguity; Regret analysis;All these keywords.
Statistics
Access and download statisticsCorrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:ejores:v:320:y:2025:i:3:p:616-627. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/eor .
Please note that corrections may take a couple of weeks to filter through the various RePEc services.