Probability Matching and Reinforcement Learning
Author
Abstract
Suggested Citation
Download full text from publisher
Other versions of this item:
- Rivas, Javier, 2013. "Probability matching and reinforcement learning," Journal of Mathematical Economics, Elsevier, vol. 49(1), pages 17-21.
References listed on IDEAS
- Erev, Ido & Roth, Alvin E, 1998. "Predicting How People Play Games: Reinforcement Learning in Experimental Games with Unique, Mixed Strategy Equilibria," American Economic Review, American Economic Association, vol. 88(4), pages 848-881, September.
- Borgers, Tilman & Sarin, Rajiv, 2000.
"Naive Reinforcement Learning with Endogenous Aspirations,"
International Economic Review, Department of Economics, University of Pennsylvania and Osaka University Institute of Social and Economic Research Association, vol. 41(4), pages 921-950, November.
- Tilman Börgers & Rajiv Sarin, "undated". "Naive Reinforcement Learning With Endogenous Aspiration," ELSE working papers 037, ESRC Centre on Economics Learning and Social Evolution.
- T. Borgers & R. Sarin, 2010. "Naïve Reinforcement Learning With Endogenous Aspirations," Levine's Working Paper Archive 381, David K. Levine.
- Rubinstein, Ariel, 2002. "Irrational diversification in multiple decision problems," European Economic Review, Elsevier, vol. 46(8), pages 1369-1378, September.
- Roth, Alvin E. & Erev, Ido, 1995. "Learning in extensive-form games: Experimental data and simple dynamic models in the intermediate term," Games and Economic Behavior, Elsevier, vol. 8(1), pages 164-212.
- Javier Rivas, 2008. "Learning within a Markovian Environment," Economics Working Papers ECO2008/13, European University Institute.
- Colin Camerer & Teck-Hua Ho, 1999. "Experience-weighted Attraction Learning in Normal Form Games," Econometrica, Econometric Society, vol. 67(4), pages 827-874, July.
- Kosfeld, Michael & Droste, Edward & Voorneveld, Mark, 2002.
"A myopic adjustment process leading to best-reply matching,"
Games and Economic Behavior, Elsevier, vol. 40(2), pages 270-298, August.
- Droste, E.J.R. & Kosfeld, M. & Voorneveld, M., 1998. "A Myopic Adjustment Process Leading to Best-Reply Matching," Other publications TiSEM 20ed79ed-0621-4383-b1fb-9, Tilburg University, School of Economics and Management.
- Droste, E.J.R. & Kosfeld, M. & Voorneveld, M., 1998. "A Myopic Adjustment Process Leading to Best-Reply Matching," Discussion Paper 1998-111, Tilburg University, Center for Economic Research.
- Samuelson Larry, 1994. "Stochastic Stability in Games with Alternative Best Replies," Journal of Economic Theory, Elsevier, vol. 64(1), pages 35-65, October.
Most related items
These are the items that most often cite the same works as this one and are cited by the same works as this one.- Oyarzun, Carlos & Sarin, Rajiv, 2013.
"Learning and risk aversion,"
Journal of Economic Theory, Elsevier, vol. 148(1), pages 196-225.
- Carlos Oyarzun & Rajiv Sarin, 2005. "Learning and Risk Aversion," Levine's Bibliography 784828000000000482, UCLA Department of Economics.
- Carlos Oyarzun & Rajiv Sarin, 2012. "Learning and Risk Aversion," Levine's Working Paper Archive 786969000000000572, David K. Levine.
- Schuster, Stephan, 2012. "Applications in Agent-Based Computational Economics," MPRA Paper 47201, University Library of Munich, Germany.
- Sarin, Rajiv & Vahid, Farshid, 2001.
"Predicting How People Play Games: A Simple Dynamic Model of Choice,"
Games and Economic Behavior, Elsevier, vol. 34(1), pages 104-122, January.
- Sarin, R. & Vahid, F., 1999. "Predicting how People Play Games: a Simple Dynamic Model of Choice," Monash Econometrics and Business Statistics Working Papers 12/99, Monash University, Department of Econometrics and Business Statistics.
- Dixon, Huw D. & Sbriglia, Patrizia & Somma, Ernesto, 2006. "Learning to collude: An experiment in convergence and equilibrium selection in oligopoly," Research in Economics, Elsevier, vol. 60(3), pages 155-167, September.
- Jaspersen, Johannes G. & Montibeller, Gilberto, 2020. "On the learning patterns and adaptive behavior of terrorist organizations," European Journal of Operational Research, Elsevier, vol. 282(1), pages 221-234.
- Duffy, John, 2006.
"Agent-Based Models and Human Subject Experiments,"
Handbook of Computational Economics, in: Leigh Tesfatsion & Kenneth L. Judd (ed.), Handbook of Computational Economics, edition 1, volume 2, chapter 19, pages 949-1011,
Elsevier.
- John Duffy, 2004. "Agent-Based Models and Human Subject Experiments," Computational Economics 0412001, University Library of Munich, Germany.
- Droste, Edward & Kosfeld, Michael & Voorneveld, Mark, 2003. "Best-reply matching in games," Mathematical Social Sciences, Elsevier, vol. 46(3), pages 291-309, December.
- Ianni, A., 2002. "Reinforcement learning and the power law of practice: some analytical results," Discussion Paper Series In Economics And Econometrics 203, Economics Division, School of Social Sciences, University of Southampton.
- Sergiu Hart & Andreu Mas-Colell, 2013.
"A Simple Adaptive Procedure Leading To Correlated Equilibrium,"
World Scientific Book Chapters, in: Simple Adaptive Strategies From Regret-Matching to Uncoupled Dynamics, chapter 2, pages 17-46,
World Scientific Publishing Co. Pte. Ltd..
- Sergiu Hart & Andreu Mas-Colell, 2000. "A Simple Adaptive Procedure Leading to Correlated Equilibrium," Econometrica, Econometric Society, vol. 68(5), pages 1127-1150, September.
- Sergiu Hart & Andreu Mas-Colell, 1996. "A simple adaptive procedure leading to correlated equilibrium," Economics Working Papers 200, Department of Economics and Business, Universitat Pompeu Fabra, revised Dec 1996.
- S. Hart & A. Mas-Collel, 2010. "A Simple Adaptive Procedure Leading to Correlated Equilibrium," Levine's Working Paper Archive 572, David K. Levine.
- Sergiu Hart & Andreu Mas-Colell, 1997. "A Simple Adaptive Procedure Leading to Correlated Equilibrium," Game Theory and Information 9703006, University Library of Munich, Germany, revised 25 Nov 1997.
- Marco LiCalzi & Roland Mühlenbernd, 2022. "Feature-weighted categorized play across symmetric games," Experimental Economics, Springer;Economic Science Association, vol. 25(3), pages 1052-1078, June.
- Osili, Una Okonkwo & Paulson, Anna, 2014. "Crises and confidence: Systemic banking crises and depositor behavior," Journal of Financial Economics, Elsevier, vol. 111(3), pages 646-660.
- Rottenstreich, Yuval & Kivetz, Ran, 2006. "On decision making without likelihood judgment," Organizational Behavior and Human Decision Processes, Elsevier, vol. 101(1), pages 74-88, September.
- Ed Hopkins, 2002.
"Two Competing Models of How People Learn in Games,"
Econometrica, Econometric Society, vol. 70(6), pages 2141-2166, November.
- Ed Hopkins, 2000. "Two Competing Models of How People Learn in Games," Edinburgh School of Economics Discussion Paper Series 51, Edinburgh School of Economics, University of Edinburgh.
- Ed Hopkins, 2001. "Two Competing Models of How People Learn in Games," NajEcon Working Paper Reviews 625018000000000226, www.najecon.org.
- Ed Hopkins, 2001. "Two Competing Models of How People Learn in Games," Levine's Working Paper Archive 625018000000000226, David K. Levine.
- Francisco Gomes & Michael Haliassos & Tarun Ramadorai, 2021.
"Household Finance,"
Journal of Economic Literature, American Economic Association, vol. 59(3), pages 919-1000, September.
- Haliassos, Michael & Gomes, Francisco, 2020. "Household Finance," CEPR Discussion Papers 14502, C.E.P.R. Discussion Papers.
- Gomes, Francisco J. & Haliassos, Michael & Ramadorai, Tarun, 2020. "Household finance," IMFS Working Paper Series 138, Goethe University Frankfurt, Institute for Monetary and Financial Stability (IMFS).
- Ponti, Giovanni, 2000.
"Continuous-time evolutionary dynamics: theory and practice,"
Research in Economics, Elsevier, vol. 54(2), pages 187-214, June.
- Giovanni Ponti, 1999. "- Continuous-Time Evolutionary Dynamics: Theory And Practice," Working Papers. Serie AD 1999-31, Instituto Valenciano de Investigaciones Económicas, S.A. (Ivie).
- Simon P. Anderson & Jacob K. Goeree & Charles A. Holt, 2002.
"The Logit Equilibrium: A Perspective on Intuitive Behavioral Anomalies,"
Southern Economic Journal, John Wiley & Sons, vol. 69(1), pages 21-47, July.
- Simon P. Anderson & Jacob K. Goeree & Charles A. Holt, 1999. "The Logit Equilibrium: A Perspective on Intuitive Behavioral Anomalies," Virginia Economics Online Papers 332, University of Virginia, Department of Economics.
- Arifovic, Jasmina & Karaivanov, Alexander, 2010.
"Learning by doing vs. learning from others in a principal-agent model,"
Journal of Economic Dynamics and Control, Elsevier, vol. 34(10), pages 1967-1992, October.
- Jasmina Arifovic & Alexander Karaivanov, 2007. "Learning by Doing vs. Learning from Others in a Principal-Agent Model," Discussion Papers dp07-24, Department of Economics, Simon Fraser University.
- Dalton, Michael & Landry, Peter, 2020. "‘Overattention’ to first-hand experience in hiring decisions: Evidence from professional basketball," Journal of Economic Behavior & Organization, Elsevier, vol. 175(C), pages 98-113.
- Funai, Naoki, 2022. "Reinforcement learning with foregone payoff information in normal form games," Journal of Economic Behavior & Organization, Elsevier, vol. 200(C), pages 638-660.
- Hopkins, Ed, 2007.
"Adaptive learning models of consumer behavior,"
Journal of Economic Behavior & Organization, Elsevier, vol. 64(3-4), pages 348-368.
- Ed Hopkins, 2004. "Adaptive Learning Models of Consumer Behaviour," Edinburgh School of Economics Discussion Paper Series 121, Edinburgh School of Economics, University of Edinburgh.
- Ed Hopkins, 2006. "Adaptive Learning Models of Consumer Behaviour," Levine's Bibliography 122247000000000658, UCLA Department of Economics.
- Ed Hopkins, 2010. "Adaptive Learning Models of Consumer Behaviour," Levine's Working Paper Archive 506439000000000346, David K. Levine.
More about this item
Keywords
Probability Matching; Reinforcement Learning;JEL classification:
- C73 - Mathematical and Quantitative Methods - - Game Theory and Bargaining Theory - - - Stochastic and Dynamic Games; Evolutionary Games
NEP fields
This paper has been announced in the following NEP Reports:- NEP-CBE-2011-03-26 (Cognitive and Behavioural Economics)
- NEP-EVO-2011-03-26 (Evolutionary Economics)
- NEP-GTH-2011-03-26 (Game Theory)
- NEP-NEU-2011-03-26 (Neuroeconomics)
Statistics
Access and download statisticsCorrections
All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:lec:leecon:11/20. See general information about how to correct material in RePEc.
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Abbie Sleath (email available below). General contact details of provider: https://edirc.repec.org/data/deleiuk.html .
Please note that corrections may take a couple of weeks to filter through the various RePEc services.