Reinforcement learning in population games

My bibliography Save this article

Reinforcement learning in population games

Author

Listed:

Lahkar, Ratul
Seymour, Robert M.

Registered:

Ratul Lahkar

Abstract

We study reinforcement learning in a population game. Agents in a population game revise mixed strategies using the Cross rule of reinforcement learning. The population state—the probability distribution over the set of mixed strategies—evolves according to the replicator continuity equation which, in its simplest form, is a partial differential equation. The replicator dynamic is a special case in which the initial population state is homogeneous, i.e. when all agents use the same mixed strategy. We apply the continuity dynamic to various classes of symmetric games. Using 3×3 coordination games, we show that equilibrium selection depends on the variance of the initial strategy distribution, or initial population heterogeneity. We give an example of a 2×2 game in which heterogeneity persists even as the mean population state converges to a mixed equilibrium. Finally, we apply the dynamic to negative definite and doubly symmetric games.

Suggested Citation

Lahkar, Ratul & Seymour, Robert M., 2013. "Reinforcement learning in population games," Games and Economic Behavior, Elsevier, vol. 80(C), pages 10-38.

Handle: RePEc:eee:gamebe:v:80:y:2013:i:c:p:10-38
DOI: 10.1016/j.geb.2013.02.006

Download full text from publisher

As the access to this document is restricted, you may want to search for a different version of it.

References listed on IDEAS

Fudenberg Drew & Kreps David M., 1993. "Learning Mixed Equilibria," Games and Economic Behavior, Elsevier, vol. 5(3), pages 320-367, July.
- Fudenberg, D. & Kreps, D.M., 1992. "Learning Mixed Equilibria," Working papers 92-13, Massachusetts Institute of Technology (MIT), Department of Economics.
- Drew Fudenberg & David Kreps, 2010. "Learning Mixed Equilibria," Levine's Working Paper Archive 415, David K. Levine.
Erev, Ido & Roth, Alvin E, 1998. "Predicting How People Play Games: Reinforcement Learning in Experimental Games with Unique, Mixed Strategy Equilibria," American Economic Review, American Economic Association, vol. 88(4), pages 848-881, September.
repec:dau:papers:123456789/1014 is not listed on IDEAS
Tilman Börgers & Antonio J. Morales & Rajiv Sarin, 2004. "Expedient and Monotone Learning Rules," Econometrica, Econometric Society, vol. 72(2), pages 383-405, March.
- Tilman Börgers & Rajiv Sarin & Antonio J. Morales, 2001. "Expedient and Monotone Learning Rules," Economic Working Papers at Centro de Estudios Andaluces E2001/06, Centro de Estudios Andaluces.
- Tilman Borgers & Antonio Morales & Rajiv Sarin, 2010. "Expedient and Monotone Learning Rules," Levine's Working Paper Archive 625018000000000099, David K. Levine.
Fudenberg, Drew & Takahashi, Satoru, 2011. "Heterogeneous beliefs and local information in stochastic fictitious play," Games and Economic Behavior, Elsevier, vol. 71(1), pages 100-120, January.
- Drew Fudenberg & Satoru Takahashi, 2008. "Heterogeneous Beliefs and Local Information in Stochastic Fictitious Play," Levine's Working Paper Archive 122247000000001695, David K. Levine.
- Takahashi, Satoru & Fudenberg, Drew, 2011. "Heterogeneous beliefs and local information in stochastic fictitious play," Scholarly Articles 27755310, Harvard University Department of Economics.
Borgers, Tilman & Sarin, Rajiv, 2000. "Naive Reinforcement Learning with Endogenous Aspirations," International Economic Review, Department of Economics, University of Pennsylvania and Osaka University Institute of Social and Economic Research Association, vol. 41(4), pages 921-950, November.
- Tilman Börgers & Rajiv Sarin, "undated". "Naive Reinforcement Learning With Endogenous Aspiration," ELSE working papers 037, ESRC Centre on Economics Learning and Social Evolution.
- T. Borgers & R. Sarin, 2010. "Naïve Reinforcement Learning With Endogenous Aspirations," Levine's Working Paper Archive 381, David K. Levine.
Hopkins, Ed, 1999. "Learning, Matching, and Aggregation," Games and Economic Behavior, Elsevier, vol. 26(1), pages 79-110, January.
- Ed Hopkins, "undated". "Learning, Matching and Aggregation," Discussion Papers 1996-2, Edinburgh School of Economics, University of Edinburgh.
- Hopkins, E., 1995. "Learning, Matching and Aggregation," G.R.E.Q.A.M. 95a20, Universite Aix-Marseille III.
- Ed Hopkins, 1995. "Learning, Matching and Aggregation," Edinburgh School of Economics Discussion Paper Series 2, Edinburgh School of Economics, University of Edinburgh.
- Ed Hopkins, "undated". "Learning, Matching and Aggregation," ELSE working papers 033, ESRC Centre on Economics Learning and Social Evolution.
- Ed Hopkins, 1995. "Learning, Matching and Aggregation," Game Theory and Information 9512001, University Library of Munich, Germany.
- Ed Hopkins, "undated". "Learning, Matching and Aggregation," Department of Economics 1996 : II, Edinburgh School of Economics, University of Edinburgh.
Sergiu Hart & Andreu Mas-Colell, 2013. "A Simple Adaptive Procedure Leading To Correlated Equilibrium," World Scientific Book Chapters, in: Simple Adaptive Strategies From Regret-Matching to Uncoupled Dynamics, chapter 2, pages 17-46, World Scientific Publishing Co. Pte. Ltd..
- Sergiu Hart & Andreu Mas-Colell, 2000. "A Simple Adaptive Procedure Leading to Correlated Equilibrium," Econometrica, Econometric Society, vol. 68(5), pages 1127-1150, September.
- Sergiu Hart & Andreu Mas-Colell, 1996. "A simple adaptive procedure leading to correlated equilibrium," Economics Working Papers 200, Department of Economics and Business, Universitat Pompeu Fabra, revised Dec 1996.
- S. Hart & A. Mas-Collel, 2010. "A Simple Adaptive Procedure Leading to Correlated Equilibrium," Levine's Working Paper Archive 572, David K. Levine.
- Sergiu Hart & Andreu Mas-Colell, 1997. "A Simple Adaptive Procedure Leading to Correlated Equilibrium," Game Theory and Information 9703006, University Library of Munich, Germany, revised 25 Nov 1997.
Borgers, Tilman & Sarin, Rajiv, 1997. "Learning Through Reinforcement and Replicator Dynamics," Journal of Economic Theory, Elsevier, vol. 77(1), pages 1-14, November.
- Tilman Börgers & Rajiv Sarin, "undated". "Learning Through Reinforcement and Replicator Dynamics," ELSE working papers 051, ESRC Centre on Economics Learning and Social Evolution.
- T. Borgers & R. Sarin, 2010. "Learning Through Reinforcement and Replicator Dynamics," Levine's Working Paper Archive 380, David K. Levine.
Friedman, Daniel & Ostrov, Daniel N., 2008. "Conspicuous consumption dynamics," Games and Economic Behavior, Elsevier, vol. 64(1), pages 121-145, September.
Sandholm, William H., 2001. "Potential Games with Continuous Player Sets," Journal of Economic Theory, Elsevier, vol. 97(1), pages 81-108, March.
- Sandholm,W.H., 1999. "Potential games with continuous player sets," Working papers 23, Wisconsin Madison - Social Systems.
Sandholm, William H., 2007. "Evolution in Bayesian games II: Stability of purified equilibria," Journal of Economic Theory, Elsevier, vol. 136(1), pages 641-667, September.
- Sandholm,W.H., 2003. "Evolution in Bayesian games II : stability of purified equilibria," Working papers 21, Wisconsin Madison - Social Systems.
Ely, Jeffrey C. & Sandholm, William H., 2005. "Evolution in Bayesian games I: Theory," Games and Economic Behavior, Elsevier, vol. 53(1), pages 83-109, October.
Ellison, Glenn & Fudenberg, Drew, 2000. "Learning Purified Mixed Equilibria," Journal of Economic Theory, Elsevier, vol. 90(1), pages 84-115, January.
- Glenn Ellison & Drew Fudenberg, 1998. "Learning Purified Mixed Equilibria," Harvard Institute of Economic Research Working Papers 1817, Harvard - Institute of Economic Research.
Josef Hofbauer & Sylvain Sorin & Yannick Viossat, 2009. "Time Average Replicator and Best-Reply Dynamics," Mathematics of Operations Research, INFORMS, vol. 34(2), pages 263-269, May.
- Josef Hofbauer & Sylvain Sorin & Yannick Viossat, 2009. "Time Average Replicator and Best Reply Dynamics," Post-Print hal-00360767, HAL.
Ramsza, Michal & Seymour, Robert M., 2010. "Fictitious play in an evolutionary environment," Games and Economic Behavior, Elsevier, vol. 68(1), pages 303-324, January.
Friedman, Daniel & Ostrov, Daniel N., 2010. "Gradient dynamics in population games: Some basic results," Journal of Mathematical Economics, Elsevier, vol. 46(5), pages 691-707, September.
Hofbauer, Josef & Sandholm, William H., 2009. "Stable games and their dynamics," Journal of Economic Theory, Elsevier, vol. 144(4), pages 1665-1693.4, July.

Full references (including those not matched with items on IDEAS)

Citations

Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.

Cited by:

Dai Zusai, 2018. "Evolutionary dynamics in heterogeneous populations: a general framework for an arbitrary type distribution," Papers 1805.04897, arXiv.org, revised May 2019.
Lahkar, Ratul & Seymour, Robert M., 2014. "The dynamics of generalized reinforcement learning," Journal of Economic Theory, Elsevier, vol. 151(C), pages 584-595.
Dai Zusai, 2017. "Nonaggregable evolutionary dynamics under payoff heterogeneity," DETU Working Papers 1702, Department of Economics, Temple University.
V'ictor Gallego & Roi Naveiro & David R'ios Insua & Wolfram Rozas, 2021. "Data sharing games," Papers 2101.10721, arXiv.org.
Wei, Fangfang & Jia, Ning & Ma, Shoufeng, 2016. "Day-to-day traffic dynamics considering social interaction: From individual route choice behavior to a network flow model," Transportation Research Part B: Methodological, Elsevier, vol. 94(C), pages 335-354.
Karl D. Lewis & A. J. Shaiju, 2024. "Asymmetric Replicator Dynamics on Polish Spaces: Invariance, Stability, and Convergence," Dynamic Games and Applications, Springer, vol. 14(5), pages 1160-1190, November.
Lai, Chong & Li, Rui & Gao, Xiujuan, 2024. "Bank competition with technological innovation based on evolutionary games," International Review of Economics & Finance, Elsevier, vol. 89(PA), pages 742-759.

Most related items

These are the items that most often cite the same works as this one and are cited by the same works as this one.

Sandholm, William H., 2015. "Population Games and Deterministic Evolutionary Dynamics," Handbook of Game Theory with Economic Applications,, Elsevier.
Jonathan Newton, 2018. "Evolutionary Game Theory: A Renaissance," Games, MDPI, vol. 9(2), pages 1-67, May.
Fudenberg, Drew & Takahashi, Satoru, 2011. "Heterogeneous beliefs and local information in stochastic fictitious play," Games and Economic Behavior, Elsevier, vol. 71(1), pages 100-120, January.
- Drew Fudenberg & Satoru Takahashi, 2008. "Heterogeneous Beliefs and Local Information in Stochastic Fictitious Play," Levine's Working Paper Archive 122247000000001695, David K. Levine.
- Takahashi, Satoru & Fudenberg, Drew, 2011. "Heterogeneous beliefs and local information in stochastic fictitious play," Scholarly Articles 27755310, Harvard University Department of Economics.
Benaïm, Michel & Hofbauer, Josef & Hopkins, Ed, 2009. "Learning in games with unstable equilibria," Journal of Economic Theory, Elsevier, vol. 144(4), pages 1694-1709, July.
- Ed Hopkins & Josef Hofbauer & Michel Benaim, 2005. "Learning in Games with Unstable Equilibria," Edinburgh School of Economics Discussion Paper Series 135, Edinburgh School of Economics, University of Edinburgh.
- Michel Benaim & Josef Hofbauer & Ed Hopkins, 2006. "Learning in Games with Unstable Equilibria," Levine's Bibliography 321307000000000547, UCLA Department of Economics.
- Michel Benaim & Josef Hofbauer & Ed Hopkins, 2005. "Learning in Games with Unstable Equilibria," Levine's Bibliography 784828000000000609, UCLA Department of Economics.
Ed Hopkins, 2002. "Two Competing Models of How People Learn in Games," Econometrica, Econometric Society, vol. 70(6), pages 2141-2166, November.
- Ed Hopkins, 1999. "Two Competing Models of How People Learn in Games," Edinburgh School of Economics Discussion Paper Series 42, Edinburgh School of Economics, University of Edinburgh, revised Dec 2000.
- Ed Hopkins, 2001. "Two Competing Models of How People Learn in Games," NajEcon Working Paper Reviews 625018000000000226, www.najecon.org.
- Ed Hopkins, 2000. "Two Competing Models of How People Learn in Games," Edinburgh School of Economics Discussion Paper Series 51, Edinburgh School of Economics, University of Edinburgh, revised Dec 2000.
- Ed Hopkins, 2001. "Two Competing Models of How People Learn in Games," Levine's Working Paper Archive 625018000000000226, David K. Levine.
Panayotis Mertikopoulos & William H. Sandholm, 2016. "Learning in Games via Reinforcement and Regularization," Mathematics of Operations Research, INFORMS, vol. 41(4), pages 1297-1324, November.
Funai, Naoki, 2022. "Reinforcement learning with foregone payoff information in normal form games," Journal of Economic Behavior & Organization, Elsevier, vol. 200(C), pages 638-660.
Hopkins, Ed, 2007. "Adaptive learning models of consumer behavior," Journal of Economic Behavior & Organization, Elsevier, vol. 64(3-4), pages 348-368.
- Ed Hopkins, 2002. "Adaptive Learning Models of Consumer Behaviour," Edinburgh School of Economics Discussion Paper Series 80, Edinburgh School of Economics, University of Edinburgh.
- Ed Hopkins, 2004. "Adaptive Learning Models of Consumer Behaviour," Edinburgh School of Economics Discussion Paper Series 121, Edinburgh School of Economics, University of Edinburgh, revised Nov 2004.
- Ed Hopkins, 2006. "Adaptive Learning Models of Consumer Behaviour," Levine's Bibliography 122247000000000658, UCLA Department of Economics.
- Ed Hopkins, 2010. "Adaptive Learning Models of Consumer Behaviour," Levine's Working Paper Archive 506439000000000346, David K. Levine.
Dai Zusai, 2018. "Evolutionary dynamics in heterogeneous populations: a general framework for an arbitrary type distribution," Papers 1805.04897, arXiv.org, revised May 2019.
Naoki Funai, 2019. "Convergence results on stochastic adaptive learning," Economic Theory, Springer;Society for the Advancement of Economic Theory (SAET), vol. 68(4), pages 907-934, November.
Oyarzun, Carlos & Sarin, Rajiv, 2013. "Learning and risk aversion," Journal of Economic Theory, Elsevier, vol. 148(1), pages 196-225.
- Carlos Oyarzun & Rajiv Sarin, 2005. "Learning and Risk Aversion," Levine's Bibliography 784828000000000482, UCLA Department of Economics.
- Carlos Oyarzun & Rajiv Sarin, 2012. "Learning and Risk Aversion," Levine's Working Paper Archive 786969000000000572, David K. Levine.
Ed Hopkins & Robert M. Seymour, 2002. "The Stability of Price Dispersion under Seller and Consumer Learning," International Economic Review, Department of Economics, University of Pennsylvania and Osaka University Institute of Social and Economic Research Association, vol. 43(4), pages 1157-1190, November.
- Ed Hopkins & Robert M Seymour, 1999. "The Stability of Price Dispersion under Seller and Consumer Learning," Edinburgh School of Economics Discussion Paper Series 45, Edinburgh School of Economics, University of Edinburgh, revised Dec 2000.
- Ed Hopkins & Roberty M. Seymour, 2002. "The Stability of Price Dispersion under Seller and Consumer Learning," Game Theory and Information 0203002, University Library of Munich, Germany.
- Ed Hopkins & Robert M Seymour, 2000. "The Stability of Price Dispersion under Seller and Consumer Learning," Edinburgh School of Economics Discussion Paper Series 52, Edinburgh School of Economics, University of Edinburgh, revised Dec 2000.
Dai Zusai, 2018. "Tempered best response dynamics," International Journal of Game Theory, Springer;Game Theory Society, vol. 47(1), pages 1-34, March.
Mertikopoulos, Panayotis & Sandholm, William H., 2018. "Riemannian game dynamics," Journal of Economic Theory, Elsevier, vol. 177(C), pages 315-364.
Mengel, Friederike, 2012. "Learning across games," Games and Economic Behavior, Elsevier, vol. 74(2), pages 601-619.
- Friederike Mengel, 2007. "Learning Across Games," Working Papers. Serie AD 2007-05, Instituto Valenciano de Investigaciones Económicas, S.A. (Ivie).
Ulrich Doraszelski & Gregory Lewis & Ariel Pakes, 2018. "Just Starting Out: Learning and Equilibrium in a New Market," American Economic Review, American Economic Association, vol. 108(3), pages 565-615, March.
- Ulrich Doraszelski & Gregory Lewis & Ariel Pakes, 2016. "Just Starting Out: Learning and Equilibrium in a New Market," NBER Working Papers 21996, National Bureau of Economic Research, Inc.
Tassos Patokos, 2014. "Introducing Disappointment Dynamics and Comparing Behaviors in Evolutionary Games: Some Simulation Results," Games, MDPI, vol. 5(1), pages 1-25, January.
Lahkar, Ratul & Seymour, Robert M., 2014. "The dynamics of generalized reinforcement learning," Journal of Economic Theory, Elsevier, vol. 151(C), pages 584-595.
Mario Bravo & Mathieu Faure, 2013. "Reinforcement Learning with Restrictions on the Action Set," AMSE Working Papers 1335, Aix-Marseille School of Economics, France, revised 01 Jul 2013.
- Mario Bravo & Mathieu Faure, 2015. "Reinforcement Learning with Restrictions on the Action Set," Post-Print hal-01457301, HAL.
Sandholm, William H., 2007. "Evolution in Bayesian games II: Stability of purified equilibria," Journal of Economic Theory, Elsevier, vol. 136(1), pages 641-667, September.
- Sandholm,W.H., 2003. "Evolution in Bayesian games II : stability of purified equilibria," Working papers 21, Wisconsin Madison - Social Systems.

More about this item

Keywords

Reinforcement learning; Continuity equation; Replicator dynamics;
All these keywords.

JEL classification:

C72 - Mathematical and Quantitative Methods - - Game Theory and Bargaining Theory - - - Noncooperative Games
C73 - Mathematical and Quantitative Methods - - Game Theory and Bargaining Theory - - - Stochastic and Dynamic Games; Evolutionary Games

Statistics

Access and download statistics

Corrections

All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:gamebe:v:80:y:2013:i:c:p:10-38. See general information about how to correct material in RePEc.

If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/inca/622836 .

Please note that corrections may take a couple of weeks to filter through the various RePEc services.

IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.

Browse Econ Literature

More features

Reinforcement learning in population games

Author

Abstract

Suggested Citation

Download full text from publisher

References listed on IDEAS

Citations

Most related items

More about this item

Keywords

JEL classification:

Statistics

Corrections

More services and features

MyIDEAS

Author registration

Rankings

RePEc Genealogy

RePEc Biblio

MPRA

New papers by email

EconAcademics

Plagiarism

About RePEc

RePEc home

Blog

Help/FAQ

RePEc team

Participating archives

Privacy statement

Help us

Corrections

Volunteers

Get papers listed

Open a RePEc archive

Get RePEc data