Learning to be Indifferent in Complex Decisions: A Coarse Payoff-Assessment Model

My bibliography Save this paper

Learning to be Indifferent in Complex Decisions: A Coarse Payoff-Assessment Model

Author

Listed:

Philippe Jehiel
Aviman Satpathy

Registered:

Philippe Jehiel

Abstract

We introduce the Coarse Payoff-Assessment Learning (CPAL) model, which captures reinforcement learning by boundedly rational decision-makers who focus on the aggregate outcomes of choosing among exogenously defined clusters of alternatives (similarity classes), rather than evaluating each alternative individually. Analyzing a smooth approximation of the model, we show that the learning dynamics exhibit steady-states corresponding to smooth Valuation Equilibria (Jehiel and Samet, 2007). We demonstrate the existence of multiple equilibria in decision trees with generic payoffs and establish the local asymptotic stability of pure equilibria when they occur. Conversely, when trivial choices featuring alternatives within the same similarity class yield sufficiently high payoffs, a unique mixed equilibrium emerges, characterized by indifferences between similarity classes, even under acute sensitivity to payoff differences. Finally, we prove that this unique mixed equilibrium is globally asymptotically stable under the CPAL dynamics.

Suggested Citation

Philippe Jehiel & Aviman Satpathy, 2024. "Learning to be Indifferent in Complex Decisions: A Coarse Payoff-Assessment Model," Papers 2412.09321, arXiv.org, revised Dec 2024.

Handle: RePEc:arx:papers:2412.09321

Download full text from publisher

References listed on IDEAS

Pedro Bordalo & Nicola Gennaioli & Andrei Shleifer, 2013. "Salience and Consumer Choice," Journal of Political Economy, University of Chicago Press, vol. 121(5), pages 803-843.
- Pedro Bordalo & Nicola Gennaioli & Andrei Shleifer, "undated". "Salience and Consumer Choice," Working Paper 62321, Harvard University OpenScholar.
- Pedro Bordalo & Nicola Gennaioli & Andrei Shleifer, 2012. "Salience and Consumer Choice," NBER Working Papers 17947, National Bureau of Economic Research, Inc.
- Bordalo, Pedro & Gennaioli, Nicola & Shleifer, Andrei, 2013. "Salience and Consumer Choice," Scholarly Articles 27814563, Harvard University Department of Economics.
- Pedro Bordado & Nicola Gennaioli & Andrei Shleifer, 2015. "Salience and Consumer Choice," Working Papers 501, Barcelona School of Economics.
- Pedro Bordalo & Nicola Gennaioli & Andrei Shleifer, 2012. "Salience and Consumer Choice," Working Papers 463, IGIER (Innocenzo Gasparini Institute for Economic Research), Bocconi University.
- Pedro Bordalo & Nicola Gennaioli & Andrei Shleifer, 2010. "Salience and consumer choice," Economics Working Papers 1252, Department of Economics and Business, Universitat Pompeu Fabra, revised May 2012.
Fudenberg Drew & Kreps David M., 1993. "Learning Mixed Equilibria," Games and Economic Behavior, Elsevier, vol. 5(3), pages 320-367, July.
- Fudenberg, D. & Kreps, D.M., 1992. "Learning Mixed Equilibria," Working papers 92-13, Massachusetts Institute of Technology (MIT), Department of Economics.
- Drew Fudenberg & David Kreps, 2010. "Learning Mixed Equilibria," Levine's Working Paper Archive 415, David K. Levine.
Erev, Ido & Roth, Alvin E, 1998. "Predicting How People Play Games: Reinforcement Learning in Experimental Games with Unique, Mixed Strategy Equilibria," American Economic Review, American Economic Association, vol. 88(4), pages 848-881, September.
Cominetti, Roberto & Melo, Emerson & Sorin, Sylvain, 2010. "A payoff-based learning procedure and its application to traffic games," Games and Economic Behavior, Elsevier, vol. 70(1), pages 71-83, September.
, & ,, 2007. "Valuation equilibrium," Theoretical Economics, Econometric Society, vol. 2(2), June.
- Philippe Jehiel & Dov Samet, 2003. "Valuation Equilibria," Game Theory and Information 0310003, University Library of Munich, Germany.
- Philippe Jehiel & Dov Samet, 2007. "Valuation Equilibrium," PSE-Ecole d'économie de Paris (Postprint) halshs-00754229, HAL.
- Philippe Jehiel & Dov Samet, 2007. "Valuation Equilibrium," Post-Print halshs-00754229, HAL.
- Philippe Jehiel & Dov Samet, 2006. "Valuation Equilibria," Levine's Bibliography 784828000000000111, UCLA Department of Economics.
- Philippe Jehiel & Dov Samet, 2003. "Valuation Equilibria," Levine's Bibliography 666156000000000046, UCLA Department of Economics.
Jehiel, Philippe, 2005. "Analogy-based expectation equilibrium," Journal of Economic Theory, Elsevier, vol. 123(2), pages 81-104, August.
- Philippe Jeniel, 2001. "Analogy-Based Expectation Equilibrium," Economics Working Papers 0003, Institute for Advanced Study, School of Social Science.
- Philippe Jehiel, 2005. "Analogy-Based Expectation Equilibrium," Levine's Bibliography 784828000000000106, UCLA Department of Economics.
- Philippe Jehiel, 2005. "Analogy-based Expectation Equilibrium," Post-Print halshs-00754070, HAL.
Jehiel, Philippe & Singh, Juni, 2021. "Multi-state choices with aggregate feedback on unfamiliar alternatives," Games and Economic Behavior, Elsevier, vol. 130(C), pages 1-24.
- Philippe Jehiel & Juni Singh, 2019. "Multi-state choices with aggregate feedback on unfamiliar alternatives," PSE Working Papers halshs-02183444, HAL.
- Philippe Jehiel & Juni Singh, 2021. "Multi-state choices with aggregate feedback on unfamiliar alternatives," PSE-Ecole d'économie de Paris (Postprint) halshs-03672197, HAL.
- Philippe Jehiel & Juni Singh, 2021. "Multi-state choices with aggregate feedback on unfamiliar alternatives," Post-Print halshs-03672197, HAL.
- Philippe Jehiel & Juni Singh, 2019. "Multi-state choices with aggregate feedback on unfamiliar alternatives," Working Papers halshs-02183444, HAL.
Michel Benaim & Josef Hofbauer & Sylvain Sorin, 2005. "Stochastic Approximations and Differential Inclusions II: Applications," Levine's Bibliography 784828000000000098, UCLA Department of Economics.
Hausman, Jerry & McFadden, Daniel, 1984. "Specification Tests for the Multinomial Logit Model," Econometrica, Econometric Society, vol. 52(5), pages 1219-1240, September.
- D. McFadden & J. Hausman, 1981. "Specification Tests for the Multinominal Logit Model," Working papers 292, Massachusetts Institute of Technology (MIT), Department of Economics.
Borgers, Tilman & Sarin, Rajiv, 1997. "Learning Through Reinforcement and Replicator Dynamics," Journal of Economic Theory, Elsevier, vol. 77(1), pages 1-14, November.
- Tilman Börgers & Rajiv Sarin, "undated". "Learning Through Reinforcement and Replicator Dynamics," ELSE working papers 051, ESRC Centre on Economics Learning and Social Evolution.
- T. Borgers & R. Sarin, 2010. "Learning Through Reinforcement and Replicator Dynamics," Levine's Working Paper Archive 380, David K. Levine.
Pedro Bordalo & Nicola Gennaioli & Andrei Shleifer, 2012. "Salience Theory of Choice Under Risk," The Quarterly Journal of Economics, President and Fellows of Harvard College, vol. 127(3), pages 1243-1285.
- Pedro Bordalo & Nicola Gennaioli & Andrei Shleifer, "undated". "Salience Theory of Choice Under Risk," Working Paper 29210, Harvard University OpenScholar.
- Andrei Shleifer & Nicola Gennaioli & Pedro Bordalo, 2011. "Salience theory of choice under risk," 2011 Meeting Papers 1442, Society for Economic Dynamics.
- Shleifer, Andrei & Bordalo, Pedro & Gennaioli, Nicola, 2012. "Salience Theory of Choice Under Risk," Scholarly Articles 10636303, Harvard University Department of Economics.
- Pedro Bordalo & Nicola Gennaioli & Andrei Shleifer, 2010. "Salience Theory of Choice Under Risk," NBER Working Papers 16387, National Bureau of Economic Research, Inc.
Josef Hofbauer & William H. Sandholm, 2002. "On the Global Convergence of Stochastic Fictitious Play," Econometrica, Econometric Society, vol. 70(6), pages 2265-2294, November.
Stefano DellaVigna, 2009. "Psychology and Economics: Evidence from the Field," Journal of Economic Literature, American Economic Association, vol. 47(2), pages 315-372, June.
- Stefano DellaVigna, 2007. "Psychology and Economics: Evidence from the Field," NBER Working Papers 13420, National Bureau of Economic Research, Inc.
Roth, Alvin E. & Erev, Ido, 1995. "Learning in extensive-form games: Experimental data and simple dynamic models in the intermediate term," Games and Economic Behavior, Elsevier, vol. 8(1), pages 164-212.
Jacob K. Goeree & Charles A. Holt & Thomas R. Palfrey, 2016. "Quantal Response Equilibrium:A Stochastic Theory of Games," Economics Books, Princeton University Press, edition 1, number 10743.
Nachbar, J H, 1990. ""Evolutionary" Selection Dynamics in Games: Convergence and Limit Properties," International Journal of Game Theory, Springer;Game Theory Society, vol. 19(1), pages 59-89.
Sarin, Rajiv & Vahid, Farshid, 1999. "Payoff Assessments without Probabilities: A Simple Dynamic Model of Choice," Games and Economic Behavior, Elsevier, vol. 28(2), pages 294-309, August.
Monderer, Dov & Shapley, Lloyd S., 1996. "Fictitious Play Property for Games with Identical Interests," Journal of Economic Theory, Elsevier, vol. 68(1), pages 258-265, January.

Full references (including those not matched with items on IDEAS)

Most related items

These are the items that most often cite the same works as this one and are cited by the same works as this one.

Funai, Naoki, 2022. "Reinforcement learning with foregone payoff information in normal form games," Journal of Economic Behavior & Organization, Elsevier, vol. 200(C), pages 638-660.
Naoki Funai, 2019. "Convergence results on stochastic adaptive learning," Economic Theory, Springer;Society for the Advancement of Economic Theory (SAET), vol. 68(4), pages 907-934, November.
Leslie, David S. & Collins, E.J., 2006. "Generalised weakened fictitious play," Games and Economic Behavior, Elsevier, vol. 56(2), pages 285-298, August.
Panayotis Mertikopoulos & William H. Sandholm, 2016. "Learning in Games via Reinforcement and Regularization," Mathematics of Operations Research, INFORMS, vol. 41(4), pages 1297-1324, November.
Benaïm, Michel & Hofbauer, Josef & Hopkins, Ed, 2009. "Learning in games with unstable equilibria," Journal of Economic Theory, Elsevier, vol. 144(4), pages 1694-1709, July.
- Ed Hopkins & Josef Hofbauer & Michel Benaim, 2005. "Learning in Games with Unstable Equilibria," Edinburgh School of Economics Discussion Paper Series 135, Edinburgh School of Economics, University of Edinburgh.
- Michel Benaim & Josef Hofbauer & Ed Hopkins, 2006. "Learning in Games with Unstable Equilibria," Levine's Bibliography 321307000000000547, UCLA Department of Economics.
- Michel Benaim & Josef Hofbauer & Ed Hopkins, 2005. "Learning in Games with Unstable Equilibria," Levine's Bibliography 784828000000000609, UCLA Department of Economics.
Hopkins, Ed, 1999. "Learning, Matching, and Aggregation," Games and Economic Behavior, Elsevier, vol. 26(1), pages 79-110, January.
- Ed Hopkins, "undated". "Learning, Matching and Aggregation," Discussion Papers 1996-2, Edinburgh School of Economics, University of Edinburgh.
- Hopkins, E., 1995. "Learning, Matching and Aggregation," G.R.E.Q.A.M. 95a20, Universite Aix-Marseille III.
- Ed Hopkins, 1995. "Learning, Matching and Aggregation," Edinburgh School of Economics Discussion Paper Series 2, Edinburgh School of Economics, University of Edinburgh.
- Ed Hopkins, "undated". "Learning, Matching and Aggregation," ELSE working papers 033, ESRC Centre on Economics Learning and Social Evolution.
- Ed Hopkins, 1995. "Learning, Matching and Aggregation," Game Theory and Information 9512001, University Library of Munich, Germany.
- Ed Hopkins, "undated". "Learning, Matching and Aggregation," Department of Economics 1996 : II, Edinburgh School of Economics, University of Edinburgh.
Jonathan Newton, 2018. "Evolutionary Game Theory: A Renaissance," Games, MDPI, vol. 9(2), pages 1-67, May.
Daskalova, Vessela & Vriend, Nicolaas J., 2021. "Learning frames," Journal of Economic Behavior & Organization, Elsevier, vol. 191(C), pages 78-96.
- Vessela Daskalova & Nicolaas J. Vriend, 2021. "Learning Frames," Working Papers 202118, School of Economics, University College Dublin.
- Vessela Daskalova & Nicolaas J.Vriend, 2021. "Learning frames," Working Papers 929, Queen Mary University of London, School of Economics and Finance.
Jakub Bielawski & Thiparat Chotibut & Fryderyk Falniowski & Michal Misiurewicz & Georgios Piliouras, 2022. "Unpredictable dynamics in congestion games: memory loss can prevent chaos," Papers 2201.10992, arXiv.org, revised Jan 2022.
Pangallo, Marco & Sanders, James B.T. & Galla, Tobias & Farmer, J. Doyne, 2022. "Towards a taxonomy of learning dynamics in 2 × 2 games," Games and Economic Behavior, Elsevier, vol. 132(C), pages 1-21.
- Marco Pangallo & James Sanders & Tobias Galla & Doyne Farmer, 2017. "Towards a taxonomy of learning dynamics in 2 x 2 games," Papers 1701.09043, arXiv.org, revised Sep 2021.
Mertikopoulos, Panayotis & Sandholm, William H., 2024. "Nested replicator dynamics, nested logit choice, and similarity-based learning," Journal of Economic Theory, Elsevier, vol. 220(C).
Sandholm, William H., 2015. "Population Games and Deterministic Evolutionary Dynamics," Handbook of Game Theory with Economic Applications,, Elsevier.
Bravo, Mario & Mertikopoulos, Panayotis, 2017. "On the robustness of learning in games with stochastically perturbed payoff observations," Games and Economic Behavior, Elsevier, vol. 103(C), pages 41-66.
Ianni, A., 2002. "Reinforcement learning and the power law of practice: some analytical results," Discussion Paper Series In Economics And Econometrics 203, Economics Division, School of Social Sciences, University of Southampton.
DeJong, D.V. & Blume, A. & Neumann, G., 1998. "Learning in Sender-Receiver Games," Other publications TiSEM 4a8b4f46-f30b-4ad2-bb0c-1, Tilburg University, School of Economics and Management.
- Blume, A. & DeJong, D.V. & Neumann, G.R. & Savin, N.E., 1998. "Learning in Sender-Receiver Games," Working Papers 98-02, University of Iowa, Department of Economics.
- DeJong, D.V. & Blume, A. & Neumann, G., 1998. "Learning in Sender-Receiver Games," Discussion Paper 1998-28, Tilburg University, Center for Economic Research.
- Andreas Blume & Douglas V. DeJong & George R. Neumann & Nathan E. Savin, 1998. "Learning in Sender-Receiver Games," CIG Working Papers FS IV 98-13, Wissenschaftszentrum Berlin (WZB), Research Unit: Competition and Innovation (CIG).
Antonio Cabrales & Giovanni Ponti, 2000. "Implementation, Elimination of Weakly Dominated Strategies and Evolutionary Dynamics," Review of Economic Dynamics, Elsevier for the Society for Economic Dynamics, vol. 3(2), pages 247-282, April.
Cason, Timothy N. & Friedman, Daniel & Hopkins, Ed, 2010. "Testing the TASP: An experimental investigation of learning in games with unstable equilibria," Journal of Economic Theory, Elsevier, vol. 145(6), pages 2309-2331, November.
- Timothy N. Cason & Daniel Friedman & Ed Hopkins, 2009. "Testing the TASP: An Experimental Investigation of Learning in Games with Unstable Equilibria," Edinburgh School of Economics Discussion Paper Series 188, Edinburgh School of Economics, University of Edinburgh.
- Cason, Timothy N. & Friedman, Daniel & Hopkins, Ed H, 2009. "Testing the TASP: An Experimental Investigation of Learning in Games with Unstable Equilibria," Santa Cruz Department of Economics, Working Paper Series qt8kp6c049, Department of Economics, UC Santa Cruz.
- Timothy N. Cason & Daniel Friedman & Ed Hopkins, 2010. "Testing the TASP: An Experimental Investigation of Learning in Games with Unstable Equilibria," Purdue University Economics Working Papers 1233, Purdue University, Department of Economics.
- Cason, Timothy N. & Friedman, Daniel UC & Hopkins, Ed, 2009. "Testing the TASP: An Experimental Investigation of Learning in Games with Unstable Equilibria," SIRE Discussion Papers 2009-15, Scottish Institute for Research in Economics (SIRE).
Mohlin, Erik & Östling, Robert & Wang, Joseph Tao-yi, 2020. "Learning by similarity-weighted imitation in winner-takes-all games," Games and Economic Behavior, Elsevier, vol. 120(C), pages 225-245.
Ed Hopkins, 2002. "Two Competing Models of How People Learn in Games," Econometrica, Econometric Society, vol. 70(6), pages 2141-2166, November.
- Ed Hopkins, 1999. "Two Competing Models of How People Learn in Games," Edinburgh School of Economics Discussion Paper Series 42, Edinburgh School of Economics, University of Edinburgh, revised Dec 2000.
- Ed Hopkins, 2001. "Two Competing Models of How People Learn in Games," NajEcon Working Paper Reviews 625018000000000226, www.najecon.org.
- Ed Hopkins, 2000. "Two Competing Models of How People Learn in Games," Edinburgh School of Economics Discussion Paper Series 51, Edinburgh School of Economics, University of Edinburgh, revised Dec 2000.
- Ed Hopkins, 2001. "Two Competing Models of How People Learn in Games," Levine's Working Paper Archive 625018000000000226, David K. Levine.
Funai Naoki, 2014. "An Adaptive Learning Model with Foregone Payoff Information," The B.E. Journal of Theoretical Economics, De Gruyter, vol. 14(1), pages 149-176, January.

More about this item

NEP fields

This paper has been announced in the following NEP Reports:

NEP-MIC-2025-01-27 (Microeconomics)

Statistics

Access and download statistics

Corrections

All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:arx:papers:2412.09321. See general information about how to correct material in RePEc.

If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: arXiv administrators (email available below). General contact details of provider: http://arxiv.org/ .

Please note that corrections may take a couple of weeks to filter through the various RePEc services.

IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.

Browse Econ Literature

More features

Learning to be Indifferent in Complex Decisions: A Coarse Payoff-Assessment Model

Author

Abstract

Suggested Citation

Download full text from publisher

References listed on IDEAS

Most related items

More about this item

NEP fields

Statistics

Corrections

More services and features

MyIDEAS

Author registration

Rankings

RePEc Genealogy

RePEc Biblio

MPRA

New papers by email

EconAcademics

Plagiarism

About RePEc

RePEc home

Blog

Help/FAQ

RePEc team

Participating archives

Privacy statement

Help us

Corrections

Volunteers

Get papers listed

Open a RePEc archive

Get RePEc data