Inefficient Learning Leads to Efficient Coordination: A Repeated Stag-Hunt Game with Learning Agents
Abstract
1. Introduction
2. Materials and Methods
2.1. General Configuration
2.2. Dynamics
2.2.1. Initial Conditions
2.2.2. Update Rule
2.2.3. Error Probability
2.3. Simulations
3. Results
3.1. Baseline Model
3.2. Learning Rate and Error Probability
3.2.1. Learning Rate Effect
3.2.2. Error Probability Effect
3.2.3. Combined Effect of Learning Rate and Error Probability
3.3. Comparison with Experimental Results
4. Discussion
Author Contributions
Funding
Informed Consent Statement
Data Availability Statement
Acknowledgments
Conflicts of Interest
Abbreviations
| SH | Stag-hunt game |
| Learning rate | |
| Error probability | |
| Probability that agents coordinate on Stag | |
| Stag’s payoff |
Appendix A

Appendix B

Appendix C
| 0 | 0.1 | 0.2 | 0.3 | 0.4 | 0.5 | |
| 0 | 0.023 | 0.002 | 0.011 | 0.035 | 0.079 | 0.129 |
| (0.149) | (0.006) | (0.026) | (0.059) | (0.104) | (0.138) | |
| 0.1 | 0.002 | 0.005 | 0.022 | 0.074 | 0.167 | 0.235 |
| (0.007) | (0.009) | (0.030) | (0.075) | (0.134) | (0.146) | |
| 0.2 | 0.011 | 0.023 | 0.050 | 0.068 | 0.131 | 0.195 |
| (0.024) | (0.030) | (0.075) | (0.075) | (0.113) | (0.145) | |
| 0.3 | 0.034 | 0.074 | 0.067 | 0.089 | 0.138 | 0.179 |
| (0.060) | (0.077) | (0.071) | (0.083) | (0.111) | (0.131) | |
| 0.4 | 0.079 | 0.157 | 0.135 | 0.136 | 0.156 | 0.167 |
| (0.108) | (0.125) | (0.117) | (0.110) | (0.117) | (0.130) | |
| 0.5 | 0.132 | 0.244 | 0.213 | 0.177 | 0.166 | 0.162 |
| (0.142) | (0.147) | (0.147) | (0.134) | (0.126) | (0.121) | |
| , | ||||||
| 0 | 0.1 | 0.2 | 0.3 | 0.4 | 0.5 | |
| 0 | 0.020 | 0.020 | 0.036 | 0.066 | 0.123 | 0.185 |
| (0.110) | (0.049) | (0.045) | (0.042) | (0.053) | (0.066) | |
| 0.1 | 0.006 | 0.005 | 0.019 | 0.069 | 0.153 | 0.229 |
| (0.050) | (0.006) | (0.013) | (0.037) | (0.070) | (0.074) | |
| 0.2 | 0.018 | 0.020 | 0.031 | 0.053 | 0.126 | 0.204 |
| (0.025) | (0.026) | (0.040) | (0.034) | (0.060) | (0.071) | |
| 0.3 | 0.050 | 0.064 | 0.059 | 0.061 | 0.102 | 0.162 |
| (0.056) | (0.062) | (0.064) | (0.052) | (0.069) | (0.088) | |
| 0.4 | 0.114 | 0.137 | 0.133 | 0.106 | 0.101 | 0.124 |
| (0.100) | (0.104) | (0.103) | (0.089) | (0.080) | (0.084) | |
| 0.5 | 0.173 | 0.210 | 0.208 | 0.161 | 0.122 | 0.114 |
| (0.122) | (0.123) | (0.131) | (0.116) | (0.097) | (0.084) | |
| 0 | 0.1 | 0.2 | 0.3 | 0.4 | 0.5 | |
| 0 | 0.016 | 0.030 | 0.064 | 0.103 | 0.140 | 0.189 |
| (0.065) | (0.062) | (0.061) | (0.059) | (0.063) | (0.059) | |
| 0.1 | 0.034 | 0.022 | 0.042 | 0.084 | 0.126 | 0.182 |
| (0.080) | (0.033) | (0.037) | (0.044) | (0.047) | (0.063) | |
| 0.2 | 0.065 | 0.044 | 0.039 | 0.060 | 0.099 | 0.142 |
| (0.070) | (0.039) | (0.036) | (0.043) | (0.050) | (0.059) | |
| 0.3 | 0.105 | 0.084 | 0.062 | 0.050 | 0.065 | 0.088 |
| (0.067) | (0.044) | (0.044) | (0.035) | (0.044) | (0.052) | |
| 0.4 | 0.139 | 0.126 | 0.100 | 0.064 | 0.050 | 0.051 |
| (0.061) | (0.046) | (0.052) | (0.044) | (0.036) | (0.038) | |
| 0.5 | 0.194 | 0.181 | 0.142 | 0.086 | 0.054 | 0.045 |
| (0.066) | (0.062) | (0.061) | (0.049) | (0.039) | (0.032) | |
References
- Arad, A., & Rubinstein, A. (2012). The 11–20 money request game: A level-k reasoning study. American Economic Review, 102(7), 3561–3573. [Google Scholar] [CrossRef] [Scilit]
- Axelrod, R. (1981). The emergence of cooperation among egoists. American Political Science Review, 75(2), 306–318. [Google Scholar] [CrossRef] [Scilit]
- Bardsley, N., Mehta, J., Starmer, C., & Sugden, R. (2010). Explaining focal points: Cognitive hierarchy theory versus team reasoning. The Economic Journal, 120(543), 40–79. [Google Scholar]
- Battalio, R., Samuelson, L., & Van Huyck, J. (2001). Optimization incentives and coordination failure in laboratory stag hunt games. Econometrica, 69(3), 749–764. [Google Scholar] [CrossRef] [Scilit]
- Belloc, M., Bilancini, E., Boncinelli, L., & D’Alessandro, S. (2019). Intuition and deliberation in the stag hunt game. Scientific Reports, 9(1), 14833. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Bilancini, E., Boncinelli, L., & Nax, H. H. (2021). What noise matters? Experimental evidence for stochastic deviations in social norms. Journal of Behavioral and Experimental Economics, 90, 101626. [Google Scholar] [CrossRef] [Scilit]
- Camerer, C. F., Ho, T.-H., & Chong, J.-K. (2004). A cognitive hierarchy model of games. The Quarterly Journal of Economics, 119(3), 861–898. [Google Scholar] [CrossRef] [Scilit]
- Challet, D., & Zhang, Y.-C. (1997). Emergence of cooperation and organization in an evolutionary game. Physica A: Statistical Mechanics and Its Applications, 246(3–4), 407–418. [Google Scholar] [CrossRef] [Scilit]
- Chen, R., & Chen, Y. (2011). The potential of social identity for equilibrium selection. American Economic Review, 101(6), 2562–2589. [Google Scholar] [CrossRef] [Scilit]
- Cooper, D. J., & Weber, R. A. (2020). Recent advances in experimental coordination games. In Handbook of experimental game theory (pp. 149–183). Edward Elgar Publishing. [Google Scholar] [CrossRef] [Scilit]
- De Kwaadsteniet, E. W., & Van Dijk, E. (2010). Social status as a cue for tacit coordination. Journal of Experimental Social Psychology, 46(3), 515–524. [Google Scholar] [CrossRef] [Scilit]
- Devetag, G., & Ortmann, A. (2014). Chapter 16—Solving coordination problems experimentally. In M. Webster, & J. Sell (Eds.), Laboratory experiments in the social sciences (2nd ed., pp. 357–384). Academic Press. Available online: https://www.sciencedirect.com/science/article/pii/B9780124046818000169 (accessed on 18 June 2026).
- Engelmann, D., & Normann, H.-T. (2010). Maximum effort in the minimum-effort game. Experimental Economics, 13(3), 249–259. [Google Scholar] [CrossRef] [Scilit]
- Fallucchi, F., & Nosenzo, D. (2022). The coordinating power of social norms. Experimental Economics, 25(1), 1–25. [Google Scholar] [CrossRef] [Scilit]
- Flache, A., & Macy, M. W. (2002). Stochastic collusion and the power law of learning: A general reinforcement learning model of cooperation. Journal of Conflict Resolution, 46(5), 629–653. [Google Scholar] [CrossRef] [Scilit]
- Hofbauer, J., & Sigmund, K. (1998). Evolutionary games and population dynamics. Cambridge University Press. [Google Scholar]
- Krupka, E. L., Weber, R., Crosno, R. T., & Hoover, H. (2022). “When in Rome”: Identifying social norms usingcoordination games. Judgment and Decision Making, 17(2), 263–283. [Google Scholar] [CrossRef] [Scilit]
- Krupka, E. L., & Weber, R. A. (2013). Identifying social norms using coordination games: Why does dictator game sharing vary? Journal of the European Economic Association, 11(3), 495–524. [Google Scholar] [CrossRef] [Scilit]
- Macy, M. W., & Flache, A. (2002). Learning dynamics in social dilemmas. Proceedings of the National Academy of Sciences of the United States of America, 99(Suppl. S3), 7229–7236. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Mäs, M., & Nax, H. H. (2016). A behavioral study of “noise” in coordination games. Journal of Economic Theory, 162, 195–208. [Google Scholar] [CrossRef] [Scilit]
- Nagel, R. (1995). Unraveling in guessing games: An experimental study. The American Economic Review, 85(5), 1313–1326. [Google Scholar]
- Raducha, T., & San Miguel, M. (2022). Coordination and equilibrium selection in games: The role of local effects. Scientific Reports, 12(1), 3373. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Roca, C. P., Cuesta, J. A., & Sánchez, A. (2009a). Effect of spatial structure on the evolution of cooperation. Physical Review E—Statistical, Nonlinear, and Soft Matter Physics, 80(4), 046106. [Google Scholar] [CrossRef] [Scilit]
- Roca, C. P., Cuesta, J. A., & Sánchez, A. (2009b). Evolutionary game theory: Temporal and spatial effects beyond replicator dynamics. Physics of Life Reviews, 6(4), 208–249. [Google Scholar] [CrossRef] [Scilit]
- Roth, A. E., & Erev, I. (1995). Learning in extensive-form games: Experimental data and simple dynamic models in the intermediate term. Games and Economic Behavior, 8(1), 164–212. [Google Scholar] [CrossRef] [Scilit]
- Rubinstein, A. (2016). A typology of players: Between instinctive and contemplative. The Quarterly Journal of Economics, 131(2), 859–890. [Google Scholar] [CrossRef] [Scilit]
- Silva, R. (2024). Coordination in stag hunt games. Journal of Behavioral and Experimental Economics, 113, 102290. [Google Scholar] [CrossRef] [Scilit]
- Skyrms, B. (2001). The stag hunt (Vol. 75, No. 2, pp. 31–41). American Philosophical Association. [Google Scholar]
- Skyrms, B. (2004). The stag hunt and the evolution of social structure. Cambridge University Press. [Google Scholar]
- Stahl, D. O., & Wilson, P. W. (1995). On players’ models of other players: Theory and experimental evidence. Games and Economic Behavior, 10(1), 218–254. [Google Scholar] [CrossRef] [Scilit]
- Tomassini, M., & Pestelacci, E. (2010). Coordination games on dynamical networks. Games, 1(3), 242–261. [Google Scholar] [CrossRef] [Scilit]
- Van Huyck, J. B., Battalio, R. C., & Beil, R. O. (1990). Tacit coordination games, strategic uncertainty, and coordination failure. The American Economic Review, 80(1), 234–248. [Google Scholar]
- Vilone, D., Realpe-Gómez, J., & Andrighetto, G. (2021). Evolutionary advantages of turning points in human cooperative behaviour. PLoS ONE, 16(2), e0246278. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Weidenholzer, S. (2010). Coordination games and local interactions: A survey of the game theoretic literature. Games, 1(4), 551–585. [Google Scholar] [CrossRef] [Scilit]








| Stag | Hare | |
|---|---|---|
| Stag | 0 | |
| Hare | r | r |
| t | T | |||||
|---|---|---|---|---|---|---|
| 0 | 0 | 4 |
| t | T | |||||
|---|---|---|---|---|---|---|
| 1 | 1 | 4 |
| Stag | Hare | |
|---|---|---|
| Stag | ||
| Hare |
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content. |
© 2026 by the authors. Licensee MDPI, Basel, Switzerland. This article is an open access article distributed under the terms and conditions of the Creative Commons Attribution (CC BY) license.
Share and Cite
Manfredi, R.; Vilone, D.; Cvetkovic, T.J.; Bagnoli, F.; Guazzini, A. Inefficient Learning Leads to Efficient Coordination: A Repeated Stag-Hunt Game with Learning Agents. Games 2026, 17, 39. https://doi.org/10.3390/g17040039
Manfredi R, Vilone D, Cvetkovic TJ, Bagnoli F, Guazzini A. Inefficient Learning Leads to Efficient Coordination: A Repeated Stag-Hunt Game with Learning Agents. Games. 2026; 17(4):39. https://doi.org/10.3390/g17040039
Chicago/Turabian StyleManfredi, Ren, Daniele Vilone, Tijan J. Cvetkovic, Franco Bagnoli, and Andrea Guazzini. 2026. "Inefficient Learning Leads to Efficient Coordination: A Repeated Stag-Hunt Game with Learning Agents" Games 17, no. 4: 39. https://doi.org/10.3390/g17040039
APA StyleManfredi, R., Vilone, D., Cvetkovic, T. J., Bagnoli, F., & Guazzini, A. (2026). Inefficient Learning Leads to Efficient Coordination: A Repeated Stag-Hunt Game with Learning Agents. Games, 17(4), 39. https://doi.org/10.3390/g17040039

