Reference health

The AI Economist: Taxation policy design via two-level deep multiagent reinforcement learning

https://doi.org/10.1126/sciadv.abk2607
CiteStamped reference-health badge
29/29 checkable references clean · checked 2026-08-07

Every reference with a DOI in the deposited reference list resolved to a known work in Crossref or DataCite at the dated check, and none carried a retraction, withdrawal, or removal notice.

16 without a DOI — not checked. A reference deposited without a DOI is never matched by title or guessed at; it stays outside the checked set, and this line discloses that.

The 29 checked references that resolve
resolves10.1093/epirev/mxh003
Income Inequality and Health: What Have We Learned So Far?
resolves10.1287/moor.6.1.58
Optimal Auction Design
resolves10.1177/1043463192004001008
On the Value of Game Theory in Social Science
resolves10.2307/j.ctvcm4j8j.18
Behavioral Game Theory:
resolves10.1016/S0167-2231(76)80003-6
Econometric policy evaluation: A critique
resolves10.1007/978-3-540-45193-8_2
Automated Mechanism Design: A New Application Area for Search Algorithms
resolves10.1038/nature24270
Mastering the game of Go without human knowledge
resolves10.1038/s41586-019-1724-z
Grandmaster level in StarCraft II using multi-agent reinforcement learning
resolves10.1016/j.eswa.2018.01.039
Using deep Q-learning to understand the tax evasion behavior of risk-averse firms
resolves10.3390/g10040042
Taxation with Mobile High-Income Agents: Experimental Evidence on Tax Compliance and Equity Perceptions
resolves10.1073/pnas.082080899
Agent-based modeling: Methods and techniques for simulating human systems
resolves10.1257/aer.102.3.53
Getting at Systemic Risk via an Agent-Based Model of the Housing Market
resolves10.1145/1553374.1553380
Curriculum learning
resolves10.1080/09540099108946587
Function Optimization using Connectionist Reinforcement Learning Algorithms
resolves10.1007/s10458-009-9108-7
Evolutionary mechanism design: a review
resolves10.1016/0047-2727(76)90047-5
Optimal tax theory
resolves10.1257/jep.23.4.147
Optimal Taxation in Theory and Practice
resolves10.1016/S1573-4420(02)80025-8
Taxation and Economic Efficiency**We thank Charles Blackorby, Peter Diamond, Kenneth Judd, Louis Kaplow, Gareth Myles, Michel Strawczynski and Ronald Wendner for helpful comments on a previous draft.
resolves10.1111/1467-937X.00166
Using Elasticities to Derive Optimal Income Tax Rates
resolves10.1257/aer.20141362
Generalized Social Marginal Welfare Weights for Optimal Tax Theory
resolves10.1515/9781400835270
The New Dynamic Public Finance
resolves10.1111/j.1467-937X.2006.00367.x
Dynamic Optimal Taxation with Private Information
resolves10.3386/w17642
Optimal Dynamic Taxes
resolves10.1111/j.1467-937X.2009.00587.x
Dynamic Mirrlees Taxation under Political Economy Constraints
resolves10.1177/1091142110381640
Tax Compliance as an Evolutionary Coordination Game: An Agent-Based Approach
resolves10.1016/j.socec.2012.11.002
An agent based model for studying optimal tax collection policy using experimental data: The cases of Chile and Italy
resolves10.1016/S0047-2727(01)00085-8
The elasticity of taxable income: evidence and implications
resolves10.1145/3287560.3287596
Model Cards for Model Reporting
The 16 references without a DOI — listed, not checked
no DOI — not checkedUnited Nations Inequality Matters: Report of the World Social Situation 2013 (Department of Economic and Social Affairs 2013).
no DOI — not checkedA. M. Rivlin P. M. Timpane Ethical and Legal Issues of Social Experimentation (Brookings Institution Washington 1975) vol. 4.
no DOI — not checkedV. Conitzer T. Sandholm Complexity of mechanism design in Proceedings of the 18th Conference on Uncertainty in Artificial Intelligence (University of Alberta Edmonton Alberta Canada August 1-4 2002) pp. 103–110.
no DOI — not checkedH. Narasimhan S. Agarwal D. C. Parkes Automated mechanism design without money via machine learning in Proceedings of the 25th International Joint Conference on Artificial Intelligence (AAAI Press/International Joint Conferences on Artificial Intelligence 2016) pp. 433–439.
no DOI — not checkedT. Baumann T. Graepel J. Shawe-Taylor Adaptive mechanism design: Learning to promote cooperation. arXiv:1806.04067 [cs.GT] (11 June 2018).
no DOI — not checkedP. Duetting Z. Feng H. Narasimhan D. Parkes S. S. Ravindranath Optimal auctions through deep learning in Proceedings of the 36th International Conference on Machine Learning K. Chaudhuri R. Salakhutdinov Eds. (PMLR 2019) pp. 1706–1715.
no DOI — not checkedOpenAI Openai five (2018); https://blog.openai.com/openai-five/.
no DOI — not checkedJ. Z. Leibo V. Zambaldi M. Lanctot J. Marecki T. Graepel Multi-agent reinforcement learning in sequential social dilemmas in Proceedings of the 16th Conference on Autonomous Agents and MultiAgent Systems (AAMAS ‘17) (International Foundation for Autonomous Agents and Multiagent Systems 2017) pp. 464–473.
no DOI — not checkedM. P. Wellman Methods for empirical game-theoretic analysis. AAAI 1552–1556 (2006).
no DOI — not checkedK. Tuyls J. Perolat M. Lanctot J. Z. Leibo T. Graepel A generalised method for empirical game theoretic analysis in Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems (AAMAS ‘18) (International Foundation for Autonomous Agents and Multiagent Systems 2018) pp. 77–85.
no DOI — not checkedP. Muller S. Omidshafiei M. Rowland K. Tuyls J. Perolat S. Liu D. Hennes L. Marris M. Lanctot E. Hughes Z. Wang G. Lever N. Heess T. Graepel R. Munos A generalized training approach for multiagent learning in International Conference on Learning Representations (OpenReview.net 2020) pp. 1–35.
no DOI — not checkedP. Diamond, J. Mirrlees, Optimal taxation and public production I: Production efficiency. Am. Econ. Rev. 61, 8–27 (1971).
no DOI — not checkedR. S. Sutton A. G. Barto Reinforcement Learning: An Introduction (MIT Press 2018).
no DOI — not checkedK. J. Arrow The theory of risk aversion in Essays in the Theory of Risk-Bearing (Markham Publishing Co. Chicago 1971) pp. 90–120.
no DOI — not checkedT. Gebru J. Morgenstern B. Vecchione J. W. Vaughan H. Wallach H. Daumé III K. Crawford Datasheets for datasets. arXiv:1803.09010 [cs.DB] (23 March 2018).
no DOI — not checkedJ. Schulman F. Wolski P. Dhariwal A. Radford O. Klimov Proximal policy optimization algorithms. arXiv:1707.06347 [cs.LG] (20 July 2017).
What this badge says. CiteStamped means the CHECKABLE references of this work were clean at the dated check: each resolved to a known work in a public registry, and none carried a retraction notice at that time. It says nothing about the quality, findings, or importance of the work itself, and nothing about references deposited without a DOI.

checked 2026-08-07 — re-checked daily as this page is visited; titles and statuses come from Crossref and DataCite and are not part of the signed record

Embed this badge

Both snippets point at the live badge image and link back to this page. The badge re-renders from the daily check, so an embed never goes stale by more than a day of visits.

<a href="https://citestamp.com/citestamped/10.1126/sciadv.abk2607"><img src="https://citestamp.com/citestamped/10.1126/sciadv.abk2607/badge.svg" alt="CiteStamped reference-health badge" width="460" height="64"></a>
[![CiteStamped reference-health badge](https://citestamp.com/citestamped/10.1126/sciadv.abk2607/badge.svg)](https://citestamp.com/citestamped/10.1126/sciadv.abk2607)