Limiting dynamics for Q-learning with memory one in two-player, two-action games

Autor:	Janusz M Meylahn
Rok vydání:	2021
Předmět:	Equilibrium point Computer Science::Computer Science and Game Theory Strategy Iterated function Computer science Q-learning Repeated game Applied mathematics Graph (abstract data type) Limit (mathematics) Action (physics)
Zdroj:	SSRN Electronic Journal.
ISSN:	1556-5068
DOI:	10.2139/ssrn.3895751
Popis:	We develop a computational method to identify all pure strategy equilibrium points in the strategy space of the two-player, two-action repeated games played by Q-learners with one period memory. In order to approximate the dynamics of these Q-learners, we construct a graph of pure strategy mutual best-responses. We apply this method to the iterated prisoner’s dilemma and find that there are exactly three absorbing states. By analyzing the graph for various values of the discount factor, we find that, in addition to the absorbing states, limit cycles become possible. We confirm our results using numerical simulations.
Databáze:	OpenAIRE
Externí odkaz:	https://explore.openaire.eu/search/publication?articleId=doi_________::114a0e1380992eded8da335fb6599e3c https://doi.org/10.2139/ssrn.3895751 Zobrazit plný text záznamu