Limiting dynamics for Q-learning with memory one in two-player, two-action games

Meylahn, Janusz M

Full-text links:

Download:

Current browse context:

math.DS

< prev | next >

new | recent | 2107

Mathematics > Dynamical Systems

Title: Limiting dynamics for Q-learning with memory one in two-player, two-action games

Authors: Janusz M Meylahn

(Submitted on 29 Jul 2021 (this version), latest version 3 Oct 2022 (v2))

Abstract: We develop a computational method to identify all pure strategy equilibrium points in the strategy space of the two-player, two-action repeated games played by Q-learners with one period memory. In order to approximate the dynamics of these Q-learners, we construct a graph of pure strategy mutual best-responses. We apply this method to the iterated prisoner's dilemma and find that there are exactly three absorbing states. By analyzing the graph for various values of the discount factor, we find that, in addition to the absorbing states, limit cycles become possible. We confirm our results using numerical simulations.

Subjects:	Dynamical Systems (math.DS); Adaptation and Self-Organizing Systems (nlin.AO)
Cite as:	arXiv:2107.13995 [math.DS]
	(or arXiv:2107.13995v1 [math.DS] for this version)

Submission history

From: Janusz Meylahn [view email]
[v1] Thu, 29 Jul 2021 14:13:48 GMT (334kb,D)
[v2] Mon, 3 Oct 2022 13:38:02 GMT (22192kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> math > arXiv:2107.13995v1

Download:

Current browse context:

Change to browse by:

References & Citations

Bookmark

Mathematics > Dynamical Systems

Title: Limiting dynamics for Q-learning with memory one in two-player, two-action games

Submission history