PaperScope
LIVE · 2026-09-17 05:40 UTC

Symmetric solution of the Bellman optimality equation for repeated harmony game

Hisato Komatsu

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2609.16289 v1
Category
Submitted
2026-09-14

Abstract

In social dilemma games, additional rewards or punishments have been studied as means of promoting cooperation. Therefore, it is important to investigate the ideal situation, in which such an additional payoff would change the game. In this study, we investigated the symmetric solution of the Bellman optimality equation for a repeated harmony game. The calculations showed that three types of symmetric solutions exist. One of them corresponds to the trivial All-C strategy, and another to the Win-stay Lose-shift strategy of the prisoners dilemma game. The nontrivial behavior of the strategy corresponding to the last solution is also discussed in detail. In addition, we numerically investigated which strategy the agents actually learn by the reinforcement learning algorithm.

Comment: 22 pages, 2 figures

arXiv abs page · PDF · same-day batch