cs.GTSep 14, 2026

Symmetric solution of the Bellman optimality equation for repeated harmony game

Authors: Hisato Komatsu

Organizations: Department of Physics, Kindai University, 577-8502, Higashi-Osaka, Osaka, Japan · Data Science and AI Innovation Research Promotion Center, Shiga University, 522-8522, Hikone, Shiga, Japan

Abstract

In social dilemma games, additional rewards or punishments have been studied as means of promoting cooperation. Therefore, it is important to investigate the ideal situation, in which such an additional payoff would change the game. In this study, we investigated the symmetric solution of the Bellman optimality equation for a repeated harmony game. The calculations showed that three types of symmetric solutions exist. One of them corresponds to the trivial All-C strategy, and another to the Win-stay Lose-shift strategy of the prisoners dilemma game. The nontrivial behavior of the strategy corresponding to the last solution is also discussed in detail. In addition, we numerically investigated which strategy the agents actually learn by the reinforcement learning algorithm.

Explore similar work

CardsList