科研速览 · Science Skim继续刷下去 · Keep skimming →
◇ arXiv2026-09-14· cs.GT

Symmetric solution of the Bellman optimality equation for repeated harmony game

Hisato Komatsu

原始摘要(英文原文)· Original abstract
In social dilemma games, additional rewards or punishments have been studied as means of promoting cooperation. Therefore, it is important to investigate the ideal situation, in which such an additional payoff would change the game. In this study, we investigated the symmetric solution of the Bellman optimality equation for a repeated harmony game. The calculations showed that three types of symmetric solutions exist. One of them corresponds to the trivial All-C strategy, and another to the Win-stay Lose-shift strategy of the prisoners dilemma game. The nontrivial behavior of the strategy corresponding to the last solution is also discussed in detail. In addition, we numerically investigated which strategy the agents actually learn by the reinforcement learning algorithm.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Symmetric solution of the Bellman optimality equation for repeated harmony game — 科研速览 Science Skim