arXiv · 2609.34005
Learning to cooperate in a changing world: How caring about the future promotes cooperation across scales
Abstract
In social dilemmas, individuals need to forgo short-term temptations to achieve synergistic collective outcomes through cooperation. Previous work has examined mechanisms through which cooperation can evolve, including direct reciprocity, indirect reciprocity, environmental stochasticity, network reciprocity, and demographic stochasticity. These mechanisms have largely been studied under natural selection or social learning. As the other side of the same coin, it is equally important to study cooperation under self-learning, where individuals adapt based on their own experiences. Here, we focus on the dynamics of multi-agent reinforcement learning and derive analytical conditions under which these mechanisms can stabilize cooperation under learning dynamics across scales. We find that reinforcement learning can steer self-interested individuals toward cooperation when agents value the future over short-term temptation. Our work provides a unified approach to identifying the determinants of learning to cooperate in a changing world, thereby paving the way for the advancement of cooperative artificial intelligence.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Yuxin Geng, Xingru Chen, Xin Wang, Hongwei Zheng, Longzhao Liu, Shaoting Tang, Feng Fu. 2026-09-27. Learning to cooperate in a changing world: How caring about the future promotes cooperation across scales. https://arxiv.org/abs/2609.34005
Cite the original work for its findings. Save a collection to share your selection of sources.