arXiv ScienceSearch

arXiv · 2507.00088

How large language models judge and influence human cooperation

Abstract

Humans increasingly rely on large language models (LLMs) to support decisions in social settings. Previous work suggests that such tools shape people's moral and political judgements. However, the long-term implications of LLM-based social decision-making remain unknown. How will human cooperation be affected when the assessment of social interactions relies on language models? This is a pressing question, as human cooperation is often driven by indirect reciprocity, reputations, and the capacity to judge interactions of others. Here, we assess how state-of-the-art LLMs judge cooperative actions. We provide 21 different LLMs with an extensive set of examples where individuals cooperate -- or refuse cooperating -- in a range of social contexts, and ask how these interactions should be judged. Furthermore, through an evolutionary game-theoretical model, we evaluate cooperation dynamics in populations where the extracted LLM-driven judgements prevail, assessing the long-term impact of LLMs on human prosociality. We observe a remarkable agreement in evaluating cooperation against good opponents. On the other hand, we notice within- and between-model variance when judging cooperation with ill-reputed individuals. We show that the differences revealed between models can significantly impact the prevalence of cooperation. Finally, we test prompts to steer LLM norms, showing that such interventions can shape LLM judgements, particularly through goal-oriented prompts. Our research connects LLM-based advices and long-term social dynamics, and highlights the need to carefully align LLM norms in order to preserve human cooperation.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Alexandre S. Pires, Laurens Samson, Sennay Ghebreab, Fernando P. Santos. 2025-06-30. How large language models judge and influence human cooperation. https://arxiv.org/abs/2507.00088

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Hyperscaling of spatial fluctuations constrains the development of urban populations

Urban populations exhibit fractal organization and systematic scaling regularities, yet the scaling exponents reported across cities vary substantially, challenging existing theory. Using 100~m gridded population maps for 109 urban regions in the Netherlands (2000--2023) and 368 major world cities (1975--2020), we recursively coarse-grain each city and quantify how the mean and variance of inhabitants in square grid cells of side length $\ell$ scale with $\ell$. This yields two exponents, $β$ from $\langle N_\ell\rangle\sim \ell^β$ and $γ$ from $\mathrm{Var}(N_\ell)\sim \ell^γ$, where in the small-$\ell$ limit $β$ equals the planar fractal dimension of populated space. Across cities within a given year, $γ$ depends linearly on $β$. Compiling $>$10,000 exponent estimates over five decades shows that this hyperscaling relation is robust yet non-universal: its slope and intercept vary across continents and drift systematically in time, trending toward the limiting form $γ\simeq 2+β$. A mean-field (independent-cell) argument predicts a quadratic mean--variance mapping and cannot reproduce the observed $β$--$γ$ dependence, implying strong spatial correlations. We derive a correlation-aware variance decomposition in which $γ$ is controlled by a correlation dimension $D_c$; in the correlation-dominated regime $γ=2+D_c$. If large maturing cities, as are the ones selected in our dataset, evolve to effective monofractal ($D_c\simeq β$) cities, the asymptotic prediction becomes $γ\simeq 2+β$, consistent with the observed temporal drift. This interdependence links urban geometry and fluctuations, provides an empirical constraint for mechanistic models of urban growth, and implies scaling predictions for spatial indicators built from local means and variances.

physics.soc-ph

Dynamic probabilistic decision networks

A new type of decision networks is suggested and its operation is analyzed. The network nodes are represented by intelligent agents who can denote either some biological beings, like humans, or neurons of the brain, or the nodes of artificial intelligence. The specifics of the network are in the following: It is probabilistic in the sense that the choice, accomplished by each agent, is characterized by the related probability. It is dynamic, with the probabilities varying in time due to the exchange of information between the agents. It is affective, because the agents choose between alternatives by taking account of utility as well as of biases and emotions. In general, it is heterogeneous, being composed of the groups of agents with different properties, for instance having long-term memory and short-term memory. The network dynamics, caused by the information exchange, results in decision error decrease. The network operation is illustrated by the example starting with the Allais paradox, its resolution, and the decision error diminution in the process of decision dynamics with information exchange. Resorting to machine-learning techniques it is possible to regulate the behavior of the network agents forcing them to choose particular alternatives.

physics.soc-ph

Diverse Minds, Divided Networks? Personality Composition, Polarization, and Collective Intelligence in LLM-Based Social Simulations

Simulated societies of large language model agents are used to study online polarization, and separately to study collective intelligence, but the two are rarely measured in the same system. It is therefore difficult to say whether a society's personality composition shapes both, or whether reducing polarization costs collective competence. We present TraitMix, an experimental design in which the Big Five composition of a simulated social network, both trait levels and trait heterogeneity, is a controlled experimental variable, and in which polarization and collective performance are measured in the same runs. Across 991 simulations of hundred-agent societies, spanning six contested topics and six language models, trait heterogeneity has the largest measured effects, acting in opposite directions on two faces of polarization: varied societies hold more dispersed opinions while being less segregated into camps, so homogeneous societies are not moderate but consensual echo chambers. Trait effects are not additive, as Agreeableness determines the sign of Openness, an interaction that replicates across models although the primary model's estimate is influence-driven. Contrary to the trade-off the study was designed to measure, no polarization measure predicts poorer collective performance, and cross-cutting interaction is the only one of four whose association with collective accuracy survives partialling on the aggregation identity. We report ablations removing two potential measurement circularities, an induction gate applied to every model, and the measures that failed them.

physics.soc-ph