arXiv ScienceSearch

arXiv subjects

Ruqiang Guo

Publications and source records attributed to Ruqiang Guo.

3 recordsLinked to original sources

Payoffs and perception mediate environmental feedback in an N-player trust game with Q-learning

Trust develops through learning, while collective behavior can alter the environment in which later decisions are made. Reinforcement-learning models describe adaptation, and eco-evolutionary models describe behavior-environment feedback, but how an endogenous environment changes trust through material incentives and perceived states remains unclear. We couple a fixed-role, two-population hierarchical trust game populated by heterogeneous tabular $Q$-learning agents to a centered endogenous environment. For the main feedback comparisons, we initialize the fixed-payoff baseline and feedback branches from the same learned state to control for prior learning. We then compare responses when payoff parameters or the observed environmental tier vary. As the payoff multiplication factor increased, successful trust in the fixed-payoff game crossed from a low to a higher finite-horizon level, led by risk-seeking investors. Environmental feedback promoted trust near perception-tier boundaries and reinforced its decline at the high-return operating point. At the high-return operating point, payoff feedback reinforced the learning-driven decline in trust. Stronger drive and longer horizons deepened this negative effect. Compared with a fixed-payoff trust game, environmental feedback carries past collective outcomes into future payoffs and observed states.

math-ph

Strategy Evolution in the Adoption of Conservation Tillage Technology under Time Preference Heterogeneity and Lemon Market: Insights from Evolutionary Dynamics

The promotion of Conservation Tillage Technology (CTT) is critical for mitigating global soil degradation, yet their actual adoption rates remain substantially lower than anticipated targets. Existing research predominantly focuses on static factor analyses, failing to adequately capture the dynamic evolutionary mechanisms of farmer strategic interactions and the impacts of information asymmetries in agricultural markets. This study constructs an evolutionary game model integrating heterogeneous time preferences and lemon market effects to reveal the dynamic equilibrium of technology adoption within farmer groups operating under bounded rationality. Key findings indicate that farmers with high time preferences significantly impede CTT adoption due to the excessive discounting of long-term benefits. Furthermore, the lemon market effect dictates the system's equilibrium states: 1) When the lemon market benefit ($P$) exceeds the lemon market loss ($Q$) ($P > Q$), stable tripartite coexistence of adoption strategies emerges; 2) When $P < Q$, the system evolves unpredictably, exhibiting dynamics characterized by heteroclinic cycles; 3) At the critical threshold $P = Q$, the system transforms into a conservative Hamiltonian system, yielding stable periodic oscillation solutions. Based on these insights, policy recommendations are proposed: implementing ecological certification schemes to mitigate information asymmetry, offering subsidies and insurance to reduce adoption risks, and utilizing environmental taxes to internalize the negative externalities associated with conventional tillage. This research not only provides a dynamic analytical paradigm for the diffusion of green agricultural technologies but also furnishes a theoretical foundation for designing sustainable agricultural policies in developing countries.

physics.soc-ph

The paradigm of tax-reward and tax-punishment strategies in the advancement of public resource management dynamics

In contemporary society, the effective utilization of public resources remains a subject of significant concern. A common issue arises from defectors seeking to obtain an excessive share of these resources for personal gain, potentially leading to resource depletion. To mitigate this tragedy and ensure sustainable development of resources, implementing mechanisms to either reward those who adhere to distribution rules or penalize those who do not, appears advantageous. We introduce two models: a tax-reward model and a tax-punishment model, to address this issue. Our analysis reveals that in the tax-reward model, the evolutionary trajectory of the system is influenced not only by the tax revenue collected but also by the natural growth rate of the resources. Conversely, the tax-punishment model exhibits distinct characteristics when compared to the tax-reward model, notably the potential for bistability. In such scenarios, the selection of initial conditions is critical, as it can determine the system's path. Furthermore, our study identifies instances where the system lacks stable points, exemplified by a limit cycle phenomenon, underscoring the complexity and dynamism inherent in managing public resources using these models.

math.DS