arXiv Science⌕ Search

arXiv · 2610.03253

Prompt framing governs LLM default following in collective-action

Abstract

Large language models are increasingly deployed as agents that make or recommend decisions on behalf of users, often operating through interfaces that pre-fill suggested values or default options. Whether models treat such defaults as merely informational or as suggestions that systematically alter their choices remains unclear. We study default deference in two one-shot social dilemmas: a common-pool resource (CPR) extraction game and a threshold public-good (TPG) contribution game. We measure how defaults shift each model's choice distribution relative to its no-default baseline, across default values, wordings, and action-space granularities. We find that pre-filled defaults pull probability mass on the default value in both games, but the magnitude depends strongly on wording: the same model can show high pull under one formulation and near-zero pull under another. Permission-style wording reduces default pull in both games, more strongly in CPR than in TPG. Default pull is weaker in coarse action spaces, and conflict defaults attract more mass than agreement ones. These results indicate that default deference depends on the model, the wording of the interface, and the structure of available choices. For agentic systems, evaluating model behaviour without controlling the surrounding choice architecture can miss an important source of behavioural variation.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Eladio Montero-Porras, Axel Abels, Tom Lenaerts. 2026-10-02. Prompt framing governs LLM default following in collective-action. https://arxiv.org/abs/2610.03253

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Path Polymatrix Games are PPAD-hard

We prove that it is PPAD-hard to compute a Nash equilibrium in a tree polymatrix game with at most twenty actions per player, and path polymatrix game with at most four hundred actions per player. These are the first PPAD hardness results for a game with a constant number of actions per player where the interaction graph is acyclic. Along the way we show PPAD-hardness for finding an $ε$-fixed point of a 2D LinearFIXP instance, when $ε$ is any constant less than $0.5$. This result is tight, since finding a 0.5-fixed point in a 2D LinearFIXP instance is trivial.

cs.GT↗

LLM-Guided Reinforcement Learning with Representative Agents for Traffic Modeling

Large language models (LLMs) are increasingly used as behavioral proxies for self-interested travelers in agent-based traffic models. Although more flexible and generalizable than conventional models, the practical use of these approaches remains limited by scalability due to the cost of calling one LLM for every traveler. Moreover, it has been found that LLM agents often make opaque choices and produce unstable day-to-day dynamics. To address these challenges, we propose to model each homogeneous traveler group facing the same decision context with a single representative LLM agent who behaves like the population's average, maintaining and updating a mixed strategy over routes that coincides with the group's aggregate flow proportions. Each day, the LLM reviews the travel experience and flags routes with positive reinforcement that they hope to use more often, and an interpretable update rule then converts this judgment into strategy adjustments using a tunable (progressively decaying) step size. The representative-agent design improves scalability, while the separation of reasoning from updating clarifies the decision logic while stabilizing learning. In classic traffic assignment settings, we find that the proposed approach converges rapidly to the user equilibrium. In richer settings with income heterogeneity, multi-criteria costs, and multi-modal choices, the generated dynamics remain stable and interpretable, reproducing plausible behavioral patterns well-documented in psychology and economics, for example, the decoy effect in toll versus non-toll road selection, and higher willingness-to-pay for convenience among higher-income travelers when choosing between driving, transit, and park-and-ride options.

cs.GT↗

Pricing Time, Not Just Tokens: Latency-Aware Mechanism Design for LLM Inference

The economic theory of LLM pricing treats tokens as a homogeneous commodity considering aggregate token count as the main features buyers and sellers consider. We model inference as a service market where buyers have three-dimensional private information - willingness-to-pay, task volume, and time preference - and utility depends on latency slack alongside token quantities. Our main result is a separation theorem: discrete hardware tiers induce endogenous self-selection on time preferences, reducing three-dimensional screening to standard one-dimensional screening within each tier. We derive the cost structure from GPU inference physics - compute-bound prefill and bandwidth-bound decode - and characterize optimal tiered mechanisms via virtual-value techniques. Optimal per-task prices are volume-independent, providing theoretical grounding for flat per-token API pricing. We verify the mechanism empirically by calibrating to 8-GPU clusters of H100 and B200 hardware. The separation theorem holds in 83% of 105 tested configurations overall, rising to 96% at economically relevant WTP scales. A seller adopting two-tier pricing under the optimal mechanism captures 26-66% higher profit than the best single-tier alternative, with gains driven by efficient cross-tier allocation in regimes where hardware costs are a significant fraction of per-request value.

cs.GT↗