arXiv · 2609.14827
Multi-Agent Reinforcement Learning in Markets with Congestion
Abstract
This paper investigates multi-agent reinforcement learning (MARL) in settings where firms compete for customers using congestible resources. We consider Bertrand competition in which firms compete by announcing prices and customers choose among firms based on both price and congestion. The relationship between price, congestion and the quantity of customers willing to accept service is governed by an unknown inverse demand curve, which firms must learn through experience. Each firm is modeled as a self-interested learning agent that chooses its price to maximize profit. A growing literature has shown that independently learning MARL agents can develop tacitly collusive behavior. We examine how such behavior emerges in markets with congestible resources. Our results provide insight into how learning dynamics, state representation, and strategic interaction jointly shape competition, with implications for both economic learning and the design of learning-enabled markets.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Qixuan Zai, Randall Berry. 2026-09-13. Multi-Agent Reinforcement Learning in Markets with Congestion. https://arxiv.org/abs/2609.14827
Cite the original work for its findings. Save a collection to share your selection of sources.