arXiv · 2604.10953
Diffusion Reinforcement Learning Based Online 3D Bin Packing Spatial Strategy Optimization
Abstract
The online 3D bin packing problem is important in logistics, warehousing and intelligent manufacturing, with solutions shifting to deep reinforcement learning (DRL) which faces challenges like low sample efficiency. This paper proposes a diffusion reinforcement learning-based algorithm, using a Markov decision chain for packing modeling, height map-based state representation and a diffusion model-based actor network. Experiments show it significantly improves the average number of packed items compared to state-of-the-art DRL methods, with excellent application potential in complex online scenarios.
Explore related subjects
Keep this discovery
Jie Han, Tong Li, Qingyang Xu, Yong Song, Bao Pang, Xianfeng Yuan. 2026-04-13. Diffusion Reinforcement Learning Based Online 3D Bin Packing Spatial Strategy Optimization. https://arxiv.org/abs/2604.10953
Cite the original work for its findings. Save a collection to share your selection of sources.