A Double Deep Q-Learning Approach to the Container Relocation Problem

Vol 57, 2025 - 341182
Complete Articles (CA)
Favorite this paper
How to cite this paper?
Abstract

Efficient management of container stacks in maritime terminals is critical to global supply chains, as relocating "blocking'' containers incurs substantial time and cost penalties. This work addresses the static, distinct, deterministic, restricted Container Relocation Problem (CRP), in which containers must be efficiently retrieved in a predetermined order. Following a novel approach, the CRP is modelled as a Markov Decision Process. To learn effective relocation policies without domain-specific heuristics, a Double Deep Q-Learning (DDQN) agent is implemented that leverages neural networks to evaluate the layout of containers and choose the best action. Bayesian Optimization (BO) is implemented to tune crucial hyperparameters systematically and efficiently. Computational experiments on benchmark instances show promising results, with low computational times. This study is, to the best of the authors' knowledge, the first to apply DDQN with BO to the CRP, opening a promising avenue for machine-learning-driven yard management.

Share your ideas or questions with the authors!

Did you know that the greatest stimulus in scientific and cultural development is curiosity? Leave your questions or suggestions to the author!

Sign in to interact

Have a question or suggestion? Share your feedback with the authors!

Institutions
  • 1 Departamento de Engenharia de Produção - Universidade de São Paulo
Track
  • IA- OR and AI
Keywords
Container Relocation Problem
Double Deep Q-Learning
Bayesian Optimization