Q-LEARNING ALGORITHM OPTIMIZATION VIA PARTICLE SWARM OPTIMIZATION APPLIED TO THE K-SERVO PROBLEM

Vol 57, 2025 - 341020
Extended Abstracts (EA)
Favorite this paper
How to cite this paper?
Abstract

The k-server problem is an online combinatorial optimization problem in which each movement decision is irrevocable and made without information about future demands. Q-Learning is a promising alternative for this scenario, but its performance is highly sensitive to hyperparameter calibration. This paper proposes a hybrid approach that uses Particle Swarm Optimization (PSO) to tune Q-Learning hyperparameters, combined with Double Q-Learning and Experience Replay. The approach was evaluated on instances parameterized by the Gini Index, covering Multifocal and Migratory dynamics. The experiments show that PSO-based tuning reduces movement cost and decreases the coefficient of variation by up to 45\% compared to fixed configurations, indicating greater learning stability.

Share your ideas or questions with the authors!

Did you know that the greatest stimulus in scientific and cultural development is curiosity? Leave your questions or suggestions to the author!

Sign in to interact

Have a question or suggestion? Share your feedback with the authors!

Institutions
  • 1 Universidade Federal Rural do Semi-Árido
Track
  • EST&AM – OR Analytics in Statistics and Machine Learning
Keywords
Problem of k-servos
Q-Learning
Particle Swarm Optimization
Reinforcement Learning
Metaheuristics