Reinforcement Learning for Optimizing Can-Order Policy with the Rolling Horizon Method

Jiseong Noh

doi:10.3390/systems11070350

Abstract

This study presents a novel approach to a mixed-integer linear programming (MILP) model for periodic inventory management that combines reinforcement learning algorithms. The rolling horizon method (RHM) is a multi-period optimization approach that is applied to handle new information in updated markets. The RHM faces a limitation in easily determining a prediction horizon; to overcome this, a dynamic RHM is developed in which RL algorithms optimize the prediction horizon of the RHM. The state vector consisted of the order-up-to-level, real demand, total cost, holding cost, and backorder cost, whereas the action included the prediction horizon and forecasting demand for the next time step. The performance of the proposed model was validated through two experiments conducted in cases with stable and uncertain demand patterns. The results showed the effectiveness of the proposed approach in inventory management, particularly when the proximal policy optimization (PPO) algorithm was used for training compared with other reinforcement learning algorithms. This study signifies important advancements in both the theoretical and practical aspects of multi-item inventory management.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: Systems	Publication Date: Jul 7, 2023
Citations: 2	License type: CC BY 4.0

R Discovery Prime

R Discovery Prime

Reinforcement Learning for Optimizing Can-Order Policy with the Rolling Horizon Method

Abstract

Talk to us

Similar Papers

More From: Systems

Lead the way for us

Similar Papers

Finite Capacity Scheduling of Make-Pack Production: Case Study of Adhesive Factory
T Chotpradit ... P Yenradee
Journal of Engineering, Project, and Production Management | VOL. 4
T Chotpradit, et. al.T Chotpradit ... P Yenradee
31 Jul 2014
Journal of Engineering, Project, and Production Management | VOL. 4

Traffic Navigation for Urban Air Mobility with Reinforcement Learning
Jaeho Lee ... Junyoung Noh
-
Jaeho Lee, et. al.Jaeho Lee ... Junyoung Noh
30 Sep 2022
30 Sep 2022

Process control of mAb production using multi-actor proximal policy optimization
Nikita Gupta ... Hariprasad Kodamana
Digital Chemical Engineering | VOL. 8
Nikita Gupta, et. al.Nikita Gupta ... Hariprasad Kodamana
03 Jun 2023
Digital Chemical Engineering | VOL. 8

Dynamic Economic Optimization of a Continuously Stirred Tank Reactor Using Reinforcement Learning
Derek Machalek ... Titus Quah
-
Derek Machalek, et. al.Derek Machalek ... Titus Quah
01 Jul 2020
01 Jul 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Reinforcement Learning for Optimizing Can-Order Policy with the Rolling Horizon Method

Abstract

Talk to us

Similar Papers

More From: Systems