Online Learning for Dual-Index Policies in Dual-Sourcing Systems

Jingwen Tang,Boxiao Chen,Cong Shi

doi:10.1287/msom.2022.0323

Abstract

Problem definition: We consider a periodic-review dual-sourcing inventory system with a regular source (lower unit cost but longer lead time) and an expedited source (shorter lead time but higher unit cost) under carried-over supply and backlogged demand. Unlike existing literature, we assume that the firm does not have access to the demand distribution a priori and relies solely on past demand realizations. Even with complete information on the demand distribution, it is well known in the literature that the optimal inventory replenishment policy is complex and state dependent. Therefore, we focus our attention on a class of popular, easy-to-implement, and near-optimal heuristic policies called the dual-index policy. Methodology/results: The performance measure is the regret, defined as the cost difference of any feasible learning algorithm against the full-information optimal dual-index policy. We develop a nonparametric online learning algorithm that admits a regret upper bound of [Formula: see text], which matches the regret lower bound for any feasible learning algorithms up to a logarithmic factor. Our algorithm integrates stochastic bandits and sample average approximation techniques in an innovative way. As part of our regret analysis, we explicitly prove that the underlying Markov chain is ergodic and converges to its steady state exponentially fast via coupling arguments, which could be of independent interest. Managerial implications: Our work provides practitioners with an easy-to-implement, robust, and provably good online decision support system for managing a dual-sourcing inventory system. Funding: This work was supported by the Amazon Research Award. Supplemental Material: The online appendix is available at https://doi.org/10.1287/msom.2022.0323 .

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Online Learning for Dual-Index Policies in Dual-Sourcing Systems

Abstract

Talk to us

Similar Papers

More From: Manufacturing & Service Operations Management

Lead the way for us

Journal: Manufacturing & Service Operations Management	Publication Date: Dec 12, 2023
Citations: 1

Similar Papers

Order quantities for style goods with two order opportunities and Bayesian updating of demand. Part II: capacity constraints
J Miltenburg ... H C Pong
International Journal of Production Research | VOL. 45
J Miltenburg, et. al.J Miltenburg ... H C Pong
15 Apr 2007
International Journal of Production Research | VOL. 45

Order quantities for style goods with two order opportunities and Bayesian updating of demand. Part I: no capacity constraints
J Miltenburg ... H C Pong
International Journal of Production Research | VOL. 45
J Miltenburg, et. al.J Miltenburg ... H C Pong
01 Apr 2007
International Journal of Production Research | VOL. 45

Practising in the Margins: High Turnover, Low Unit Cost
Leon Van Schaik
Architectural Design | VOL. 88
Leon Van SchaikLeon Van Schaik
01 Sep 2018
Architectural Design | VOL. 88

Comparing NHS Hospital Unit Costs
Diane Dawson ... Andrew Street
Public Money & Management | VOL. 20
Diane Dawson, et. al.Diane Dawson ... Andrew Street
01 Oct 2000
Public Money & Management | VOL. 20

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Online Learning for Dual-Index Policies in Dual-Sourcing Systems

Abstract

Talk to us

Similar Papers

More From: Manufacturing &amp; Service Operations Management

More From: Manufacturing & Service Operations Management