Fetching the paper…
Reading the bibliography…
In many operations management problems, we need to make decisions sequentially to minimize the cost while satisfying certain constraints.
1901
Earlier work this paper cites.
1901
Earlier work this paper cites.
1903
Earlier work this paper cites.
1906
Earlier work this paper cites.
1909
Earlier work this paper cites.
1909
Earlier work this paper cites.
1910
Earlier work this paper cites.
Sion M, et al. (1958) On general minimax theorems. Pacific Journal of mathematics 8(1):171–176
1958
Earlier work this paper cites.
Bertsimas D, Orlin JB (1994) A technique for speeding up the solution of the Lagrangian dual. Mathematical Programming 63(1-3):23–45
1994
Earlier work this paper cites.
Singh SP, Cohn D (1998) How to dynamically merge Markov decision processes. Advances in neural information processing systems , 1057–1063
1998
Earlier work this paper cites.
Altman E (1999) Constrained Markov decision processes , volume 7 (CRC Press)
1999
Earlier work this paper cites.
Kakade S, Langford J (2002) Approximately optimal approximate reinforcement learning. ICML , volume 2, 267–274
2002
Earlier work this paper cites.
Harrison JM, Zeevi A (2004) Dynamic scheduling of a multiclass queue in the halfin-whitt heavy traffic regime. Operations Research 52(2):243–257
2004
Earlier work this paper cites.
Mandelbaum A, Stolyar AL (2004) Scheduling flexible servers with convex delay costs: Heavy-traffic optimality of the generalized c μ \mu -rule. Operations Research 52(6):836–855
2004
Earlier work this paper cites.
Stolyar AL, et al. (2004) Maxweight scheduling in a generalized switch: State space collapse and workload minimization in heavy traffic. The Annals of Applied Probability 14(1):1–53
2004
Cited alongside, same era.
Borkar VS (2005) An actor-critic algorithm for constrained Markov decision processes. Systems & control letters 54(3):207–213
2005
Cited alongside, same era.
Adelman D, Mersereau AJ (2008) Relaxations of weakly coupled stochastic dynamic programs. Operations Research 56(3):712–727
2008
Cited alongside, same era.
Dai JG, Lin W, et al. (2008) Asymptotic optimality of maximum pressure policies in stochastic processing networks. The Annals of Applied Probability 18(6):2239–2299
2008
Cited alongside, same era.
Nedić A, Ozdaglar A (2009) Subgradient methods for saddle-point problems. Journal of optimization theory and applications 142(1):205–228
2009
2016
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Neely MJ (2011) Online fractional programming for Markov decision systems. 2011 49th Annual Allerton Conference on Communication, Control, and Computing (Allerton) , 353–360 (IEEE)
2011
Cited alongside, same era.
Bhatnagar S, Lakshmanan K (2012) An online actor–critic algorithm with function approximation for constrained Markov decision processes. Journal of Optimization Theory and Applications 153(3):688–708
2012
Cited alongside, same era.
Nemirovski A (2012) Tutorial: Mirror descent algorithms for large-scale deterministic and stochastic convex optimization. Conference on Learning Theory (COLT)
2012
Cited alongside, same era.
Turken N, Tan Y, Vakharia AJ, Wang L, Wang R, Yenipazarli A (2012) The multi-product newsvendor problem: Review, extensions, and directions for future research. Handbook of newsvendor problems , 3–39 (Springer)
2012
Cited alongside, same era.
Bubeck S (2014) Convex optimization: Algorithms and complexity. arXiv preprint arXiv:1405.4980
2014
Cited alongside, same era.
Caramanis C, Dimitrov NB, Morton DP (2014) Efficient algorithms for budget-constrained Markov decision processes. IEEE Transactions on Automatic Control 59(10):2813–2817
2014
Cited alongside, same era.
Bertsekas DP, Scientific A (2015) Convex optimization algorithms (Athena Scientific Belmont)
2015
Cited alongside, same era.
2018
Later among the works it cites.
Sutton RS, Barto AG (2018) Reinforcement learning: An introduction (MIT press)
2018
Later among the works it cites.
2018
Later among the works it cites.
Balseiro SR, Brown DB, Chen C (2019) Dynamic pricing of relocating resources in large networks. ACM SIGMETRICS Performance Evaluation Review 47(1):29–30
2019
Later among the works it cites.
Dai J, Shi P (2019) Inpatient overflow: An approximate dynamic programming approach. Manufacturing & Service Operations Management 21(4):894–911
2019
Later among the works it cites.
Miryoosefi S, Brantley K, Daume III H, Dudik M, Schapire RE (2019) Reinforcement learning with convex constraints. Advances in Neural Information Processing Systems , 14093–14102
2019
Later among the works it cites.
Brown DB, Smith JE (2020) Index policies and performance bounds for dynamic selection problems. Management Science
2020
Later among the works it cites.
Chen J, Dong J, Shi P (2020) A survey on skill-based routing with applications to service operations management. Queueing Systems 1–30
2020
Later among the works it cites.
Song H, Tucker AL, Graue R, Moravick S, Yang JJ (2020) Capacity pooling in hospitals: The hidden consequences of off-service placement. Management Science 66(9):3825–3842
2020
Later among the works it cites.