Fetching the paper…
Reading the bibliography…
In the theory of Partially Observed Markov Decision Processes (POMDPs), existence of optimal policies have in general been established via converting the original partially observed stochastic control problem to a fully observed one on the belief space, leading to a belief-MDP.
Central limit theorem for nonstationary Markov chains. i
R.L. Dobrushin · 1956
Earlier work this paper cites.
Probability Measures on Metric Spaces
K.R. Parthasarathy · 1967
Earlier work this paper cites.
Incomplete information in Markovian decision models
D. Rhenius · 1974
Earlier work this paper cites.
Convergence of discretization procedures in dynamic programming
D.P. Bertsekas · 1975
Earlier work this paper cites.
Reduction of a controlled Markov model with incomplete data to a problem with complete information in the case of Borel state and control spaces
A.A. Yushkevich · 1976
Earlier work this paper cites.
Controlled Markov processes with arbitrary numerical criteria
E. A. Feinberg · 1982
Earlier work this paper cites.
Adaptive Markov Control Processes
O. Hernández-Lerma · 1989
Earlier work this paper cites.
An optimal one-way multigrid algorithm for discrete-time stochastic control
C.S. Chow and J. N. Tsitsiklis · 1991
Earlier work this paper cites.
A survey of algorithmic methods for partially observed Markov decision processes
W.S. Lovejoy · 1991
Earlier work this paper cites.
A survey of solution techniques for the partially observed Markov decision process
C.C. White · 1991
Earlier work this paper cites.
Finite-memory suboptimal design for partially observed markov decision processes
C. C. White-III and W. T. Scherer · 1994
Earlier work this paper cites.
Discrete-Time Markov Control Processes: Basic Optimality Criteria
O. Hernandez-Lerma and J. B. Lasserre · 1996
Earlier work this paper cites.
Convergence of probability measures
P. Billingsley · 1999
Earlier work this paper cites.
An improved grid-based approximation algorithm for POMDPs
R. Zhou and E.A. Hansen · 2001
Cited alongside, same era.
Perseus: Randomized point-based value iteration for pomdps
N. Vlassis and M. T. J. Spaan · 2005
Cited alongside, same era.
Anytime point-based approximations for large pomdps
J. Pineau, G. Gordon, and S. Thrun · 2006
Cited alongside, same era.
Point-based value iteration for continuous pomdps
J. M. Porta, N. Vlassis, M. T. J. Spaan, and P. Poupart · 2006
Cited alongside, same era.
Optimal transport: old and new
C. Villani · 2008
Cited alongside, same era.
On near optimality of the set of finite-state controllers for average cost pomdp
H. Yu and D.P. Bertsekas · 2008
Cited alongside, same era.
Covering number for efficient heuristic-based pomdp planning
Z. Zhang, D. Hsu, and W. S. Lee · 2014
Later among the works it cites.
Partially observable total-cost Markov decision process with weakly continuous transition probabilities
E.A. Feinberg, P.O. Kasyanov, and M.Z. Zgurovsky · 2016
Later among the works it cites.
Partially observed Markov decision processes: from filtering to controlled sensing
V. Krishnamurthy · 2016
Later among the works it cites.
On the asymptotic optimality of finite approximations to markov decision processes with borel spaces
N. Saldi, S. Yüksel, and T. Linder · 2017
Later among the works it cites.
Finite Approximations in discrete-time stochastic control
N. Saldi, T. Linder, and S. Yuksel · 2018
Later among the works it cites.
Weak feller property of non-linear filters
A. D. Kara, N. Saldi, and S. Yüksel · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A density projection approach to dimension reduction for continuous-state POMDPs
E. Zhou, M. C. Fu, and S. I. Marcus · 2008
Cited alongside, same era.
Intrinsic methods in filter stability
P. Chigansky, R. Liptser, and R. van Handel · 2009
Cited alongside, same era.
Solving continuous-state POMDPs via density projection
E. Zhou, M. C. Fu, and S. I. Marcus · 2010
Cited alongside, same era.
Approximation of Markov decision processes with general state space
F. Dufour and T. Prieto-Rumeau · 2012
Cited alongside, same era.
Average cost Markov decision processes with weakly continuous transition probabilities
E.A. Feinberg, P.O. Kasyanov, and N.V. Zadioanchuk · 2012
Cited alongside, same era.
Point-based pomdp algorithms: Improved analysis and implementation
T. Smith and R. Simmons · 2012
Cited alongside, same era.
Later among the works it cites.
Observability and filter stability for partially observed markov processes
C. McDonald and S. Yüksel · 2019
Later among the works it cites.
Approximate information state for partially observed systems
J. Subramanian and A. Mahajan · 2019
Later among the works it cites.
Information state embedding in partially observable cooperative multi-agent reinforcement learning
W. Mao, K. Zhang, E. Miehling, and T. Başar · 2020
Closest in time.
Exponential filter stability via Dobrushin’s coefficient
C. McDonald and S. Yüksel · 2020
Closest in time.
Finite model approximations for partially observed markov decision processes with discounted cost
N. Saldi, S. Yüksel, and T. Linder · 2020
Closest in time.
A. D. Kara and S. Yüksel · 2021
Closest in time.
C. McDonald and S. Yüksel · 2022
Closest in time.