Fetching the paper…
Reading the bibliography…
In real-world control settings, the observation space is often unnecessarily high-dimensional and subject to time-correlated noise.
On a problem of partitions
Alfred Brauer · 1942
Earlier work this paper cites.
A theorem on regular matrices
Peter Perkins · 1961
Earlier work this paper cites.
Variants of the selberg sieve, and bounded intervals containing many primes
DHJ Polymath · 2014
Earlier work this paper cites.
Curiosity-driven exploration by self-supervised prediction
Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell · 2017
Earlier work this paper cites.
Provably efficient rl with rich observations via latent state decoding
Simon Du, Akshay Krishnamurthy, Nan Jiang, Alekh Agarwal, Miroslav Dudik, and John Langford · 2019
Earlier work this paper cites.
Deepmdp: Learning continuous latent space models for representation learning
Carles Gelada, Saurabh Kumar, Jacob Buckman, Ofir Nachum, and Marc G. Bellemare · 2019
Earlier work this paper cites.
Learning latent dynamics for planning from pixels
Danijar Hafner, Timothy Lillicrap, Ian Fischer, Ruben Villegas, David Ha, Honglak Lee, and James Davidson · 2019
Earlier work this paper cites.
Deep reinforcement and infomax learning
Bogdan Mazoure, Remi Tachet des Combes, Thang Long Doan, Philip Bachman, and R Devon Hjelm · 2020
Cited alongside, same era.
Learning the linear quadratic regulator from nonlinear observations
Zakaria Mhammedi, Dylan J Foster, Max Simchowitz, Dipendra Misra, Wen Sun, Akshay Krishnamurthy, Alexander Rakhlin, and John Langford · 2020
Cited alongside, same era.
Kinematic state abstraction and provably efficient rich-observation reinforcement learning
Dipendra Misra, Mikael Henaff, Akshay Krishnamurthy, and John Langford · 2020
Cited alongside, same era.
Learning invariant representations for reinforcement learning without reconstruction
Amy Zhang, Rowan Thomas McAllister, Roberto Calandra, Yarin Gal, and Sergey Levine · 2020
Cited alongside, same era.
Uniqueness and complexity of inverse mdp models, 2022
Marcus Hutter and Steven Hansen · 2022
Cited alongside, same era.
Denoised mdps: Learning world models better than the world itself
Tongzhou Wang, Simon Du, Antonio Torralba, Phillip Isola, Amy Zhang, and Yuandong Tian · 2022
Later among the works it cites.
Principled offline RL in the presence of rich exogenous information
Riashat Islam, Manan Tomar, Alex Lamb, Yonathan Efroni, Hongyu Zang, Aniket Rajiv Didolkar, Dipendra Misra, Xin Li, Harm Van Seijen, Remi Tachet Des Combes, and John Langford · 2023
Later among the works it cites.
Interpretable (un)controllable features in MDP’s
Jacob Eeuwe Kooi, Mark Hoogendoorn, and Vincent Francois-Lavet · 2023
Later among the works it cites.
Pclast: Discovering plannable continuous latent states, 2023
Anurag Koul, Shivakanth Sujit, Shaoru Chen, Ben Evans, Lili Wu, Byron Xu, Rajan Chari, Riashat Islam, Raihan Seraj, Yonathan Efroni, Lekan Molu, Miro Dudik, John Langford, and Alex Lamb · 2023
Later among the works it cites.
Zakaria Mhammedi, Dylan J Foster, and Alexander Rakhlin · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alex Lamb, Riashat Islam, Yonathan Efroni, Aniket Rajiv Didolkar, Dipendra Misra, Dylan J Foster, Lekan P Molu, Rajan Chari, Akshay Krishnamurthy, and John Langford · 2022
Cited alongside, same era.
Sample-efficient reinforcement learning in the presence of exogenous information
Yonathan Efroni, Dylan J Foster, Dipendra Misra, Akshay Krishnamurthy, and John Langford
Cited in the paper.
Provably filtering exogenous distractors using multistep inverse dynamics
Yonathan Efroni, Dipendra Misra, Akshay Krishnamurthy, Alekh Agarwal, and John Langford
Cited in the paper.
Later among the works it cites.
Agent-centric state discovery for finite-memory POMDPs
Lili Wu, Ben Evans, Riashat Islam, Raihan Seraj, Yonathan Efroni, and Alex Lamb · 2023
Later among the works it cites.