Fetching the paper…
Reading the bibliography…
The interaction between an artificial agent and its environment is bi-directional.
Computation of channel capacity and rate-distortion functions
Blahut, Richard · 1972
Earlier work this paper cites.
The bidirectional communication theory–a generalization of information theory
Marko, Hans · 1973
Earlier work this paper cites.
Dynamic programming and optimal control
Bertsekas, Dimitri P., et al · 1995
Earlier work this paper cites.
Directed information for channels with feedback
Kramer, Gerhard · 1998
Earlier work this paper cites.
Reinforcement learning: An introduction
Sutton, Richard S and Barto, Andrew G · 1998
Earlier work this paper cites.
Lyapunov design for safe reinforcement learning
Perkins, Theodore J and Barto, Andrew G · 2002
Earlier work this paper cites.
Control under communication constraints
Tatikonda, Sekhar and Mitter, Sanjoy · 2004
Earlier work this paper cites.
Empowerment: A universal agent-centric measure of control
Klyubin, Alexander S, Polani, Daniel, and Nehaniv, Chrystopher L · 2005
Earlier work this paper cites.
Conservation of mutual and directed information
Massey, James L and Massey, Peter C · 2005
Cited alongside, same era.
Relevant information in optimized persistence vs. progeny strategies
Polani, Daniel, Nehaniv, C, Martinetz, Thomas, and Kim, Jan T · 2006
Cited alongside, same era.
Linearly-solvable markov decision problems
Todorov, Emanuel et al · 2006
Cited alongside, same era.
Introduction to algorithms
Cormen, Thomas H · 2009
Cited alongside, same era.
Finite state channels with time-invariant deterministic feedback
Permuter, Haim Henry, Weissman, Tsachy, and Goldsmith, Andrea J · 2009
Cited alongside, same era.
Information theory of decisions and actions
Tishby, Naftali and Polani, Daniel · 2011
Cited alongside, same era.
Optimal control as a graphical model inference problem
Kappen, Hilbert J, Gómez, Vicenç, and Opper, Manfred · 2012
Later among the works it cites.
Trading value and information in mdps
Rubin, Jonathan, Shamir, Ohad, and Tishby, Naftali · 2012
Later among the works it cites.
Markov decision processes: discrete stochastic dynamic programming
Puterman, Martin L · 2014
Later among the works it cites.
Empowerment–an introduction
Salge, Christoph, Glackin, Cornelius, and Polani, Daniel · 2014
Later among the works it cites.
Variational information maximisation for intrinsically motivated reinforcement learning
Mohamed, Shakir and Rezende, Danilo Jimenez · 2015
Later among the works it cites.
Sdp-based joint sensor and controller design for information-regularized optimal lqg control
Tanaka, Takashi and Sandberg, Henrik · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Elements of information theory
Cover, Thomas M and Thomas, Joy A · 2012
Cited alongside, same era.
All else being equal be empowered
Klyubin, Alexander S, Polani, Daniel, and Nehaniv, Chrystopher L
Cited in the paper.
Tanaka, Takashi, Esfahani, Peyman Mohajerin, and Mitter, Sanjoy K · 2015
Later among the works it cites.