Fetching the paper…
Reading the bibliography…
We report the "Recurrent Deterioration" (RD) phenomenon observed in online recommender systems.
The optimal control of partially observable markov processes over a finite horizon
R. D. Smallwood and E. J. Sondik · 1973
Earlier work this paper cites.
Dynamic Programming: Deterministic and Stochastic Models
D. P. Bertsekas · 1987
Earlier work this paper cites.
Q-learning
C. J. Watkins and P. Dayan · 1992
Earlier work this paper cites.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 1998
Earlier work this paper cites.
Value-function approximations for partially observable markov decision processes
M. Hauskrecht · 2000
Earlier work this paper cites.
Content-boosted collaborative filtering for improved recommendations
P. Melville, R. J. Mooney, and R. Nagarajan · 2002
Earlier work this paper cites.
Toward the next generation of recommender systems: A survey of the state-of-the-art and possible extensions
G. Adomavicius and A. Tuzhilin · 2005
Earlier work this paper cites.
Tree-based batch mode reinforcement learning
D. Ernst, P. Geurts, and L. Wehenkel · 2005
Earlier work this paper cites.
Neural fitted q iteration–first experiences with a data efficient neural reinforcement learning method
M. Riedmiller · 2005
Earlier work this paper cites.
An mdp-based recommender system
G. Shani, D. Heckerman, and R. I. Brafman · 2005
Earlier work this paper cites.
The netflix prize
J. Bennett and S. Lanning · 2007
Cited alongside, same era.
Probabilistic matrix factorization
A. Mnih and R. Salakhutdinov · 2007
Cited alongside, same era.
One-class collaborative filtering
R. Pan, Y. Zhou, B. Cao, N. N. Liu, R. Lukose, M. Scholz, and Q. Yang · 2008
Cited alongside, same era.
Collaborative filtering with temporal dynamics
Y. Koren · 2009
Cited alongside, same era.
Collaborative prediction and ranking with non-random missing data
B. M. Marlin and R. S. Zemel · 2009
Cited alongside, same era.
Double q-learning
H. V. Hasselt · 2010
Cited alongside, same era.
A contextual-bandit approach to personalized news article recommendation
L. Li, W. Chu, J. Langford, and R. E. Schapire · 2010
A hidden markov model for collaborative filtering
N. Sahoo, P. V. Singh, and T. Mukhopadhyay · 2012
Later among the works it cites.
Hierarchical exploration for accelerating contextual bandits
Y. Yue, S. A. Hong, and C. Guestrin · 2012
Later among the works it cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Later among the works it cites.
A fast parallel sgd for matrix factorization in shared memory systems
Y. Zhuang, W.-S. Chin, Y.-C. Juan, and C.-J. Lin · 2013
Later among the works it cites.
Probabilistic matrix factorization with non-random missing data
J. M. Hernandez-lobato, N. Houlsby, and Z. Ghahramani · 2014
Later among the works it cites.
Exploration vs. exploitation in the information filtering problem
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Factorization machines
S. Rendle · 2010
Cited alongside, same era.
A linear ensemble of individual and blended models for music rating prediction
P.-L. Chen, C.-T. Tsai, Y.-N. Chen, K.-C. Chou, C.-L. Li, C.-H. Tsai, K.-W. Wu, Y.-C. Chou, C.-Y. Li, W.-S. Lin, et al · 2012
Cited alongside, same era.
The yahoo! music dataset and kdd-cup’11
G. Dror, N. Koenigstein, Y. Koren, and M. Weimer · 2012
Cited alongside, same era.
X. Zhao and P. I. Frazier · 2014
Later among the works it cites.
The movielens datasets: History and context
F. M. Harper and J. A. Konstan · 2015
Later among the works it cites.
Deep recurrent q-learning for partially observable mdps
M. Hausknecht and P. Stone · 2015
Later among the works it cites.
Continuous control with deep reinforcement learning
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Later among the works it cites.