Fetching the paper…
Reading the bibliography…
Besides independent learning, human learning process is highly improved by summarizing what has been learned, communicating it with peers, and subsequently fusing knowledge from different sources to assist the current learning goal.
Multi-agent reinforcement learning: Independent vs. cooperative agents. In ICML
Ming Tan. 1993 · 1993
Earlier work this paper cites.
Collaborative learning enhances critical thinking
Anuradha A Gokhale. 1995 · 1995
Earlier work this paper cites.
Collaborative Learning: Cognitive and Computational Approaches. Advances in Learning and Instruction Series
Pierre Dillenbourg. 1999 · 1999
Earlier work this paper cites.
Coordinated reinforcement learning. In ICML
Carlos Guestrin, Michail Lagoudakis, and Ronald Parr. 2002 · 2002
Earlier work this paper cites.
Cross channel optimized marketing by reinforcement learning. In SIGKDD
Naoki Abe, Naval Verma, Chid Apte, and Robert Schroko. 2004 · 2004
Earlier work this paper cites.
Regularized multi–task learning. In SIGKDD
Theodoros Evgeniou and Massimiliano Pontil. 2004 · 2004
Earlier work this paper cites.
A behavior-based scheme using reinforcement learning for autonomous underwater vehicles
Marc Carreras, Junku Yuh, Joan Batlle, and Pere Ridao. 2005 · 2005
Earlier work this paper cites.
Agent based decision support system using reinforcement learning under emergency circumstances. In International Conference on Natural Computation
Devinder Thapa, In-Sung Jung, and Gi-Nam Wang. 2005 · 2005
Earlier work this paper cites.
Model compression. In SIGKDD
Cristian Bucilu, Rich Caruana, and Alexandru Niculescu-Mizil. 2006 · 2006
Earlier work this paper cites.
Collaborative multiagent reinforcement learning by payoff propagation
Jelle R Kok and Nikos Vlassis. 2006 · 2006
Earlier work this paper cites.
Multi-task feature learning
A Evgeniou and Massimiliano Pontil. 2007 · 2007
Earlier work this paper cites.
Transfer learning for reinforcement learning domains: A survey
Matthew E Taylor and Peter Stone. 2009 · 2009
Earlier work this paper cites.
Optimizing debt collections using constrained reinforcement learning. In SIGKDD
Naoki Abe, Prem Melville, Cezar Pendus, Chandan K Reddy, David L Jensen, Vince P Thomas, James J Bennett, Gary F Anderson, Brent R Cooley, Melissa Kowalczyk, and others. 2010 · 2010
Earlier work this paper cites.
Multi-agent reinforcement learning: An overview
Lucian Buşoniu, Robert Babuška, and Bart De Schutter. 2010 · 2010
Cited alongside, same era.
A survey on transfer learning
Sinno Jialin Pan and Qiang Yang. 2010 · 2010
Cited alongside, same era.
A two-stage weighting framework for multi-source domain adaptation. In NIPS
Qian Sun, Rita Chattopadhyay, Sethuraman Panchanathan, and Jieping Ye. 2011 · 2011
Cited alongside, same era.
MALSAR: Multi-task learning via structural regularization
Jiayu Zhou, Jianhui Chen, and Jieping Ye. 2011 · 2011
Cited alongside, same era.
A convex formulation for learning task relationships in multi-task learning
Yu Zhang and Dit-Yan Yeung. 2012 · 2012
Cited alongside, same era.
The Arcade Learning Environment: An evaluation platform for general agents
Janarthanan Rajendran, Aravind Lakshminarayanan, Mitesh M Khapra, Balaraman Ravindran, and others. 2015 · 2015
Later among the works it cites.
Andrei A Rusu, Sergio Gomez Colmenarejo, Caglar Gulcehre, Guillaume Desjardins, James Kirkpatrick, Razvan Pascanu, Volodymyr Mnih, Koray Kavukcuoglu, and Raia Hadsell. 2015 · 2015
Later among the works it cites.
High-dimensional continuous control using generalized advantage estimation
John Schulman, Philipp Moritz, Sergey Levine, Michael Jordan, and Pieter Abbeel. 2015 · 2015
Later among the works it cites.
Simultaneous deep transfer across domains and tasks. In ICCV
Eric Tzeng, Judy Hoffman, Trevor Darrell, and Kate Saenko. 2015 · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling. 2013 · 2013
Cited alongside, same era.
Facial landmark detection by deep multi-task learning. In ECCV
Zhanpeng Zhang, Ping Luo, Chen Change Loy, and Xiaoou Tang. 2014 · 2014
Cited alongside, same era.
Sapiens: a brief history of humankind
Yuval N. Harari. 2015 · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2015 · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, and others. 2015 · 2015
Cited alongside, same era.
Massively parallel methods for deep reinforcement learning
Arun Nair, Praveen Srinivasan, Sam Blackwell, Cagdas Alcicek, Rory Fearon, Alessandro De Maria, Vedavyas Panneershelvam, Mustafa Suleyman, Charles Beattie, Stig Petersen, and others. 2015 · 2015
Cited alongside, same era.
Actor-mimic: Deep multitask and transfer reinforcement learning
Emilio Parisotto, Jimmy Lei Ba, and Ruslan Salakhutdinov. 2015 · 2015
Cited alongside, same era.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Later among the works it cites.
Reinforcement learning with unsupervised auxiliary tasks
Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki, Tom Schaul, Joel Z Leibo, David Silver, and Koray Kavukcuoglu. 2016 · 2016
Later among the works it cites.
Playing FPS games with deep reinforcement learning
Guillaume Lample and Devendra Singh Chaplot. 2016 · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy P Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 2016
Later among the works it cites.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, and others. 2016 · 2016
Later among the works it cites.
Target-driven visual navigation in indoor scenes using deep reinforcement learning
Yuke Zhu, Roozbeh Mottaghi, Eric Kolve, Joseph J Lim, Abhinav Gupta, Li Fei-Fei, and Ali Farhadi. 2016 · 2016
Later among the works it cites.
Learning Invariant Feature Spaces to Transfer Skills with Reinforcement Learning. In Under review as a conference paper at ICLR 2017
YuXuan Liu Pieter Abbeel†‡ Sergey Levine Abhishek Gupta†, Coline Devin†. 2017 · 2017
Closest in time.
OpenAI universe-starter-agent
OpenAI. 2017 · 2017
Closest in time.