Fetching the paper…
Reading the bibliography…
Deep learning techniques have been widely applied, achieving state-of-the-art results in various fields of study.
ACM SIGART Bulletin 2(4): 160–163
Sutton RS (1991) Dyna, an integrated architecture for learning, planning, and reacting · 1991
Earlier work this paper cites.
In: Reinforcement Learning . Springer, pp. 5–32
Williams RJ (1992) Simple statistical gradient-following algorithms for connectionist reinforcement learning · 1992
Earlier work this paper cites.
Neural Computation 5(4): 613–624
Dayan P (1993) Improving generalization for temporal difference learning: The successor representation · 1993
Earlier work this paper cites.
MIT press Cambridge
Sutton RS and Barto AG (1998) Reinforcement learning: An introduction , volume 1 · 1998
Earlier work this paper cites.
In: ICML , volume 2. pp. 267–274
Kakade S and Langford J (2002) Approximately optimal approximate reinforcement learning · 2002
Earlier work this paper cites.
In: Intelligent Robots and Systems, 2004.(IROS 2004). Proceedings. 2004 IEEE/RSJ International Conference on , volume 3. IEEE, pp. 2149–2154
Koenig N, A B and Howard A (2004) Design and use paradigms for gazebo, an open-source multi-robot simulator · 2004
Earlier work this paper cites.
In: Proceedings of the 37th conference on Winter simulation . Winter Simulation Conference, pp. 83–95
Fu MC, Glover FW and April J (2005) Simulation optimization: a review, new developments, and applications · 2005
Earlier work this paper cites.
Thrun S, Burgard W and Fox D (2005) Probabilistic robotics
2005
Earlier work this paper cites.
Neural computation 18(12): 2936–2941
Szita I and Lörincz A (2006) Learning tetris using the noisy cross-entropy method · 2006
Earlier work this paper cites.
In: AAAI , volume 8. Chicago, IL, USA, pp. 1433–1438
Ziebart BD, Maas AL, Bagnell JA and Dey AK (2008) Maximum entropy inverse reinforcement learning · 2008
Earlier work this paper cites.
In: Proceedings of the 26th annual international conference on machine learning . ACM, pp. 41–48
Bengio Y, Louradour J, Collobert R and Weston J (2009) Curriculum learning · 2009
Earlier work this paper cites.
In: Robotics and Automation (ICRA), 2011 IEEE International Conference on . IEEE, pp. 3607–3613
Kümmerle R, Grisetti G, Strasdat H, Konolige K and Burgard W (2011) g 2 o: A general framework for graph optimization · 2011
Earlier work this paper cites.
In: Proceedings of the fourteenth international conference on artificial intelligence and statistics . pp. 627–635
Ross S, Gordon G and Bagnell D (2011) A reduction of imitation learning and structured prediction to no-regret online learning · 2011
Earlier work this paper cites.
In: Advances in neural information processing systems . pp. 1097–1105
Krizhevsky A, Sutskever I and Hinton GE (2012) Imagenet classification with deep convolutional neural networks · 2012
Earlier work this paper cites.
In: Intelligent Robots and Systems (IROS), 2012 IEEE/RSJ International Conference on . IEEE, pp. 5026–5033
Todorov E, Erez T and Tassa Y (2012) Mujoco: A physics engine for model-based control · 2012
Earlier work this paper cites.
Foundations and Trends® in Robotics 2(1–2): 1–142
Deisenroth MP, Neumann G, Peters J et al. (2013) A survey on policy search for robotics · 2013
Earlier work this paper cites.
In: ICML (3) . pp. 1–9
Levine S and Koltun V (2013) Guided policy search · 2013
Earlier work this paper cites.
In: Intelligent Robots and Systems (IROS), 2013 IEEE/RSJ International Conference on . IEEE, pp. 1321–1326
Rohmer E, Singh SP and Freese M (2013) V-rep: A versatile and scalable robot simulation framework · 2013
Earlier work this paper cites.
APSIPA Transactions on Signal and Information Processing 3
Deng L (2014) A tutorial survey of architectures, algorithms, and applications for deep learning · 2014
Earlier work this paper cites.
In: Advances in neural information processing systems . pp. 2672–2680
Goodfellow I, Pouget-Abadie J, Mirza M, Xu B, Warde-Farley D, Ozair S, Courville A and Bengio Y (2014) Generative adversarial nets · 2014
Earlier work this paper cites.
In: Proceedings of the 31st International Conference on Machine Learning (ICML-14) . pp. 387–395
Silver D, Lever G, Heess N, Degris T, Wierstra D and Riedmiller M (2014) Deterministic policy gradient algorithms · 2014
Earlier work this paper cites.
arXiv preprint arXiv:1412.3474
Tzeng E, Hoffman J, Zhang N, Saenko K and Darrell T (2014) Deep domain confusion: Maximizing for domain invariance · 2014
Earlier work this paper cites.
Technical report, CARNEGIE-MELLON UNIV PITTSBURGH PA ROBOTICS INST
Bagnell JA (2015) An invitation to imitation · 2015
Earlier work this paper cites.
In: Proceedings of the International Conference on Learning Representations (ICLR)
Lillicrap TP, Hunt JJ, Pritzel A, Heess N, Erez T, Tassa Y, Silver D and Wierstra D (2015) Continuous control with deep reinforcement learning · 2015
Earlier work this paper cites.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . pp. 3431–3440
Long J, Shelhamer E and Darrell T (2015) Fully convolutional networks for semantic segmentation · 2015
Earlier work this paper cites.
Nature 518(7540): 529–533
Mnih V, Kavukcuoglu K, Silver D, Rusu AA, Veness J, Bellemare MG, Graves A, Riedmiller M, Fidjeland AK, Ostrovski G et al. (2015) Human-level control through deep reinforcement learning · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1511.06434
Radford A, Metz L and Chintala S (2015) Unsupervised representation learning with deep convolutional generative adversarial networks · 2015
Earlier work this paper cites.
Neural networks 61: 85–117
Schmidhuber J (2015) Deep learning in neural networks: An overview · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1511.07111
Tzeng E, Devin C, Hoffman J, Finn C, Abbeel P, Levine S, Saenko K and Darrell T (2015) Towards adapting deep visuomotor representations from simulated to real environments · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1507.04888
Wulfmeier M, Ondruska P and Posner I (2015) Maximum entropy deep inverse reinforcement learning · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1612.02179
Baram N, Anschel O and Mannor S (2016) Model-based adversarial imitation learning · 2016
Earlier work this paper cites.
arXiv preprint arXiv:1604.07316
Bojarski M, Del Testa D, Dworakowski D, Firner B, Flepp B, Goyal P, Jackel LD, Monfort M, Muller U, Zhang J et al. (2016) End to end learning for self-driving cars · 2016
Earlier work this paper cites.
arXiv preprint arXiv:1606.00915
Chen LC, Papandreou G, Kokkinos I, Murphy K and Yuille AL (2016) Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs · 2016
Earlier work this paper cites.
In: Robotics and Automation (ICRA), 2016 IEEE International Conference on . IEEE, pp. 512–519
Finn C, Tan XY, Duan Y, Darrell T, Levine S and Abbeel P (2016) Deep spatial autoencoders for visuomotor learning · 2016
Earlier work this paper cites.
In: Intelligent Robots and Systems (IROS), 2016 IEEE/RSJ International Conference on . IEEE, pp. 4019–4026
Fu J, Levine S and Abbeel P (2016) One-shot learning of manipulation skills with online dynamics adaptation and neural network priors · 2016
Earlier work this paper cites.
IEEE Robotics and Automation Letters 1(2): 661–667
Giusti A, Guzzi J, Cireşan DC, He FL, Rodríguez JP, Fontana F, Faessler M, Forster C, Schmidhuber J, Di Caro G et al. (2016) A machine learning approach to visual perception of forest trails for mobile robots · 2016
Earlier work this paper cites.
MIT Press
Goodfellow I, Bengio Y and Courville A (2016) Deep Learning · 2016
Earlier work this paper cites.
In: International Conference on Machine Learning (ICML-16)
Gu S, Lillicrap T, Sutskever I and Levine S (2016) Continuous deep q-learning with model-based acceleration · 2016
Earlier work this paper cites.
Neurocomputing 187: 27–48
Guo Y, Liu Y, Oerlemans A, Lao S, Wu S and Lew MS (2016) Deep learning for visual understanding: A review · 2016
Cited alongside, same era.
In: Intelligent Robots and Systems (IROS), 2016 IEEE/RSJ International Conference on . IEEE, pp. 3786–3793
Gupta A, Eppner C, Levine S and Abbeel P (2016) Learning dexterous manipulation for a soft robotic hand from human demonstrations · 2016
Cited alongside, same era.
In: Proceedings of the IEEE conference on computer vision and pattern recognition . pp. 770–778
He K, Zhang X, Ren S and Sun J (2016) Deep residual learning for image recognition · 2016
Cited alongside, same era.
In: Advances in Neural Information Processing Systems . pp. 4565–4573
Ho J and Ermon S (2016) Generative adversarial imitation learning · 2016
Cited alongside, same era.
arXiv preprint arXiv:1611.05397
Jaderberg M, Mnih V, Czarnecki WM, Schaul T, Leibo JZ, Silver D and Kavukcuoglu K (2016) Reinforcement learning with unsupervised auxiliary tasks · 2016
Cited alongside, same era.
arXiv preprint arXiv:1702.08360
Parisotto E and Salakhutdinov R (2017) Neural map: Structured memory for deep reinforcement learning · 2017
Closest in time.
In: International Conference on Machine Learning (ICML) , volume 2017
Pathak D, Agrawal P, Efros AA and Darrell T (2017) Curiosity-driven exploration by self-supervised prediction · 2017
Closest in time.
arXiv preprint arXiv:1710.06537
Peng XB, Andrychowicz M, Zaremba W and Abbeel P (2017) Sim-to-real transfer of robotic control with dynamics randomization · 2017
Closest in time.
arXiv preprint arXiv:1704.03073
Popov I, Heess N, Lillicrap T, Hafner R, Barth-Maron G, Vecerik M, Lampe T, Tassa Y, Erez T and Riedmiller M (2017) Data-efficient deep reinforcement learning for dexterous manipulation · 2017
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The International Journal of Robotics Research 35(11): 1289–1307
Kretzschmar H, Spies M, Sprunk C and Burgard W (2016) Socially compliant mobile robot navigation via inverse reinforcement learning · 2016
Cited alongside, same era.
arXiv preprint arXiv:1606.02396
Kulkarni TD, Saeedi A, Gautam S and Gershman SJ (2016) Deep successor reinforcement learning · 2016
Cited alongside, same era.
In: Robotics and Automation (ICRA), 2016 IEEE International Conference on . IEEE, pp. 378–383
Kumar V, Todorov E and Levine S (2016) Optimal control with learned local models: Application to dexterous manipulation · 2016
Cited alongside, same era.
The Journal of Machine Learning Research 17(1): 1334–1373
Levine S, Finn C, Darrell T and Abbeel P (2016) End-to-end training of deep visuomotor policies · 2016
Cited alongside, same era.
arXiv preprint arXiv:1611.03673
Mirowski P, Pascanu R, Viola F, Soyer H, Ballard A, Banino A, Denil M, Goroshin R, Sifre L, Kavukcuoglu K et al. (2016) Learning to navigate in complex environments · 2016
Cited alongside, same era.
arXiv preprint arXiv:1602.01783
Mnih V, Badia AP, Mirza M, Graves A, Lillicrap TP, Harley T, Silver D and Kavukcuoglu K (2016) Asynchronous methods for deep reinforcement learning · 2016
Cited alongside, same era.
In: ICRA . pp. 2889–2895
Okal B and Arras KO (2016) Learning socially normative robot navigation behaviors with bayesian inverse reinforcement learning · 2016
Cited alongside, same era.
Ruder M, Dosovitskiy A and Brox T (2017) Artistic style transfer for videos and spherical images · 2017
Closest in time.
In: Conference on Robot Learning . pp. 262–270
Rusu AA, Večerík M, Rothörl T, Heess N, Pascanu R and Hadsell R (2017) Sim-to-real robot learning from pixels with progressive nets · 2017
Closest in time.
arXiv preprint arXiv:1712.03931
Savva M, Chang AX, Dosovitskiy A, Funkhouser T and Koltun V (2017) Minos: Multimodal indoor simulator for navigation in complex environments · 2017
Closest in time.
arXiv preprint arXiv:1707.06347
Schulman J, Wolski F, Dhariwal P, Radford A and Klimov O (2017) Proximal policy optimization algorithms · 2017
Closest in time.
In: Field and Service Robotics
Shah S, Dey D, Lovett C and Kapoor A (2017) Airsim: High-fidelity visual and physical simulation for autonomous vehicles · 2017
Closest in time.
Nature neuroscience 20(11): 1643
Stachenfeld KL, Botvinick MM and Gershman SJ (2017) The hippocampus as a predictive map · 2017
Closest in time.
arXiv preprint arXiv:1703.01703
Stadie BC, Abbeel P and Sutskever I (2017) Third-person imitation learning · 2017
Closest in time.
arXiv preprint arXiv:1703.05407
Sukhbaatar S, Kostrikov I, Szlam A and Fergus R (2017) Intrinsic motivation and automatic curricula via asymmetric self-play · 2017
Closest in time.
In: 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . pp. 31–36
Tai L, Paolo G and Liu M (2017) Virtual-to-real deep reinforcement learning: Continuous control of mobile robots for mapless navigation · 2017
Closest in time.
In: Intelligent Robots and Systems (IROS), 2017 IEEE/RSJ International Conference on . IEEE, pp. 23–30
Tobin J, Fong R, Ray A, Schneider J, Zaremba W and Abbeel P (2017) Domain randomization for transferring deep neural networks from simulation to the real world · 2017
Closest in time.
arXiv preprint arXiv:1707.08817
Večerík M, Hester T, Scholz J, Wang F, Pietquin O, Piot B, Heess N, Rothörl T, Lampe T and Riedmiller M (2017) Leveraging demonstrations for deep reinforcement learning on robotics problems with sparse rewards · 2017
Closest in time.
In: Advances in Neural Information Processing Systems . pp. 5326–5335
Wang Z, Merel JS, Reed SE, de Freitas N, Wayne G and Heess N (2017) Robust imitation of diverse behaviors · 2017
Closest in time.
arXiv preprint arXiv:1707.06203
Weber T, Racanière S, Reichert DP, Buesing L, Guez A, Rezende DJ, Badia AP, Vinyals O, Heess N, Li Y et al. (2017) Imagination-augmented agents for deep reinforcement learning · 2017
Closest in time.
In: Advances in neural information processing systems . pp. 5285–5294
Wu Y, Mansimov E, Grosse RB, Liao S and Ba J (2017) Scalable trust-region method for deep reinforcement learning using kronecker-factored approximation · 2017
Closest in time.
arXiv preprint arXiv:1704.03952
You Y, Pan X, Wang Z and Lu C (2017) Virtual to real reinforcement learning for autonomous driving · 2017
Closest in time.
In: 2017 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . pp. 2371–2378
Zhang J, Springenberg JT, Boedecker J and Burgard W (2017a) Deep reinforcement learning with successor features for navigation across similar environments · 2017
Closest in time.
In: Robotics and Automation (ICRA), 2017 IEEE International Conference on . IEEE, pp. 3357–3364
Zhu Y, Mottaghi R, Kolve E, Lim JJ, Gupta A, Fei-Fei L and Farhadi A (2017b) Target-driven visual navigation in indoor scenes using deep reinforcement learning · 2017
Closest in time.
arXiv preprint arXiv:1801.08214
Chaplot DS, Parisotto E and Salakhutdinov R (2018) Active neural localization · 2018
Closest in time.
In: International Conference on Learning Representations
Fortunato M, Azar MG, Piot B, Menick J, Osband I, Graves A, Mnih V, Munos R, Hassabis D, Pietquin O et al. (2018) Noisy networks for exploration · 2018
Closest in time.
Gao Y, Xu HH, Lin J, Yu F, Levine S and Darrell T (2018) Reinforcement learning from imperfect demonstrations
2018
Closest in time.
arXiv preprint arXiv:1803.02999
Nichol A and Schulman J (2018) Reptile: a scalable metalearning algorithm · 2018
Closest in time.
arXiv preprint arXiv:1802.06857
Parisotto E, Chaplot DS, Zhang J and Salakhutdinov R (2018) Global pose estimation with an attention-based recurrent network · 2018
Closest in time.
In: International Conference on Learning Representations
Plappert M, Houthooft R, Dhariwal P, Sidor S, Chen RY, Chen X, Asfour T, Abbeel P and Andrychowicz M (2018) Parameter space noise for exploration · 2018
Closest in time.
arXiv preprint arXiv:1802.10567
Riedmiller M, Hafner R, Lampe T, Neunert M, Degrave J, Van de Wiele V Tom adn Mnih, Heess N and Springenberg JT (2018) Learning by playing - solving sparse reward tasks from scratch · 2018
Closest in time.
In: International Conference on Learning Representations
Savinov N, Dosovitskiy A and Koltun V (2018) Semi-parametric topological memory for navigation · 2018
Closest in time.
In: Robotics and Automation (ICRA), 2018 IEEE International Conference on
Tai L, Zhang J, Liu M and Burgard W (2018) Socially-compliant navigation through raw depth inputs with generative adversarial imitation learning · 2018
Closest in time.
arXiv preprint arXiv:1801.02209
Wu Y, Wu Y, Gkioxari G and Tian Y (2018) Building generalizable agents with a realistic and rich 3d environment · 2018
Closest in time.
arXiv preprint arXiv:1801.03458
Yang L, Liang X and Xing E (2018) Unsupervised real-to-virtual domain unification for end-to-end highway driving · 2018
Closest in time.
arXiv preprint arXiv:1802.01557
Yu T, Finn C, Xie A, Dasari S, Zhang T, Abbeel P and Levine S (2018) One-shot imitation from observing humans via domain-adaptive meta-learning · 2018
Closest in time.
arXiv preprint arXiv:1802.00265
Zhang J, Tai L, Xiong Y, Liu M, Boedecker J and Burgard W (2018) Vr goggles for robots: Real-to-sim domain adaptation for visual control · 2018
Closest in time.
arXiv preprint arXiv:1802.09564
Zhu Y, Wang Z, Merel J, Rusu A, Erez T, Cabi S, Tunyasuvunakool S, Kramár J, Hadsell R, de Freitas N et al. (2018) Reinforcement and imitation learning for diverse visuomotor skills · 2018
Closest in time.
In: AAAI , volume 16. pp. 2094–2100
Van Hasselt H, Guez A and Silver D (2016) Deep reinforcement learning with double q-learning · 2094
Closest in time.
In: 2016 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . pp. 2096–2101
Pfeiffer M, Schwesinger U, Sommer H, Galceran E and Siegwart R (2016) Predicting actions to act predictably: Cooperative partial motion planning with maximum entropy models · 2096
Closest in time.