Fetching the paper…
Reading the bibliography…
The ability of neural networks to perform robotic perception and control tasks such as depth and optical flow estimation, simultaneous localization and mapping (SLAM), and automatic control has led to their widespread adoption in recent years.
A. W. Moore, “Efficient memory-based learning for robot control,” University of Cambridge, Computer Laboratory, Tech. Rep., 1990
1990
Earlier work this paper cites.
M. Mundhenk, J. Goldsmith, C. Lusena, and E. Allender, “Complexity of finite-horizon markov decision process problems,” Journal of the ACM (JACM) , vol. 47, no. 4, pp. 681–720, 2000
2000
Earlier work this paper cites.
C. Lusena, J. Goldsmith, and M. Mundhenk, “Nonapproximability results for partially observable markov decision processes,” Journal of artificial intelligence research , vol. 14, pp. 83–103, 2001
2001
Earlier work this paper cites.
R. Coulom, “Reinforcement learning using neural networks, with applications to motor control,” Ph.D. dissertation, Institut National Polytechnique de Grenoble-INPG, 2002
2002
Earlier work this paper cites.
D. Koller and N. Friedman, Probabilistic graphical models: principles and techniques . MIT press, 2009
2009
Earlier work this paper cites.
D. G. Myers, “Intuition’s powers and perils,” Psychological Inquiry , vol. 21, no. 4, pp. 371–377, 2010
2010
Earlier work this paper cites.
M. Deisenroth and C. E. Rasmussen, “Pilco: A model-based and data-efficient approach to policy search,” in Proceedings of the 28th International Conference on machine learning (ICML-11) , 2011, pp. 465–472
2011
Earlier work this paper cites.
I. Osband, D. Russo, and B. Van Roy, “(more) efficient reinforcement learning via posterior sampling,” Advances in Neural Information Processing Systems , vol. 26, 2013
2013
Earlier work this paper cites.
I. Osband and B. Van Roy, “Near-optimal reinforcement learning in factored mdps,” Advances in Neural Information Processing Systems , vol. 27, 2014
2014
Earlier work this paper cites.
M. T. Ribeiro, S. Singh, and C. Guestrin, “” why should i trust you?” explaining the predictions of any classifier,” in Proceedings of the 22nd ACM SIGKDD international conference on knowledge discovery and data mining , 2016, pp. 1135–1144
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
Y. Gal, R. McAllister, and C. E. Rasmussen, “Improving pilco with bayesian neural network dynamics models,” in Data-efficient machine learning workshop, ICML , vol. 4, no. 34, 2016, p. 25
2016
Earlier work this paper cites.
T. Haarnoja, H. Tang, P. Abbeel, and S. Levine, “Reinforcement learning with deep energy-based policies,” in International conference on machine learning . PMLR, 2017, pp. 1352–1361
2017
Earlier work this paper cites.
T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor,” in International conference on machine learning . PMLR, 2018, pp. 1861–1870
2018
Cited alongside, same era.
K. Chua, R. Calandra, R. McAllister, and S. Levine, “Deep reinforcement learning in a handful of trials using probabilistic dynamics models,” Advances in neural information processing systems , vol. 31, 2018
2018
Cited alongside, same era.
A. Z. Zhu, L. Yuan, K. Chaney, and K. Daniilidis, “Unsupervised event-based learning of optical flow, depth, and egomotion,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2019
2019
Cited alongside, same era.
A. Kirillov, K. He, R. Girshick, C. Rother, and P. Dollar, “Panoptic segmentation,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2019
2019
Cited alongside, same era.
T. Gebru, J. Morgenstern, B. Vecchione, J. W. Vaughan, H. Wallach, H. D. Iii, and K. Crawford, “Datasheets for datasets,” Communications of the ACM , vol. 64, no. 12, pp. 86–92, 2021
2021
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
C. Lee, A. K. Kosta, and K. Roy, “Fusion-flownet: Energy-efficient optical flow estimation using sensor fusion and deep fused spiking-analog network architectures,” in 2022 International Conference on Robotics and Automation (ICRA) . IEEE, 2022, pp. 6504–6510
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Haarnoja, S. Ha, A. Zhou, J. Tan, G. Tucker, and S. Levine, “Learning to walk via deep reinforcement learning,” Robotics: Science and Systems , 2019
2019
Cited alongside, same era.
M. Janner, J. Fu, M. Zhang, and S. Levine, “When to trust your model: Model-based policy optimization,” Advances in neural information processing systems , vol. 32, 2019
2019
Cited alongside, same era.
J. Czarnowski, T. Laidlow, R. Clark, and A. J. Davison, “Deepfactors: Real-time probabilistic dense monocular slam,” IEEE Robotics and Automation Letters , vol. 5, no. 2, pp. 721–728, 2020
2020
Cited alongside, same era.
O. M. Andrychowicz, B. Baker, M. Chociej, R. Jozefowicz, B. McGrew, J. Pachocki, A. Petron, M. Plappert, G. Powell, A. Ray, et al. , “Learning dexterous in-hand manipulation,” The International Journal of Robotics Research , vol. 39, no. 1, pp. 3–20, 2020
2020
Cited alongside, same era.
D. Hafner, T. Lillicrap, J. Ba, and M. Norouzi, “Dream to control: Learning behaviors by latent imagination,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=S1lOTC4tDS
2020
Cited alongside, same era.
Łukasz Kaiser, M. Babaeizadeh, P. Miłos, B. Osiński, R. H. Campbell, K. Czechowski, D. Erhan, C. Finn, P. Kozakowski, S. Levine, A. Mohiuddin, R. Sepassi, G. Tucker, and H. Michalewski, “Model based reinforcement learning for atari,” in International Conference on Learning Representations , 2020. [Online]. Available: https://openreview.net/forum?id=S1xCPJHtDB
2020
Cited alongside, same era.
2020
Cited alongside, same era.
M. Gehrig, M. Millhäusler, D. Gehrig, and D. Scaramuzza, “E-raft: Dense optical flow from event cameras,” in International Conference on 3D Vision (3DV) , 2021
2021
Cited alongside, same era.
Z. Huang, X. Shi, C. Zhang, Q. Wang, K. C. Cheung, H. Qin, J. Dai, and H. Li, “Flowformer: A transformer architecture for optical flow,” in Computer Vision – ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23–27, 2022, Proceedings, Part XVII , 2022, p. 668–685. [Online]. Available: https://doi.org/10.1007/978-3-031-19790-1_40
2022
Later among the works it cites.
W. Ponghiran, C. M. Liyanagedera, and K. Roy, “Event-based temporally dense optical flow estimation with sequential learning,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, pp. 9827–9836
2023
Later among the works it cites.
A. Kirillov, E. Mintun, N. Ravi, H. Mao, C. Rolland, L. Gustafson, T. Xiao, S. Whitehead, A. C. Berg, W.-Y. Lo, P. Dollar, and R. Girshick, “Segment anything,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , October 2023, pp. 4015–4026
2023
Later among the works it cites.
2024
Closest in time.
R. Rahman, D. Owen, and J. You, “Tracking large-scale ai models,” 2024, accessed: 2024-09-03. [Online]. Available: https://epochai.org/blog/tracking-large-scale-ai-models
2024
Closest in time.
M. Mutti, R. D. Santi, M. Restelli, A. Marx, and G. Ramponi, “Exploiting causal graph priors with posterior sampling for reinforcement learning,” in The Twelfth International Conference on Learning Representations , 2024. [Online]. Available: https://openreview.net/forum?id=M0xK8nPGvt
2024
Closest in time.
“Proximal policy optimization,” https://stable-baselines3.readthedocs.io/en/master/modules/ppo.html , accessed: 2024-09-14
2024
Closest in time.
“Deep q-network,” https://stable-baselines3.readthedocs.io/en/master/modules/dqn.html , accessed: 2024-09-14
2024
Closest in time.
“Lunar lander environment,” https://gymnasium.farama.org/environments/box2d/lunar_lander/ , accessed: 2024-09-14
2024
Closest in time.