Fetching the paper…
Reading the bibliography…
The problem of continuous inverse optimal control (over finite time horizon) is to learn the unknown cost function over the sequence of continuous control variables from expert demonstrations.
H. Robbins and S. Monro, “A stochastic approximation method,” The annals of mathematical statistics , pp. 400–407, 1951
1951
Earlier work this paper cites.
W. K. Hastings, “Monte carlo sampling methods using markov chains and their applications,” 1970
1970
Earlier work this paper cites.
Y. LeCun, Y. Bengio et al. , “Convolutional networks for images, speech, and time series,” The handbook of brain theory and neural networks , vol. 3361, no. 10, p. 1995, 1995
1995
Earlier work this paper cites.
S. C. Zhu, Y. Wu, and D. Mumford, “Filters, random fields and maximum entropy (frame): Towards a unified theory for texture modeling,” International Journal of Computer Vision (IJCV) , vol. 27, no. 2, pp. 107–126, 1998
1998
Earlier work this paper cites.
A. Bemporad, M. Morari, V. Dua, and E. N. Pistikopoulos, “The explicit linear quadratic regulator for constrained systems,” Automatica , vol. 38, no. 1, pp. 3–20, 2002
2002
Earlier work this paper cites.
G. E. Hinton, “Training products of experts by minimizing contrastive divergence,” Neural Computation , vol. 14, no. 8, pp. 1771–1800, 2002
2002
Earlier work this paper cites.
W. Li and E. Todorov, “Iterative linear quadratic regulator design for nonlinear biological movement systems.” in Proceedings of the First International Conference on Informatics in Control, Automation and Robotics (ICINCO) , 2004, pp. 222–229
2004
Earlier work this paper cites.
A. Hyvärinen, “Estimation of non-normalized statistical models by score matching,” Journal of Machine Learning Research , vol. 6, pp. 695–709, 2005
2005
Earlier work this paper cites.
E. Todorov, “Optimal control theory,” Bayesian brain: probabilistic approaches to neural coding , pp. 269–298, 2006
2006
Earlier work this paper cites.
M. Richardson and P. Domingos, “Markov logic networks,” Machine learning , vol. 62, no. 1-2, pp. 107–136, 2006
2006
Earlier work this paper cites.
T. M. Cover and J. A. Thomas, Elements of information theory, Second Edition . Wiley, 2006
2006
Earlier work this paper cites.
J. Colyar and J. Halkias, “US highway 101 dataset,” vol. Federal Highway Administration (FHWA), Tech. Rep. FHWA-HRT-07-030, 2007
2007
Earlier work this paper cites.
B. D. Ziebart, A. L. Maas, J. A. Bagnell, and A. K. Dey, “Maximum entropy inverse reinforcement learning.” in Proceedings of the Twenty-Third AAAI Conference on Artificial Intelligence (AAAI) , vol. 8. Chicago, IL, USA, 2008, pp. 1433–1438
2008
Earlier work this paper cites.
R. M. Neal et al. , “Mcmc using hamiltonian dynamics,” Handbook of markov chain monte carlo , vol. 2, no. 11, p. 2, 2011
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in neural information processing systems (NIPS) , 2012, pp. 1097–1105
2012
Earlier work this paper cites.
S. Levine and V. Koltun, “Continuous inverse optimal control with locally optimal examples,” in International Conference on Machine Learning (ICML) , 2012, pp. 475–482
2012
Earlier work this paper cites.
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio, “Generative adversarial nets,” in Advances in neural information processing systems (NIPS) , 2014, pp. 2672–2680
2014
Earlier work this paper cites.
D. P. Kingma and M. Welling, “Auto-encoding variational bayes,” in International Conference on Learning Representations (ICLR) , 2014
2014
Earlier work this paper cites.
T. Chen, E. B. Fox, and C. Guestrin, “Stochastic gradient hamiltonian monte carlo,” in International Conference on Machine Learning (ICML) , vol. 32, 2014, pp. 1683–1691
2014
Earlier work this paper cites.
P. J. Bickel and K. A. Doksum, Mathematical statistics: basic ideas and selected topics, volumes I-II package . CRC Press, 2015
2015
Earlier work this paper cites.
2015
Cited alongside, same era.
J. Xie, W. Hu, S.-C. Zhu, and Y. N. Wu, “Learning sparse frame models for natural image patterns,” International Journal of Computer Vision (IJCV) , vol. 114, no. 2-3, pp. 91–112, 2015
2015
Cited alongside, same era.
M. Monfort, A. Liu, and B. D. Ziebart, “Intent prediction and trajectory forecasting via predictive inverse linear-quadratic regulation,” in Proceedings of the Twenty-Ninth AAAI Conference on Artificial Intelligence (AAAI) , 2015, pp. 3672–3678
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in IEEE International Conference on Computer Vision (ICCV) , 2015, pp. 1026–1034
2015
Cited alongside, same era.
J. Xie, Z. Zheng, R. Gao, W. Wang, S. Zhu, and Y. N. Wu, “Learning descriptor networks for 3d shape synthesis and analysis,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018, pp. 8629–8638
2018
Later among the works it cites.
A. Gupta, J. Johnson, L. Fei-Fei, S. Savarese, and A. Alahi, “Social gan: Socially acceptable trajectories with generative adversarial networks,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Later among the works it cites.
A. Vemula, K. Muelling, and J. Oh, “Social attention: Modeling attention in human crowds,” in Proceedings of the International Conference on Robotics and Automation (ICRA) 2018 , May 2018
2018
Later among the works it cites.
N. Deo, A. Rangesh, and M. M. Trivedi, “How would surround vehicles move? a unified framework for maneuver classification and motion prediction,” IEEE Transactions on Intelligent Vehicles , vol. 3, no. 2, pp. 129–140, 2018
2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz, “Trust region policy optimization,” in International conference on machine learning (ICML) , 2015, pp. 1889–1897
2015
Cited alongside, same era.
L. Dinh, D. Krueger, and Y. Bengio, “NICE: non-linear independent components estimation,” in International Conference on Learning Representations (ICLR) Workshop , 2015
2015
Cited alongside, same era.
J. Xie, Y. Lu, S.-C. Zhu, and Y. Wu, “A theory of generative convnet,” in International Conference on Machine Learning (ICML) , 2016, pp. 2635–2644
2016
Cited alongside, same era.
C. Finn, S. Levine, and P. Abbeel, “Guided cost learning: Deep inverse optimal control via policy optimization,” in International Conference on Machine Learning (ICML) , 2016, pp. 49–58
2016
Cited alongside, same era.
2016
Cited alongside, same era.
J. Ho and S. Ermon, “Generative adversarial imitation learning,” in Advances in Neural Information Processing Systems (NIPS) , 2016, pp. 4565–4573
2016
Cited alongside, same era.
A. Alahi, K. Goel, V. Ramanathan, A. Robicquet, L. Fei-Fei, and S. Savarese, “Social lstm: Human trajectory prediction in crowded spaces,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016
2016
Cited alongside, same era.
J. Xie, S.-C. Zhu, and Y. N. Wu, “Synthesizing dynamic patterns by spatial-temporal generative convnet,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2017, pp. 7093–7101
2017
Cited alongside, same era.
Later among the works it cites.
D. P. Kingma and P. Dhariwal, “Glow: Generative flow with invertible 1x1 convolutions,” Advances in neural information processing systems (NeurIPS) , vol. 31, 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
R. P. Bhattacharyya, D. J. Phillips, B. Wulfe, J. Morton, A. Kuefler, and M. J. Kochenderfer, “Multi-agent imitation learning for driving simulation,” in International Conference on Intelligent Robots and Systems (IROS) , 2018, pp. 1534–1539
2018
Later among the works it cites.
J. Xie, S.-C. Zhu, and Y. N. Wu, “Learning energy-based spatial-temporal generative convnets for dynamic patterns,” IEEE transactions on pattern analysis and machine intelligence (TPAMI) , 2019
2019
Closest in time.
T. Zhao, Y. Xu, M. Monfort, W. Choi, C. Baker, Y. Zhao, Y. Wang, and Y. N. Wu, “Multi-agent tensor fusion for contextual trajectory prediction,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , June 2019
2019
Closest in time.
J. Xie, R. Gao, Z. Zheng, S.-C. Zhu, and Y. N. Wu, “Learning dynamic generator model by alternating back-propagation through time,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , vol. 33, 2019, pp. 5498–5507
2019
Closest in time.
R. P. Bhattacharyya, D. J. Phillips, C. Liu, J. K. Gupta, K. Driggs-Campbell, and M. J. Kochenderfer, “Simulating emergent properties of human driving behavior using multi-agent reward augmented imitation learning,” in Proceedings of the International Conference on Robotics and Automation (ICRA) , May 2019
2019
Closest in time.
J. Xie, Z. Zheng, R. Gao, W. Wang, S.-C. Zhu, and Y. N. Wu, “Generative voxelnet: Learning energy-based models for 3d shape synthesis and analysis,” IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) , 2020
2020
Closest in time.
E. Nijkamp, M. Hill, T. Han, S. Zhu, and Y. N. Wu, “On the anatomy of mcmc-based maximum likelihood learning of energy-based models,” in Proceedings of the AAAI Conference on Artificial Intelligence (AAAI) , vol. 34, no. 04, 2020, pp. 5272–5280
2020
Closest in time.
J. Xie, Z. Zheng, X. Fang, S.-C. Zhu, and Y. N. Wu, “Cooperative training of fast thinking initializer and slow thinking solver for conditional learning,” IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) , 2021
2021
Closest in time.
J. Xie, Y. Xu, Z. Zheng, S. Zhu, and Y. N. Wu, “Generative pointnet: Deep energy-based learning on unordered point sets for 3d generation, reconstruction and classification,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 14 976–14 985
2021
Closest in time.
J. Xie, Z. Zheng, X. Fang, S. Zhu, and Y. N. Wu, “Learning cycle-consistent cooperative networks via alternating MCMC teaching for unsupervised cross-domain translation,” in Thirty-Fifth AAAI Conference on Artificial Intelligence (AAAI) , 2021, pp. 10 430–10 440
2021
Closest in time.
J. Xie, Z. Zheng, and P. Li, “Learning energy-based model with variational auto-encoder as amortized sampler,” in Thirty-Fifth AAAI Conference on Artificial Intelligence (AAAI) , 2021, pp. 10 441–10 451
2021
Closest in time.
J. Xie, Y. Zhu, J. Li, and P. Li, “A tale of two flows: Cooperative learning of langevin flow and normalizing flow toward energy-based model,” in International Conference on Learning Representations (ICLR) , 2022
2022
Closest in time.
L. Dinh, J. Sohl-Dickstein, and S. Bengio, “Density estimation using real NVP,” in International Conference on Learning Representations (ICLR) , 2017
2022
Closest in time.