Fetching the paper…
Reading the bibliography…
This paper presents the input convex neural network architecture.
Some methods of speeding up the convergence of iteration methods
Polyak, Boris T · 1964
Earlier work this paper cites.
Projected newton methods for optimization problems with simple constraints
Bertsekas, Dimitri P · 1982
Earlier work this paper cites.
Two-point step size gradient methods
Barzilai, Jonathan and Borwein, Jonathan M · 1988
Earlier work this paper cites.
Learning representations by back-propagating errors
Rumelhart, David E, Hinton, Geoffrey E, and Williams, Ronald J · 1988
Earlier work this paper cites.
Reverse tdnn: an architecture for trajectory generation
Simard, Patrice and LeCun, Yann · 1991
Earlier work this paper cites.
Globally trained handwritten word recognizer using spatial representation, convolutional neural networks, and hidden markov models
Bengio, Yoshua, LeCun, Yann, and Henderson, Donnie · 1994
Earlier work this paper cites.
Parameterisation of a stochastic model for human face identification
Samaria, Ferdinando S and Harter, Andy C · 1994
Earlier work this paper cites.
Python reference manual
Van Rossum, Guido and Drake Jr, Fred L · 1995
Earlier work this paper cites.
Primal-dual interior-point methods
Wright, Stephen J · 1997
Earlier work this paper cites.
Nonmonotone spectral projected gradient methods on convex sets
Birgin, Ernesto G, Martínez, José Mario, and Raydan, Marcos · 2000
Earlier work this paper cites.
Convex optimization
Boyd, Stephen and Vandenberghe, Lieven · 2004
Earlier work this paper cites.
Learning structured prediction models: A large margin approach
Taskar, Ben, Chatalbashev, Vassil, Koller, Daphne, and Guestrin, Carlos · 2005
Earlier work this paper cites.
Large margin methods for structured and interdependent output variables
Tsochantaridis, Ioannis, Joachims, Thorsten, Hofmann, Thomas, and Altun, Yasemin · 2005
Earlier work this paper cites.
A tutorial on energy-based learning
LeCun, Yann, Chopra, Sumit, Hadsell, Raia, Ranzato, M, and Huang, F · 2006
Earlier work this paper cites.
A guide to NumPy , volume 1
Oliphant, Travis E · 2006
Earlier work this paper cites.
(Approximate) subgradient methods for structured prediction
Ratliff, Nathan D, Bagnell, J Andrew, and Zinkevich, Martin · 2007
Cited alongside, same era.
Multilabel text classification for automated tag suggestion
Katakis, Ioannis, Tsoumakas, Grigorios, and Vlahavas, Ioannis · 2008
Cited alongside, same era.
Bundle methods for machine learning
Smola, Alex J., Vishwanathan, S.v.n., and Le, Quoc V · 2008
Cited alongside, same era.
Probabilistic graphical models: principles and techniques
Koller, Daphne and Friedman, Nir · 2009
Cited alongside, same era.
Convex piecewise-linear fitting
Magnani, Alessandro and Boyd, Stephen P · 2009
Cited alongside, same era.
Conditional neural fields
Peng, Jian, Bo, Liefeng, and Xu, Jinbo · 2009
Cited alongside, same era.
Generative adversarial nets
Goodfellow, Ian, Pouget-Abadie, Jean, Mirza, Mehdi, Xu, Bing, Warde-Farley, David, Ozair, Sherjil, Courville, Aaron, and Bengio, Yoshua · 2014
Later among the works it cites.
Adam: A method for stochastic optimization
Kingma, Diederik and Ba, Jimmy · 2014
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, Karen and Zisserman, Andrew · 2014
Later among the works it cites.
Learning deep structured models
Chen, Liang-Chieh, Schwing, Alexander G, Yuille, Alan L, and Urtasun, Raquel · 2015
Later among the works it cites.
Deep residual learning for image recognition
He, Kaiming, Zhang, Xiangyu, Ren, Shaoqing, and Sun, Jian · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Rectified linear units improve restricted boltzmann machines
Nair, Vinod and Hinton, Geoffrey E · 2010
Cited alongside, same era.
Distributed optimization and statistical learning via the alternating direction method of multipliers
Boyd, Stephen, Parikh, Neal, Chu, Eric, Peleato, Borja, and Eckstein, Jonathan · 2011
Cited alongside, same era.
Adaptive subgradient methods for online learning and stochastic optimization
Duchi, John, Hazan, Elad, and Singer, Yoram · 2011
Cited alongside, same era.
Scikit-learn: Machine learning in python
Pedregosa, Fabian, Varoquaux, Gaël, Gramfort, Alexandre, Michel, Vincent, Thirion, Bertrand, Grisel, Olivier, Blondel, Mathieu, Prettenhofer, Peter, Weiss, Ron, Dubourg, Vincent, et al · 2011
Cited alongside, same era.
Sum-product networks: A new deep architecture
Poon, Hoifung and Domingos, Pedro · 2011
Cited alongside, same era.
Mulan: A java library for multi-label learning
Tsoumakas, Grigorios, Spyromitros-Xioufis, Eleftherios, Vilcek, Jozef, and Vlahavas, Ioannis · 2011
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, Sergey and Szegedy, Christian · 2015
Later among the works it cites.
Continuous control with deep reinforcement learning
Lillicrap, Timothy P, Hunt, Jonathan J, Pritzel, Alexander, Heess, Nicolas, Erez, Tom, Tassa, Yuval, Silver, David, and Wierstra, Daan · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
Mnih, Volodymyr, Kavukcuoglu, Koray, Silver, David, Rusu, Andrei A, Veness, Joel, Bellemare, Marc G, Graves, Alex, Riedmiller, Martin, Fidjeland, Andreas K, Ostrovski, Georg, et al · 2015
Later among the works it cites.
Going deeper with convolutions
Szegedy, Christian, Liu, Wei, Jia, Yangqing, Sermanet, Pierre, Reed, Scott, Anguelov, Dragomir, Erhan, Dumitru, Vanhoucke, Vincent, and Rabinovich, Andrew · 2015
Later among the works it cites.
Conditional random fields as recurrent neural networks
Zheng, Shuai, Jayasumana, Sadeep, Romera-Paredes, Bernardino, Vineet, Vibhav, Su, Zhizhong, Du, Dalong, Huang, Chang, and Torr, Philip HS · 2015
Later among the works it cites.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
Abadi, Martın, Agarwal, Ashish, Barham, Paul, Brevdo, Eugene, Chen, Zhifeng, Citro, Craig, Corrado, Greg S, Davis, Andy, Dean, Jeffrey, Devin, Matthieu, et al · 2016
Closest in time.
Structured prediction energy networks
Belanger, David and McCallum, Andrew · 2016
Closest in time.
Brockman, Greg, Cheung, Vicki, Pettersson, Ludwig, Schneider, Jonas, Schulman, John, Tang, Jie, and Zaremba, Wojciech · 2016
Closest in time.
Continuous deep q-learning with model-based acceleration
Gu, Shixiang, Lillicrap, Timothy, Sutskever, Ilya, and Levine, Sergey · 2016
Closest in time.
Densely connected convolutional networks
Huang, Gao, Liu, Zhuang, and Weinberger, Kilian Q · 2016
Closest in time.