Fetching the paper…
Reading the bibliography…
In this report, we describe a Theano-based AlexNet (Krizhevsky et al., 2012) implementation and its naive data parallelism on multiple GPUs.
Gradient-based learning applied to document recognition
LeCun, Yann, Bottou, Léon, Bengio, Yoshua, and Haffner, Patrick · 1998
Earlier work this paper cites.
Theano: a cpu and gpu math expression compiler
Bergstra, James, Breuleux, Olivier, Bastien, Frédéric, Lamblin, Pascal, Pascanu, Razvan, Desjardins, Guillaume, Turian, Joseph, Warde-Farley, David, and Bengio, Yoshua · 2010
Earlier work this paper cites.
Torch7: A matlab-like environment for machine learning
Collobert, Ronan, Kavukcuoglu, Koray, and Farabet, Clément · 2011
Earlier work this paper cites.
Theano: new features and speed improvements
Bastien, Frédéric, Lamblin, Pascal, Pascanu, Razvan, Bergstra, James, Goodfellow, Ian, Bergeron, Arnaud, Bouchard, Nicolas, Warde-Farley, David, and Bengio, Yoshua · 2012
Earlier work this paper cites.
Pycuda and pyopencl: A scripting-based approach to gpu run-time code generation
Klöckner, Andreas, Pinto, Nicolas, Lee, Yunsup, Catanzaro, Bryan, Ivanov, Paul, and Fasih, Ahmed · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoffrey E · 2012
Cited alongside, same era.
Pylearn2: a machine learning research library
Goodfellow, Ian J, Warde-Farley, David, Lamblin, Pascal, Dumoulin, Vincent, Mirza, Mehdi, Pascanu, Razvan, Bergstra, James, Bastien, Frédéric, and Bengio, Yoshua · 2013
Cited alongside, same era.
Gpu asynchronous stochastic gradient descent to speed up neural network training
Paine, Thomas, Jin, Hailin, Yang, Jianchao, Lin, Zhe, and Huang, Thomas · 2013
Cited alongside, same era.
Multi-gpu training of convnets
Yadan, Omry, Adams, Keith, Taigman, Yaniv, and Ranzato, Marc’Aurelio · 2013
Cited alongside, same era.
Caffe: Convolutional architecture for fast feature embedding
Jia, Yangqing, Shelhamer, Evan, Donahue, Jeff, Karayev, Sergey, Long, Jonathan, Girshick, Ross, Guadarrama, Sergio, and Darrell, Trevor · 2014
Closest in time.
One weird trick for parallelizing convolutional neural networks
Krizhevsky, Alex · 2014
Closest in time.
Imagenet large scale visual recognition challenge
Russakovsky, Olga, Deng, Jia, Su, Hao, Krause, Jonathan, Satheesh, Sanjeev, Ma, Sean, Huang, Zhiheng, Karpathy, Andrej, Khosla, Aditya, Bernstein, Michael, et al · 2014
Closest in time.
Mariana: Tencent deep learning platform and its applications
Zou, Yongqiang, Jin, Xing, Li, Yi, Guo, Zhimao, Wang, Eryu, and Xiao, Bin · 2014
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chetlur, Sharan, Woolley, Cliff, Vandermersch, Philippe, Cohen, Jonathan, Tran, John, Catanzaro, Bryan, and Shelhamer, Evan · 2014
Cited alongside, same era.