Fetching the paper…
Reading the bibliography…
We investigate algorithmic progress in image classification on ImageNet, perhaps the most well-known test bed for computer vision.
“Shapley value”
Sergiu Hart · 1989
Earlier work this paper cites.
“Backpropagation applied to handwritten zip code recognition”
Yann LeCun et al · 1989
Earlier work this paper cites.
“Invariant theory of finite groups”
Mara Neusel and Larry Smith · 2002
Earlier work this paper cites.
“Imagenet: A large-scale hierarchical image database”
Jia Deng et al · 2009
Earlier work this paper cites.
“Imagenet classification with deep convolutional neural networks”
Alex Krizhevsky, Ilya Sutskever and Geoffrey Hinton · 2012
Earlier work this paper cites.
“Algorithmic progress in six domains”
Katja Grace · 2013
Earlier work this paper cites.
“Triton: an intermediate language and compiler for tiled neural network computations”
Philippe Tillet, Hsiang-Tsung Kung and David Cox · 2019
Earlier work this paper cites.
“A time leap challenge for SAT-solving”
Johannes Fichte, Markus Hecher and Stefan Szeider · 2020
Cited alongside, same era.
“Measuring the algorithmic efficiency of neural networks”
Danny Hernandez and Tom Brown · 2020
Cited alongside, same era.
“Deepspeed: System optimizations enable training deep learning models with over 100 billion parameters”
Jeff Rasley, Samyam Rajbhandari, Olatunji Ruwase and Yuxiong He · 2020
Cited alongside, same era.
“The computational limits of deep learning”
Neil Thompson, Kristjan Greenewald, Keeheon Lee and Gabriel Manso · 2020
Cited alongside, same era.
“How much chess engine progress is about adapting to bigger computers?” [Online; accessed 21-July-2022], https://www.lesswrong.com/posts/H6L7fuEN9qXDanQ6W/how-much-chess-engine-progress-is-about-adapting-to-bigger , 2021
“How Fast Do Algorithms Improve?[Point of View]”
Yash Sherry and Neil Thompson · 2021
Later among the works it cites.
“The relative importance of hardware and software progress: evidence from computer chess” [Online; accessed 21-Sept-2022], https://bayes.net/computerchess/ , 2022
Tom Adamczewski · 2022
Closest in time.
“Training Compute-Optimal Large Language Models”
Jordan Hoffmann et al · 2022
Closest in time.
“Deep Neural Nets: 33 years ago and 33 years from now” [Online; accessed 21-July-2022], http://karpathy.github.io/2022/03/14/lecun1989/ , 2022
Andrej Karpathy · 2022
Closest in time.
“Progress in mathematical programming solvers from 2001 to 2020”
Thorsten Koch, Timo Berthold, Jaap Pedersen and Charlie Vanaret · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Paul Christiano · 2021
Cited alongside, same era.
“All tokens matter: Token labeling for training better vision transformers”
Zi-Hang Jiang et al · 2021
Cited alongside, same era.
“Estimating training compute of Deep Learning models”, 2022
Jaime Sevilla et al · 2022
Closest in time.