Fetching the paper…
Reading the bibliography…
Since its emergence around 2010, deep learning has rapidly become the most important technique in Artificial Intelligence (AI), producing an array of scientific firsts in areas as diverse as protein folding, drug discovery, integrated chip design, and weather prediction.
“Language models are few-shot learners”
Tom Brown et al · 1901
Earlier work this paper cites.
“HuggingFace’s Transformers: State-of-the-art Natural Language Processing”
Thomas Wolf et al · 1910
Earlier work this paper cites.
“The translog function and the substitution of equipment, structures, and labor in U.S. manufacturing 1929-68”
Ernst. Berndt and Laurits. Christensen · 1973
Earlier work this paper cites.
“A Comparison of the Performance of Three Flexible Functional Forms”
David. Guilkey, C.. Lovell and Robin. Sickles · 1983
Earlier work this paper cites.
“Patents and R&D at the Firm Level: A First Look”
Ariel Pakes and Zvi Grilliches · 1984
Earlier work this paper cites.
“Patents and R&D: Is There A Lag?” Series: Working Paper Series, 1988, pp. 265–83
Bronwyn. Hall, Zvi Griliches and Jerry. Hausman · 1988
Earlier work this paper cites.
“Conflict and rent-seeking success functions: Ratio vs. difference models of relative success”
Jack Hirshleifer · 1989
Earlier work this paper cites.
“Endogenous Technological Change” Publisher: The University of Chicago Press
Paul. Romer · 1990
Earlier work this paper cites.
“Nonlinear principal component analysis using autoassociative neural networks”
Mark Kramer · 1991
Earlier work this paper cites.
“Inventions, R&D and industry growth” ISBN: 9798208479254, 1992
Samuel Kortum · 1992
Earlier work this paper cites.
“Equilibrium R&D and the patent–R&D ratio: Us evidence”
Samuel Kortum · 1993
Earlier work this paper cites.
“Endogenous innovation in the theory of growth”
Gene Grossman and Elhanan Helpman · 1994
Earlier work this paper cites.
“R & D-Based Models of Economic Growth” Publisher: The University of Chicago Press
Charles. Jones · 1995
Earlier work this paper cites.
“Difference-form contest success functions and effort levels in contests”
Kyung Baik · 1998
Earlier work this paper cites.
“Capital Accumulation and Innovation as Complementary Factors in Long-Run Growth”
Peter Howitt and Philippe Aghion · 1998
Earlier work this paper cites.
“Recombinant Growth”
Martin. Weitzman · 1998
Earlier work this paper cites.
“Bootstrapping: estimating confidence intervals for cost-effectiveness ratios”
Marion Campbell and David Torgerson · 1999
Earlier work this paper cites.
“Steady Endogenous Growth with Population and R & D Inputs Growing”
Peter Howitt · 1999
Earlier work this paper cites.
“Measuring the” ideas” production function: Evidence from international patent output”
Michael Porter and Scott Stern · 2000
Earlier work this paper cites.
“Scaling Laws for Neural Language Models”, 2020
Jared Kaplan et al · 2001
Earlier work this paper cites.
“Labor mobility from academe to commerce”
Lynne Zucker, Michael Darby and Maximo Torero · 2002
Earlier work this paper cites.
“Endogenous growth: Estimating the Romer model for the US and Germany”
Gang Gong, Alfred Greiner and Willi Semmler · 2004
Earlier work this paper cites.
“An index to quantify an individual’s scientific research output”
Jorge Hirsch · 2005
Earlier work this paper cites.
“The design and analysis of benchmark experiments”
Torsten Hothorn, Friedrich Leisch, Achim Zeileis and Kurt Hornik · 2005
Earlier work this paper cites.
“Relating the knowledge production function to total factor productivity: an endogenous growth puzzle”
Yasser Abdih and Frederick Joutz · 2006
Earlier work this paper cites.
“Reducing the dimensionality of data with neural networks”
Geoffrey Hinton and Ruslan Salakhutdinov · 2006
Earlier work this paper cites.
“Relative significance is insufficient: Baselines matter too”
Timothy Armstrong, Justin Zobel, William Webber and Alistair Moffat · 2009
Earlier work this paper cites.
Stefano Bianchini, Moritz Muller and Pierre Pelletier · 2009
Earlier work this paper cites.
“The knowledge production of ’R’ and ’D”’
Dirk Czarnitzki, Kornelius Kraft and Susanne Thorwarth · 2009
Earlier work this paper cites.
“Imagenet: A large-scale hierarchical image database”
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li and Li Fei-Fei · 2009
Earlier work this paper cites.
“The elements of statistical learning: data mining, inference, and prediction”
Trevor Hastie, Robert Tibshirani, Jerome Friedman and Jerome Friedman · 2009
Earlier work this paper cites.
“A new approach to the metric of journals’ scientific prestige: The SJR indicator”
Borja González-Pereira, Vicente Guerrero-Bote and Félix Moya-Anegón · 2010
Earlier work this paper cites.
“Technologies of conflict”
Hao Jia and Stergios Skaperdas · 2012
Earlier work this paper cites.
“Advanced Macroeconomics, 4e”
David Romer · 2012
Earlier work this paper cites.
“Persuasion as a contest”
Stergios Skaperdas and Samarth Vaidya · 2012
Earlier work this paper cites.
“Representation learning: A review and new perspectives”
Yoshua Bengio, Aaron Courville and Pascal Vincent · 2013
Earlier work this paper cites.
“Matthew: Effect or fable?”
Pierre Azoulay, Toby Stuart and Yanbo Wang · 2014
Earlier work this paper cites.
“Inventor data for research on migration and innovation: a survey and a pilot”
Stefano Breschi, Francesco Lissoni and Gianluca Tarasconi · 2014
Cited alongside, same era.
“Age and scientific genius”, 2014
Benjamin Jones, EJ Reedy and Bruce Weinberg · 2014
Cited alongside, same era.
“Adam: A method for stochastic optimization”
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
“Microsoft coco: Common objects in context”
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár and C Zitnick · 2014
Cited alongside, same era.
“Batch normalization: Accelerating deep network training by reducing internal covariate shift”
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
“Deep learning”
Yann LeCun, Yoshua Bengio and Geoffrey Hinton · 2015
Tamara Broderick, Ryan Giordano and Rachael Meager · 2020
Later among the works it cites.
Ahmed Elnaggar et al · 2020
Later among the works it cites.
“Scaling laws for autoregressive generative modeling”
Tom Henighan et al · 2020
Later among the works it cites.
“A survey of the recent architectures of deep convolutional neural networks”
Asifullah Khan, Anabia Sohail, Umme Zahoora and Aqsa Qureshi · 2020
Later among the works it cites.
“Gshard: Scaling giant models with conditional computation and automatic sharding”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“ImageNet Large Scale Visual Recognition Challenge”
Olga Russakovsky et al · 2015
Cited alongside, same era.
“Contest theory: Incentive mechanisms and ranking methods”
Milan Vojnović · 2015
Cited alongside, same era.
“Deep learning”
Ian Goodfellow, Yoshua Bengio and Aaron Courville · 2016
Cited alongside, same era.
“Deep residual learning for image recognition”
Kaiming He, Xiangyu Zhang, Shaoqing Ren and Jian Sun · 2016
Cited alongside, same era.
“Gaussian error linear units (gelus)”
Dan Hendrycks and Kevin Gimpel · 2016
Cited alongside, same era.
“Introduction by Ajay Agrawal”
Ajay Agrawal · 2017
Cited alongside, same era.
Dmitry Lepikhin, HyoukJoong Lee, Yuanzhong Xu, Dehao Chen, Orhan Firat, Yanping Huang, Maxim Krikun, Noam Shazeer and Zhifeng Chen · 2020
Later among the works it cites.
Zhuohan Li, Eric Wallace, Sheng Shen, Kevin Lin, Kurt Keutzer, Dan Klein and Joseph Gonzalez · 2020
Later among the works it cites.
“Is AI slowing down?”, 2020
Benjamin Muller, Peter McIntyre and Sara Altman · 2020
Later among the works it cites.
“Artificial intelligence: a modern approach, 4th Edition”, 2020
Stuart Russell and Peter Norvig · 2020
Later among the works it cites.
“The cost of training nlp models: A concise overview”
Or Sharir, Barak Peleg and Yoav Shoham · 2020
Later among the works it cites.
“A neural scaling law from the dimension of the data manifold”
Utkarsh Sharma and Jared Kaplan · 2020
Later among the works it cites.
“The computational limits of deep learning”
Neil Thompson, Kristjan Greenewald, Keeheon Lee and Gabriel Manso · 2020
Later among the works it cites.
“Economic growth under transformative AI: A guide to the vast range of possibilities for output growth, wages, and the labor share”, 2020
Philip Trammell and Anton Korinek · 2020
Later among the works it cites.
“AI Feynman: A physics-inspired method for symbolic regression” Publisher: American Association for the Advancement of Science
Silviu-Marian Udrescu and Max Tegmark · 2020
Later among the works it cites.
“Explaining neural scaling laws”
Yasaman Bahri, Ethan Dyer, Jared Kaplan, Jaehoon Lee and Utkarsh Sharma · 2021
Later among the works it cites.
“Artificial intelligence as a general-purpose technology: an historical perspective”
Nicholas Crafts · 2021
Later among the works it cites.
“Scaling Scaling Laws with Board Games”
Andy Jones · 2021
Later among the works it cites.
“Highly accurate protein structure prediction with AlphaFold”
John Jumper et al · 2021
Later among the works it cites.
“Research community dynamics behind popular AI benchmarks”
Fernando Martinez-Plumed, Pablo Barredo, Sean Heigeartaigh and Jose Hernandez-Orallo · 2021
Later among the works it cites.
“A graph placement methodology for fast chip design” Number: 7862 Publisher: Nature Publishing Group
Azalia Mirhoseini et al · 2021
Later among the works it cites.
“Deep double descent: Where bigger models and more data hurt”
Preetum Nakkiran, Gal Kaplun, Yamini Bansal, Tristan Yang, Boaz Barak and Ilya Sutskever · 2021
Later among the works it cites.
“Higher Education Research and Development: Fiscal Year 2020: Data Tables”
National Center for Science and Engineering Statistics · 2021
Later among the works it cites.
David Picard · 2021
Later among the works it cites.
“AI and the everything in the whole wide world benchmark”
Inioluwa Raji, Emily Bender, Amandalynne Paullada, Emily Denton and Alex Hanna · 2021
Later among the works it cites.
URL: https://arxiv.org/
“arXiv.org”, 2022 · 2022
Closest in time.
“Total factor productivity”
San Fed · 2022
Closest in time.
“Could Machine Learning be a General Purpose Technology? A Comparison of Emerging Technologies Using Data from Online Job Postings”
Avi Goldfarb, Bledi Taska and Florenta Teodoridis · 2022
Closest in time.
“Training Compute-Optimal Large Language Models”
Jordan Hoffmann et al · 2022
Closest in time.
“Chemformer: a pre-trained transformer for computational chemistry”
Ross Irwin, Spyridon Dimitriadis, Jiazhen He and Esben Bjerrum · 2022
Closest in time.
“Competition-level code generation with alphacode”
Yujia Li et al · 2022
Closest in time.
In Microsoft Research , 2022
“Project Academic Knowledge” · 2022
Closest in time.
Scott Reed et al · 2022
Closest in time.
URL: https://www.scopus.com/home.uri
“Scopus”, 2022 · 2022
Closest in time.
“Compute Trends Across Three Eras of Machine Learning”
Jaime Sevilla, Lennart Heim, Anson Ho, Tamay Besiroglu, Marius Hobbhahn and Pablo Villalobos · 2022
Closest in time.
“Estimating training compute of Deep Learning models”, 2022
Jaime Sevilla, Lennart Heim, Anson Ho, Marius Hobbhahn, Tamay Besiroglu and Pablo Villalobos · 2022
Closest in time.
“The importance of (exponentially more) computing power”
Neil Thompson, Shuning Ge and Gabriel Manso · 2022
Closest in time.
“Gross Domestic Product: Implicit Price Deflator” Publisher: FRED, Federal Reserve Bank of St. Louis
U.S. Bureau of Economic Analysis · 2022
Closest in time.
“My precious! The location and diffusion of scientific research: evidence from the Synchrotron Diamond Light Source”
Christian Helmers and Henry Overman · 2040
Closest in time.