Fetching the paper…
Reading the bibliography…
An important goal of AutoML is to automate-away the design of neural networks on new tasks in under-explored domains.
Improved protein structure prediction using potentials from deep learning
Andrew W. Senior, Richard Evans, John Jumper, James Kirkpatrick, Laurent Sifre, Tim Green, Chongli Qin, Augustin Žídek, Alexander W. R. Nelson, Alex Bridgland, Hugo Penedones, Stig Petersen, Karen Simonyan, Steve Crossan, Pushmeet Kohli, David T. Jones, David Silver, Koray Kavukcuoglu, and Demis Hassabis · 1923
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Object recognition with gradient-based learning
Yann LeCun, Patrick Haffner, Léon Bottou, and Yoshua Bengio · 1999
Earlier work this paper cites.
Harmonising chorales by probabilistic inference
Moray Allan and Christopher Williams · 2005
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevksy · 2009
Earlier work this paper cites.
PSICOV: precise structural contact prediction using sparse inverse covariance estimation on large multiple sequence alignments
David T. Jones, Daniel W. A. Buchan, Domenico Cozzetto, and Massimiliano Pontil · 2011
Earlier work this paper cites.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio · 2012
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2015
Earlier work this paper cites.
Butterfly factorization
Yingzhou Li, Haizhao Yang, Eileen R. Martin, Kenneth L. Ho, and Lexing Ying · 2015
Earlier work this paper cites.
ACDC: A structured efficient linear layer
Marcin Moczulski, Misha Denil, Jeremy Appleyard, and Nando de Freitas · 2015
Earlier work this paper cites.
U-Net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Earlier work this paper cites.
ImageNet large scale visual recognition challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, Alexander C. Berg, and Li Fei-Fei · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Wide residual networks
Sergey Zagoruyko and Nikos Komodakis · 2016
Cited alongside, same era.
Semi-supervised classification with graph convolutional networks
Thomas N. Kipf and Max Welling · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
An empirical evaluation of generic convolutional and recurrent networks for sequence modeling
Shaojie Bai, J. Zico Kolter, and Vladlen Koltun · 2018
Cited alongside, same era.
Efficient neural architecture search via parameter sharing
Hieu Pham, Melody Y. Guan, Barret Zoph, Quoc V. Le, and Jeff Dean · 2018
Cited alongside, same era.
DGM: A deep learning algorithm for solving partial differential equations
Justin Sirignano and Konstantinos Spiliopoulos · 2018
Cited alongside, same era.
Sparse linear networks with a fixed butterfly structure: Theory and practice
Nir Ailon, Omer Leibovich, and Vineet Nair · 2020
Later among the works it cites.
Butterfly transform: An efficient FFT based neural architecture design
Keivan Alizadeh vahid, Anish Prabhu, Ali Farhadi, and Mohammad Rastegari · 2020
Later among the works it cites.
Once-for-all: Train one network and specialize it for efficient deployment
Han Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang, and Song Han · 2020
Later among the works it cites.
Kaleidoscope: An efficient, learnable representation for all structured linear maps
Tri Dao, Nimit Sohoni, Albert Gu, Matthew Eichhorn, Amit Blonder, Megan Leszczynski, Atri Rudra, and Christopher Ré · 2020
Later among the works it cites.
NAS-Bench-201: Extending the scope of reproducible neural architecture search
Xuanyi Dong and Yi Yang · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning transferable architectures for scalable image recognition
Barret Zoph, Vijay Vasudevan, Jonathon Shlens, and Quoc V. Le · 2018
Cited alongside, same era.
Trellis networks for sequence modeling
Shaojie Bai, J. Zico Kolter, and Vladlen Koltun · 2019
Cited alongside, same era.
Learning fast algorithms for linear transforms using butterfly factorizations
Tri Dao, Albert Gu, Matthew Eichhorn, Atri Rudra, and Christopher Ré · 2019
Cited alongside, same era.
Neural architecture search: A survey
Thomas Elsken, Jan Hendrik Metzen, and Frank Hutter · 2019
Cited alongside, same era.
Random search and reproducibility for neural architecture search
Liam Li and Ameet Talwalkar · 2019
Cited alongside, same era.
Fast neural architecture search of compact semantic segmentation models via auxiliary cells
Vladimir Nekrasov, Hao Chen, Chunhua Shen, and Ian Reid · 2019
Cited alongside, same era.
Jiemin Fang, Yuzhu Sun, Qian Zhang, Yuan Li, Wenyu Liu, and Xinggang Wang · 2020
Later among the works it cites.
HiPPO: Recurrent memory with optimal polynomial projections
Albert Gu, Tri Dao, Stefano Ermon, Atri Rudra, and Christopher Ré · 2020
Later among the works it cites.
AtomNAS: Fine-grained end-to-end neural architecture search
Jieru Mei, Yingwei Li, Xiaochen Lian, Xiaojie Jin, Linjie Yang, Alan Yuille, and Jianchao Yang · 2020
Later among the works it cites.
AutoML-Zero: Evolving machine learning algorithms from scratch
Esteban Real, Chen Liang, David R. So, and Quoc V. Le · 2020
Later among the works it cites.
PC-DARTS: Partial channel connections for memory-efficient architecture search
Yuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen, Guo-Jun Qi, Qi Tian, and Hongkai Xiong · 2020
Later among the works it cites.
NAS evaluation is frustratingly hard
Antoine Yang, Pedro M. Esperança, and Fabio M. Carlucci · 2020
Later among the works it cites.
NAS-Bench-1Shot1: Benchmarking and dissecting one-shot neural architecture search
Arber Zela, Julien Siems, and Frank Hutter · 2020
Later among the works it cites.
Highly accurate protein structure prediction with alphafold
John Jumper, Richard Evans, Alexander Pritzel, Tim Green, Michael Figurnov, Olaf Ronneberger, Kathryn Tunyasuvunakool, Russ Bates, Augustin Žídek, Anna Potapenko, Alex Bridgland, Clemens Meyer, Simon A. A. Kohl, Andrew J. Ballard, Andrew Cowie, Bernardino Romera-Paredes, Stanislav Nikolov, Rishub Jain, Jonas Adler, Trevor Back, Stig Petersen, David Reiman, Ellen Clancy, Michal Zielinski, Martin Steinegger, Michalina Pacholska, Tamas Berghammer, Sebastian Bodenstein, David Silver, Oriol Vinyals, Andrew W. Senior, Koray Kavukcuoglu, Pushmeet Kohli, and Demis Hassabis · 2021
Closest in time.