Fetching the paper…
Reading the bibliography…
Recent advances demonstrate that irregularly wired neural networks from Neural Architecture Search (NAS) and Random Wiring can not only automate the design of deep neural networks but also emit models that outperform previous manual designs.
Automatic model selection for neural networks
Laredo, D., Qin, Y., Schütze, O., and Sun, J.-Q · 1905
Earlier work this paper cites.
Zhang, T., Yang, Y., Yan, F., Li, S., Teague, H., Chen, Y., et al · 1906
Earlier work this paper cites.
Dynamic programming treatment of the traveling salesman problem
Bellman, R. E · 1961
Earlier work this paper cites.
A dynamic programming approach to sequencing problems
Held, M. and Karp, R. M · 1962
Earlier work this paper cites.
Topological sorting of large networks
Kahn, A. B · 1962
Earlier work this paper cites.
A study of replacement algorithms for a virtual-storage computer
Belady, L. A · 1966
Earlier work this paper cites.
Dynamic programming
Bellman, R · 1966
Earlier work this paper cites.
Code generation for a one-register machine
Bruno, J. and Sethi, R · 1976
Earlier work this paper cites.
On the complexity of scheduling problems for parallel/pipelined machines
Bernstein, D., Rodeh, M., and Gertner, I · 1989
Earlier work this paper cites.
Optimal brain damage
LeCun, Y., Denker, J. S., and Solla, S. A · 1990
Earlier work this paper cites.
Term graph rewriting
Plump, D · 1999
Earlier work this paper cites.
Optimal instruction scheduling using integer programming
Wilken, K., Liu, J., and Heffernan, M · 2000
Earlier work this paper cites.
A dynamic programming approach to optimal integrated code generation
Keßler, C. and Bednarski, A · 2001
Earlier work this paper cites.
Compiler optimization on vliw instruction scheduling for low power
Lee, C., Lee, J. K., Hwang, T., and Tsai, S.-C · 2003
Earlier work this paper cites.
LLVM: A compilation framework for lifelong program analysis & transformation
Lattner, C. and Adve, V · 2004
Earlier work this paper cites.
Graph rewriting for hardware dependent program optimizations
Schösser, A. and Geiß, R · 2007
Earlier work this paper cites.
Dadiannao: A machine-learning supercomputer
Chen, Y., Luo, T., Liu, S., Zhang, S., He, L., Wang, J., Li, L., Chen, T., Xu, Z., Sun, N., et al · 2014
Earlier work this paper cites.
Caffe: Convolutional architecture for fast feature embedding
Jia, Y., Shelhamer, E., Donahue, J., Karayev, S., Long, J., Girshick, R., Guadarrama, S., and Darrell, T · 2014
Earlier work this paper cites.
Rigid-motion scattering for image classification
Sifre, L. and Mallat, S · 2014
Earlier work this paper cites.
Efficient and robust automated machine learning
Feurer, M., Klein, A., Eggensperger, K., Springenberg, J., Blum, M., and Hutter, F · 2015
Earlier work this paper cites.
Learning both weights and connections for efficient neural network
Han, S., Pool, J., Tran, J., and Dally, W · 2015
Earlier work this paper cites.
Tensorflow: A system for large-scale machine learning
Abadi, M., Barham, P., Chen, J., Chen, Z., Davis, A., Dean, J., Devin, M., Ghemawat, S., Irving, G., Isard, M., et al · 2016
Cited alongside, same era.
Eyeriss: An energy-efficient reconfigurable accelerator for deep convolutional neural networks
Chen, Y.-H., Krishna, T., Emer, J. S., and Sze, V · 2016
Cited alongside, same era.
Courbariaux, M., Hubara, I., Soudry, D., El-Yaniv, R., and Bengio, Y · 2016
Cited alongside, same era.
Stripes: Bit-serial deep neural network computing
Judd, P., Albericio, J., Hetherington, T., Aamodt, T. M., and Moshovos, A · 2016
Cited alongside, same era.
DoReFa-Net: Training low bitwidth convolutional neural networks with low bitwidth gradients
Zhou, S., Wu, Y., Ni, Z., Zhou, X., Wen, H., and Zou, Y · 2016
AMC: AutoML for model compression and acceleration on mobile devices
He, Y., Lin, J., Liu, Z., Wang, H., Li, L.-J., and Han, S · 2018
Later among the works it cites.
Gist: Efficient data encoding for deep neural network training
Jain, A., Phanishayee, A., Mars, J., Tang, L., and Pekhimenko, G · 2018
Later among the works it cites.
Exploring hidden dimensions in parallelizing convolutional neural networks
Jia, Z., Lin, S., Qi, C. R., and Aiken, A · 2018
Later among the works it cites.
Apprentice: Using knowledge distillation techniques to improve low-precision network accuracy
Mishra, A. and Marr, D · 2018
Later among the works it cites.
Glow: Graph lowering compiler techniques for neural networks
Rotem, N., Fix, J., Abdulrasool, S., Catron, G., Deng, S., Dzhabarov, R., Gibson, N., Hegeman, J., Lele, M., Levenstein, R., et al · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Structured pruning of deep convolutional neural networks
Anwar, S., Hwang, K., and Sung, W · 2017
Cited alongside, same era.
AdaNet: Adaptive structural learning of artificial neural networks
Cortes, C., Gonzalvo, X., Kuznetsov, V., Mohri, M., and Yang, S · 2017
Cited alongside, same era.
Machine learning for systems and systems for machine learning
Dean, J · 2017
Cited alongside, same era.
TETRIS: Scalable and efficient neural network acceleration with 3d memory
Gao, M., Pu, J., Yang, X., Horowitz, M., and Kozyrakis, C · 2017
Cited alongside, same era.
Low-power image recognition challenge
Gauen, K., Rangan, R., Mohan, A., Lu, Y.-H., Liu, W., and Berg, A. C · 2017
Cited alongside, same era.
MobileNets: Efficient convolutional neural networks for mobile vision applications
Howard, A. G., Zhu, M., Chen, B., Kalenichenko, D., Wang, W., Weyand, T., Andreetto, M., and Adam, H · 2017
Cited alongside, same era.
In-datacenter performance analysis of a tensor processing unit
Jouppi, N. P., Young, C., Patil, N., Patterson, D., Agrawal, G., Bajwa, R., Bates, S., Bhatia, S., Boden, N., Borchers, A., et al · 2017
Cited alongside, same era.
Bit Fusion: Bit-level dynamically composable architecture for accelerating deep neural networks
Sharma, H., Park, J., Suda, N., Lai, L., Chau, B., Chandra, V., and Esmaeilzadeh, H · 2018
Later among the works it cites.
Tensor Comprehensions: Framework-agnostic high-performance machine learning abstractions
Vasilache, N., Zinenko, O., Theodoridis, T., Goyal, P., DeVito, Z., Moses, W. S., Verdoolaege, S., Adams, A., and Cohen, A · 2018
Later among the works it cites.
To prune, or not to prune: exploring the efficacy of pruning for model compression
Zhu, M. and Gupta, S · 2018
Later among the works it cites.
Learning transferable architectures for scalable image recognition
Zoph, B., Vasudevan, V., Shlens, J., and Le, Q. V · 2018
Later among the works it cites.
ProxylessNAS: Direct neural architecture search on target task and hardware
Cai, H., Zhu, L., and Han, S · 2019
Later among the works it cites.
Optimizing dnn computation with relaxed graph substitutions
Jia, Z., Thomas, J., Warszawski, T., Gao, M., Zaharia, M., and Aiken, A · 2019
Later among the works it cites.
PyTorch: An imperative style, high-performance deep learning library
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., et al · 2019
Later among the works it cites.
Regularized evolution for image classifier architecture search
Real, E., Aggarwal, A., Huang, Y., and Le, Q. V · 2019
Later among the works it cites.
HAQ: Hardware-aware automated quantization with mixed precision
Wang, K., Liu, Z., Lin, Y., Lin, J., and Han, S · 2019
Later among the works it cites.
Discovering neural wirings
Wortsman, M., Farhadi, A., and Rastegari, M · 2019
Later among the works it cites.
Machine learning at facebook: Understanding inference at the edge
Wu, C.-J., Brooks, D., Chen, K., Chen, D., Choudhury, S., Dukhan, M., Hazelwood, K., Isaac, E., Jia, Y., Jia, B., et al · 2019
Later among the works it cites.
Exploring randomly wired neural networks for image recognition
Xie, S., Kirillov, A., Girshick, R., and He, K · 2019
Later among the works it cites.
Chameleon: Adaptive code optimization for expedited deep neural network compilation
Ahn, B. H., Pilligundla, P., and Esmaeilzadeh, H · 2020
Closest in time.
Learned step size quantization
Esser, S. K., McKinstry, J. L., Bablani, D., Appuswamy, R., and Modha, D. S · 2020
Closest in time.
Shredder: Learning noise distributions to protect inference privacy
Mireshghallah, F., Taram, M., Ramrakhyani, P., Jalali, A., Tullsen, D., and Esmaeilzadeh, H · 2020
Closest in time.