Fetching the paper…
Reading the bibliography…
For deep neural network accelerators, memory movement is both energetically expensive and can bound computation.
The theory of dynamic programming
Richard Bellman · 1954
Earlier work this paper cites.
Optimization by simulated annealing
Scott Kirkpatrick, C. Gelatt, and Mario Vechhi · 1983
Earlier work this paper cites.
An overview of evolutionary computation
William M Spears, Kenneth A De Jong, Thomas Bäck, David B Fogel, and Hugo De Garis · 1993
Earlier work this paper cites.
Unbounded knapsack problem: Dynamic programming revisited
Rumen Andonov, Vincent Poirriez, and Sanjay Rajopadhye · 2000
Earlier work this paper cites.
A robust optimization approach to supply chain management
Dimitris Bertsimas and Aurélie Thiele · 2004
Earlier work this paper cites.
Evolutionary computation: toward a new philosophy of machine intelligence , volume 1
David B Fogel · 2006
Earlier work this paper cites.
Neuroevolution: from architectures to learning
Dario Floreano, Peter Dürr, and Claudio Mattiussi · 2008
Earlier work this paper cites.
The graph neural network model
Franco Scarselli, Marco Gori, Ah Chung Tsoi, Markus Hagenbuchner, and Gabriele Monfardini · 2008
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
Brian D Ziebart, Andrew L Maas, J Andrew Bagnell, and Anind K Dey · 2008
Earlier work this paper cites.
Large scale distributed deep networks
Jeffrey Dean, Greg Corrado, Rajat Monga, Kai Chen, Matthieu Devin, Mark Mao, Marc’aurelio Ranzato, Andrew Senior, Paul Tucker, Ke Yang, et al · 2012
Earlier work this paper cites.
Tensorflow: A system for large-scale machine learning
Martín Abadi, Paul Barham, Jianmin Chen, Zhifeng Chen, Andy Davis, Jeffrey Dean, Matthieu Devin, Sanjay Ghemawat, Geoffrey Irving, Michael Isard, et al · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Dlau: A scalable deep learning accelerator unit on fpga
Chao Wang, Lei Gong, Qi Yu, Xi Li, Yuan Xie, and Xuehai Zhou · 2016
Earlier work this paper cites.
An alternative softmax operator for reinforcement learning
Kavosh Asadi and Michael L Littman · 2017
Earlier work this paper cites.
Deep steering: Learning end-to-end driving model from spatial and temporal visual cues
Lu Chi and Yadong Mu · 2017
Earlier work this paper cites.
In-datacenter performance analysis of a tensor processing unit
Norman P Jouppi, Cliff Young, Nishant Patil, David Patterson, Gaurav Agrawal, Raminder Bajwa, Sarah Bates, Suresh Bhatia, Nan Boden, Al Borchers, et al · 2017
Cited alongside, same era.
Continual and one-shot learning through neural networks with dynamic external memory
Benno Lüders, Mikkel Schläger, Aleksandra Korach, and Sebastian Risi · 2017
Cited alongside, same era.
Device placement optimization with reinforcement learning
Azalia Mirhoseini, Hieu Pham, Quoc V. Le, Benoit Steiner, Rasmus Larsen, Yuefeng Zhou, Naveen Kumar, Mohammad Norouzi, Samy Bengio, and Jeff Dean · 2017
Cited alongside, same era.
Placeto: Efficient progressive device placement optimization
Ravichandra Addanki, Shaileshh Bojja Venkatakrishnan, Shreyan Gupta, Hongzi Mao, and Mohammad Alizadeh · 2018
Cited alongside, same era.
Intel ngraph: An intermediate representation, compiler, and executor for deep learning
Docbert: Bert for document classification
Ashutosh Adhikari, Achyudh Ram, Raphael Tang, and Jimmy Lin · 2019
Later among the works it cites.
Hongyang Gao and Shuiwang Ji · 2019
Later among the works it cites.
Collaborative evolutionary reinforcement learning
Shauharda Khadka, Somdeb Majumdar, Tarek Nassar, Zach Dwiel, Evren Tumer, Santiago Miret, Yinyin Liu, and Kagan Tumer · 2019
Later among the works it cites.
Vijay Janapa Reddi, Christine Cheng, David Kanter, Peter Mattson, Guenther Schmuelling, Carole-Jean Wu, Brian Anderson, Maximilien Breughe, Mark Charlebois, William Chou, et al · 2019
Later among the works it cites.
Applying deep learning to the cache replacement problem
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Scott Cyphers, Arjun K Bansal, Anahita Bhiwandiwalla, Jayaram Bobba, Matthew Brookhart, Avijit Chakraborty, Will Constable, Christian Convey, Leona Cook, Omar Kanawi, et al · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
Addressing function approximation error in actor-critic methods
Scott Fujimoto, Herke van Hoof, and Dave Meger · 2018
Cited alongside, same era.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine · 2018
Cited alongside, same era.
Evolution-guided policy gradient in reinforcement learning
Shauharda Khadka and Kagan Tumer · 2018
Cited alongside, same era.
Umap: Uniform manifold approximation and projection
Leland McInnes, John Healy, Nathaniel Saul, and Lukas Großberger · 2018
Cited alongside, same era.
A hierarchical model for device placement
Azalia Mirhoseini, Anna Goldie, Hieu Pham, Benoit Steiner, Quoc V. Le, and Jeff Dean · 2018
Cited alongside, same era.
Jongsoo Park, Maxim Naumov, Protonu Basu, Summer Deng, Aravind Kalaiah, Daya Khudia, James Law, Parth Malani, Andrey Malevich, Satish Nadathur, et al · 2018
Cited alongside, same era.
Zhan Shi, Xiangru Huang, Akanksha Jain, and Calvin Lin · 2019
Later among the works it cites.
Bert rediscovers the classical nlp pipeline
Ian Tenney, Dipanjan Das, and Ellie Pavlick · 2019
Later among the works it cites.
Spring hill (nnp-i 1000) intel’s data center inference chip
O. Wechsler, M. Behar, and B. Daga · 2019
Later among the works it cites.
Simple applications of bert for ad hoc document retrieval
Wei Yang, Haotian Zhang, and Jimmy Lin · 2019
Later among the works it cites.
Guided interactive learning through chatbot using bi-directional encoder representations from transformers (bert)
Richeeka Bathija, Pranav Agarwal, Rakshith Somanna, and GB Pallavi · 2020
Closest in time.
Intel ® Nervana ™ NNP-I shows best-in-class throughput on BERT NLP model
Guy Boudoukh, Eli Kfir, Ofir Zafrir, Uzi Sarel, Michael Behar, Moshe Wasserblat, Galina Ryvchin, Peter Adams, and Kiran Atmakuri · 2020
Closest in time.
Mlperf: An industry standard benchmark suite for machine learning performance
Peter Mattson, Vijay Janapa Reddi, Christine Cheng, Cody Coleman, Greg Diamos, David Kanter, Paulius Micikevicius, David Patterson, Guenther Schmuelling, Hanlin Tang, et al · 2020
Closest in time.
Chip placement with deep reinforcement learning
Azalia Mirhoseini, Anna Goldie, Mustafa Yazgan, Joe Jiang, Ebrahim Songhori, Shen Wang, Young-Joon Lee, Eric Johnson, Omkar Pathak, Sungmin Bae, Azade Nazi, Jiwoo Pak, Andy Tong, Kavya Srinivasa, William Hang, Emre Tuncer, Anand Babu, Quoc V. Le, James Laudon, Richard Ho, Roger Carpenter, and Jeff Dean · 2020
Closest in time.
Reinforced genetic algorithm learning for optimizing computation graphs
Aditya Paliwal, Felix Gimeno, Vinod Nair, Yujia Li, Miles Lubin, Pushmeet Kohli, and Oriol Vinyals · 2020
Closest in time.
A comprehensive survey on graph neural networks
Zonghan Wu, Shirui Pan, Fengwen Chen, Guodong Long, Chengqi Zhang, and S Yu Philip · 2020
Closest in time.
Optimal data placement for heterogeneous cache, memory, and storage systems
Lei Zhang, Reza Karimi, Irfan Ahmad, and Ymir Vigfusson · 2020
Closest in time.