Fetching the paper…
Reading the bibliography…
Machine learning (ML) models are increasingly trained in clusters with non-dedicated workers possessing heterogeneous resources.
Learning hierarchical features for scene labeling
Clement Farabet, Camille Couprie, Laurent Najman, and Yann LeCun. 2013 · 1929
Earlier work this paper cites.
On the order determination of ARIMA models
T Ozaki. 1977 · 1977
Earlier work this paper cites.
Optimal static load balancing in distributed computer systems
Asser N Tantawi and Don Towsley. 1985 · 1985
Earlier work this paper cites.
Recurrent neural networks and robust time series prediction
Jerome T Connor, R Douglas Martin, and Les E Atlas. 1994 · 1994
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Scheduling multithreaded computations by work stealing
Robert D Blumofe and Charles E Leiserson. 1999 · 1999
Earlier work this paper cites.
Optimizing static job scheduling in a network of heterogeneous computers. In IEEE ICPP
Xueyan Tang and Samuel T Chanson. 2000 · 2000
Earlier work this paper cites.
GPFS: A Shared-Disk File System for Large Computing Clusters.. In USENIX FAST
Frank B Schmuck and Roger L Haskin. 2002 · 2002
Earlier work this paper cites.
NARMAX-model-based time series modeling and prediction: feedforward and recurrent fuzzy neural network approaches. In WSEAS CSECS
Yang Gao and Meng Joo Er. 2003 · 2003
Earlier work this paper cites.
ARIMA forecasting of primary energy demand by fuel in Turkey
Volkan Ş Ediger and Sertac Akar. 2007 · 2007
Earlier work this paper cites.
On early stopping in gradient descent learning
Yuan Yao, Lorenzo Rosasco, and Andrea Caponnetto. 2007 · 2007
Earlier work this paper cites.
The use of NARX neural networks to predict chaotic time series
Eugen Diaconescu. 2008 · 2008
Earlier work this paper cites.
Improving MapReduce performance in heterogeneous environments.. In USENIX OSDI
Matei Zaharia, Andy Konwinski, Anthony D Joseph, Randy H Katz, and Ion Stoica. 2008 · 2008
Earlier work this paper cites.
Scalable work stealing. In IEEE/ACM SC
James Dinan, D Brian Larkins, Ponnuswamy Sadayappan, Sriram Krishnamoorthy, and Jarek Nieplocha. 2009 · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky and Geoffrey Hinton. 2009 · 2009
Earlier work this paper cites.
Slow learners are fast. In NIPS
John Langford, Alexander J Smola, and Martin Zinkevich. 2009 · 2009
Earlier work this paper cites.
Identifying suspicious URLs: an application of large-scale online learning. In ACM ICML
Justin Ma, Lawrence K Saul, Stefan Savage, and Geoffrey M Voelker. 2009 · 2009
Earlier work this paper cites.
MapReduce: a flexible data processing tool
Jeffrey Dean and Sanjay Ghemawat. 2010 · 2010
Earlier work this paper cites.
The hadoop distributed file system. In IEEE MSST
Konstantin Shvachko, Hairong Kuang, Sanjay Radia, and Robert Chansler. 2010 · 2010
Earlier work this paper cites.
Natural language processing (almost) from scratch
Ronan Collobert, Jason Weston, Léon Bottou, Michael Karlen, Koray Kavukcuoglu, and Pavel Kuksa. 2011 · 2011
Earlier work this paper cites.
Load balanced min-min algorithm for static meta-task scheduling in grid computing
T Kokilavani, Dr DI George Amalarethinam, et al · 2011
Earlier work this paper cites.
A survey of load balancing in cloud computing: Challenges and algorithms. In IEEE NCA
Klaithem Al Nuaimi, Nader Mohamed, Mariam Al Nuaimi, and Jameela Al-Jaroodi. 2012 · 2012
Earlier work this paper cites.
Large scale distributed deep networks. In NIPS
Jeffrey Dean, Greg Corrado, Rajat Monga, Kai Chen, Matthieu Devin, Mark Mao, Andrew Senior, Paul Tucker, Ke Yang, Quoc V Le, and Andrew Ng. 2012 · 2012
Earlier work this paper cites.
ImageNet classification with deep convolutional neural networks. In NIPS
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton. 2012 · 2012
Earlier work this paper cites.
Heterogeneity and dynamicity of clouds at scale: Google trace analysis. In ACM SoCC
Charles Reiss, Alexey Tumanov, Gregory R Ganger, Randy H Katz, and Michael A Kozuch. 2012 · 2012
Earlier work this paper cites.
Resilient distributed datasets: A fault-tolerant abstraction for in-memory cluster computing. In USENIX NSDI
Matei Zaharia, Mosharaf Chowdhury, Tathagata Das, Ankur Dave, Justin Ma, Murphy McCauley, Michael J Franklin, Scott Shenker, and Ion Stoica. 2012 · 2012
Earlier work this paper cites.
Scheduling parallel programs by work stealing with private deques. In ACM SIGPLAN Notices
Umut A Acar, Arthur Charguéraud, and Mike Rainey. 2013 · 2013
Earlier work this paper cites.
Effective Straggler Mitigation: Attack of the Clones.. In USENIX NSDI
Ganesh Ananthanarayanan, Ali Ghodsi, Scott Shenker, and Ion Stoica. 2013 · 2013
Earlier work this paper cites.
More effective distributed ML via a stale synchronous parallel parameter server. In NIPS
Qirong Ho, James Cipar, Henggang Cui, Seunghak Lee, Jin Kyu Kim, Phillip B Gibbons, Garth A Gibson, Greg Ganger, and Eric P Xing. 2013 · 2013
Cited alongside, same era.
Reducing GPU offload latency via fine-grained CPU-GPU synchronization. In IEEE HPCA
Daniel Lustig and Margaret Martonosi. 2013 · 2013
Cited alongside, same era.
Analysis of variants in Round Robin Algorithms for load balancing in Cloud Computing
Pooja Samal and Pranati Mishra. 2013 · 2013
Cited alongside, same era.
Discretized streams: Fault-tolerant streaming computation at scale. In ACM SOSP
Matei Zaharia, Tathagata Das, Haoyuan Li, Timothy Hunter, Scott Shenker, and Ion Stoica. 2013 · 2013
Cited alongside, same era.
Project adam: Building an efficient and scalable deep learning training system. In USENIX OSDI
Trishul Chilimbi, Yutaka Suzue, Johnson Apacible, and Karthik Kalyanaraman. 2014 · 2014
Cited alongside, same era.
Accurate, Large Minibatch SGD: Training ImageNet in 1 Hour
Priya Goyal, Piotr Dollár, Ross Girshick, Pieter Noordhuis, Lukasz Wesolowski, Aapo Kyrola, Andrew Tulloch, Yangqing Jia, and Kaiming He. 2017 · 2017
Later among the works it cites.
DeepProf: Performance Analysis for Deep Learning Applications via Mining GPU Execution Patterns
Jiazhen Gu, Huan Liu, Yangfan Zhou, and Xin Wang. 2017 · 2017
Later among the works it cites.
Heterogeneity-aware distributed parameter servers. In ACM SIGMOD
Jiawei Jiang, Bin Cui, Ce Zhang, and Lele Yu. 2017 · 2017
Later among the works it cites.
Deep learning for short-term traffic flow prediction
Nicholas G Polson and Vadim O Sokolov. 2017 · 2017
Later among the works it cites.
Deep Learning for Fixed Model Reuse.. In AAAI . 2831–2837
Yang Yang, De-Chuan Zhan, Ying Fan, Yuan Jiang, and Zhi-Hua Zhou. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Exploiting Bounded Staleness to Speed Up Big Data Analytics. In USENIX ATC
Henggang Cui, James Cipar, Qirong Ho, Jin Kyu Kim, Seunghak Lee, Abhimanu Kumar, Jinliang Wei, Wei Dai, Gregory R Ganger, Phillip B Gibbons, et al · 2014
Cited alongside, same era.
Adaptive stream processing using dynamic batch sizing. In ACM SoCC
Tathagata Das, Yuan Zhong, Ion Stoica, and Scott Shenker. 2014 · 2014
Cited alongside, same era.
Caffe: Convolutional architecture for fast feature embedding. In ACM Multimedia
Yangqing Jia, Evan Shelhamer, Jeff Donahue, Sergey Karayev, Jonathan Long, Ross Girshick, Sergio Guadarrama, and Trevor Darrell. 2014 · 2014
Cited alongside, same era.
Exploiting GPU hardware saturation for fast compiler optimization. In ACM GPGPU
Alberto Magni, Christophe Dubach, and Michael O’Boyle. 2014 · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks. In NIPS
Ilya Sutskever, Oriol Vinyals, and Quoc V Le. 2014 · 2014
Cited alongside, same era.
Storm@ twitter. In ACM SIGMOD
Ankit Toshniwal, Siddarth Taneja, Amit Shukla, Karthik Ramasamy, Jignesh M Patel, Sanjeev Kulkarni, Jason Jackson, Krishna Gade, Maosong Fu, Jake Donham, et al · 2014
Cited alongside, same era.
Mxnet: A flexible and efficient machine learning library for heterogeneous distributed systems
Tianqi Chen, Mu Li, Yutian Li, Min Lin, Naiyan Wang, Minjie Wang, Tianjun Xiao, Bing Xu, Chiyuan Zhang, and Zheng Zhang. 2015 · 2015
Cited alongside, same era.
Applied Machine Learning at Facebook: A Datacenter Infrastructure Perspective. In IEEE HPCA
Kim Hazelwood, Sarah Bird, David Brooks, Soumith Chintala, Utku Diril, Dmytro Dzhulgakov, Mohamed Fawzy, Bill Jia, Yangqing Jia, Aditya Kalro, et al · 2018
Closest in time.
Multi-tenant GPU Clusters for Deep Learning Workloads: Analysis and Implications
Myeongjae Jeon, Shivaram Venkataraman, Junjie Qian, Amar Phanishayee, Wencong Xiao, and Fan Yang. 2018 · 2018
Closest in time.
Xianyan Jia, Shutao Song, Wei He, Yangzihao Wang, Haidong Rong, Feihu Zhou, Liqiang Xie, Zhenyu Guo, Yuanzhou Yang, Liwei Yu, et al · 2018
Closest in time.
Yuriy Kochura, Yuri Gordienko, Vlad Taran, Nikita Gordienko, Alexandr Rokovyi, Oleg Alienin, and Sergii Stirenko. 2018 · 2018
Closest in time.
Ease. ml: towards multi-tenant resource sharing for machine learning workloads
Tian Li, Jie Zhong, Ji Liu, Wentao Wu, and Ce Zhang. 2018 · 2018
Closest in time.
Optimus: an efficient dynamic resource scheduler for deep learning clusters. In ACM Eurosys
Yanghua Peng, Yixin Bao, Yangrui Chen, Chuan Wu, and Chuanxiong Guo. 2018 · 2018
Closest in time.
Throughput optimizations for FPGA-based deep neural network inference
Thorbjörn Posewsky and Daniel Ziener. 2018 · 2018
Closest in time.
Gandiva: introspective cluster scheduling for deep learning. In USENIX OSDI
Wencong Xiao, Romil Bhardwaj, Ramachandran Ramjee, Muthian Sivathanu, Nipun Kwatra, Zhenhua Han, Pratyush Patel, Xuan Peng, Hanyu Zhao, Quanlu Zhang, et al · 2018
Closest in time.
Scheduling parallel computations by work stealing: A survey
Jixiang Yang and Qingbi He. 2018 · 2018
Closest in time.
Stress-ng: a tool to load and stress a computer system
2019 · 2019
Closest in time.
Train Deep Learning Models on GPUs using Amazon EC2 Spot Instances
2019 · 2019
Closest in time.
Tiresias: A GPU Cluster Manager for Distributed Deep Learning. In USENIX NSDI
Juncheng Gu, Mosharaf Chowdhury, Kang G Shin, Yibo Zhu, Myeongjae Jeon, Junjie Qian, Hongqiang Liu, and Chuanxiong Guo. 2019 · 2019
Closest in time.
Apache Thrift
2020 · 2020
Closest in time.
EC2 Spot Instances
2020 · 2020
Closest in time.
FloydHub
2020 · 2020
Closest in time.
GlusterFS
2020 · 2020
Closest in time.
Python Psutil
2020 · 2020
Closest in time.
Wonder Shaper
2020 · 2020
Closest in time.
Balancing efficiency and fairness in heterogeneous GPU clusters for deep learning
Shubham Chaudhary, Ramachandran Ramjee, Muthian Sivathanu, Nipun Kwatra, and Srinidhi Viswanatha. 2020 · 2020
Closest in time.
AlloX: compute allocation in hybrid clusters
Tan Le, Xiao Shu Sun, Mosharaf Chowdhury, and Zhenhua Liu. 2020 · 2020
Closest in time.
Themis: Fair and Efficient GPU Cluster Scheduling. In NSDI
Kshiteej Mahajan, Arjun Balasubramanian, Arjun Singhvi, Shivaram Venkataraman, Aditya Akella, Amar Phanishayee, and Shuchi Chawla. 2020 · 2020
Closest in time.
Heterogeneity-Aware Cluster Scheduling Policies for Deep Learning Workloads. In USENIX OSDI
Deepak Narayanan, Keshav Santhanam, Fiodar Kazhamiaka, Amar Phanishayee, and Matei Zaharia. 2020 · 2020
Closest in time.
HetPipe: Enabling Large DNN Training on (Whimpy) Heterogeneous GPU Clusters through Integration of Pipelined Model Parallelism and Data Parallelism. In USENIX ATC
Jay H Park, Gyeongchan Yun, Chang M Yi, Nguyen T Nguyen, Seungmin Lee, Jaesik Choi, Sam H Noh, and Young-ri Choi. 2020 · 2020
Closest in time.