Fetching the paper…
Reading the bibliography…
Large scale cloud services use Key Performance Indicators (KPIs) for tracking and monitoring performance.
Induction of Decision Trees. Machine Learning 1
J Ross Quinlan. 1986 · 1986
Earlier work this paper cites.
Random Forests
Leo Breiman. 2001 · 2001
Earlier work this paper cites.
Fault detection by mining association rules from house-keeping data. In proceedings of the 6th International Symposium on Artificial Intelligence, Robotics and Automation in Space , Vol. 18. Citeseer, 21
Takehisa Yairi, Yoshikiyo Kato, and Koichi Hori. 2001 · 2001
Earlier work this paper cites.
Failure diagnosis using decision trees. In International Conference on Autonomic Computing, 2004. Proceedings. 36–43
M. Chen, A. X. Zheng, J. Lloyd, M. I. Jordan, and E. Brewer. 2004 · 2004
Earlier work this paper cites.
Correlating Instrumentation Data to System States: A Building Block for Automated Diagnosis and Control.. In OSDI , Vol. 4. 16–16
Ira Cohen, Jeffrey S Chase, Moises Goldszmidt, Terence Kelly, and Julie Symons. 2004 · 2004
Earlier work this paper cites.
Randomized trees for human pose detection. In 2008 IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 1–8
Grégory Rogez, Jonathan Rihan, Srikumar Ramalingam, Carlos Orrite, and Philip HS Torr. 2008 · 2008
Earlier work this paper cites.
Anomaly extraction in backbone networks using association rules. In Proceedings of the 9th ACM SIGCOMM conference on Internet measurement . ACM, 28–34
Daniela Brauckhoff, Xenofontas Dimitropoulos, Arno Wagner, and Kavè Salamatian. 2009 · 2009
Earlier work this paper cites.
Fa: A system for automating failure diagnosis. In 2009 IEEE 25th International Conference on Data Engineering . IEEE, 1012–1023
Songyun Duan, Shivnath Babu, and Kamesh Munagala. 2009 · 2009
Earlier work this paper cites.
Speed Matters
Google.com. 2019 · 2009
Earlier work this paper cites.
Customer churn prediction using improved balanced random forests
Yaya Xie, Xiu Li, EWT Ngai, and Weiyun Ying. 2009 · 2009
Earlier work this paper cites.
Detecting large-scale system problems by mining console logs. In Proceedings of the ACM SIGOPS 22nd symposium on Operating systems principles . ACM, 117–132
Wei Xu, Ling Huang, Armando Fox, David Patterson, and Michael I Jordan. 2009 · 2009
Earlier work this paper cites.
Fingerprinting the datacenter: automated classification of performance crises. In Proceedings of the 5th European conference on Computer systems . ACM, 111–124
Peter Bodik, Moises Goldszmidt, Armando Fox, Dawn B Woodard, and Hans Andersen. 2010 · 2010
Earlier work this paper cites.
Predicting the severity of a reported bug. In 2010 7th IEEE Working Conference on Mining Software Repositories (MSR 2010) . IEEE, 1–10
Ahmed Lamkanfi, Serge Demeyer, Emanuel Giger, and Bart Goethals. 2010 · 2010
Cited alongside, same era.
Log filtering and interpretation for root cause analysis. In 2010 IEEE International Conference on Software Maintenance . IEEE, 1–5
Hamzeh Zawawy, Kostas Kontogiannis, and John Mylopoulos. 2010 · 2010
Cited alongside, same era.
Algorithms for hyper-parameter optimization. In Advances in neural information processing systems . 2546–2554
James S Bergstra, Rémi Bardenet, Yoshua Bengio, and Balázs Kégl. 2011 · 2011
Cited alongside, same era.
An empirical evaluation of the comprehensibility of decision table, tree and rule based predictive models
Johan Huysmans, Karel Dejaeger, Christophe Mues, Jan Vanthienen, and Bart Baesens. 2011 · 2011
Cited alongside, same era.
Random search for hyper-parameter optimization
James Bergstra and Yoshua Bengio. 2012 · 2012
Near Real-Time Service Monitoring Using High-Dimensional Time Series. In 2015 IEEE International Conference on Data Mining Workshop (ICDMW) . IEEE, 1624–1627
Shwetabh Khanduja, Vinod Nair, S Sundararajan, Ameya Raul, Ajesh Babu Shaj, and Sathiya Keerthi. 2015 · 2015
Later among the works it cites.
LogCluster-A data clustering and pattern mining algorithm for event logs. In 2015 11th International Conference on Network and Service Management (CNSM) . IEEE, 1–7
Risto Vaarandi and Mauno Pihelgas. 2015 · 2015
Later among the works it cites.
Random forest modeling for network intrusion detection system
Nabila Farnaaz and MA Jabbar. 2016 · 2016
Later among the works it cites.
Evaluation metrics of service-level reliability monitoring rules of a big data service. In 2016 IEEE 27th International Symposium on Software Reliability Engineering (ISSRE) . IEEE, 376–387
Keun Soo Yim. 2016 · 2016
Later among the works it cites.
A parallel random forest algorithm for big data in a spark cloud computing environment
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Latent fault detection in large scale services. In IEEE/IFIP International Conference on Dependable Systems and Networks (DSN 2012) . IEEE, 1–12
Moshe Gabel, Assaf Schuster, Ran-Gilad Bachrach, and Nikolaj Bjørner. 2012 · 2012
Cited alongside, same era.
Scalable random forests for massive data. In Pacific-Asia Conference on Knowledge Discovery and Data Mining . Springer, 135–146
Bingguo Li, Xiaojun Chen, Mark Junjie Li, Joshua Zhexue Huang, and Shengzhong Feng. 2012 · 2012
Cited alongside, same era.
Structured comparative analysis of systems logs to diagnose performance problems. In Proceedings of the 9th USENIX conference on Networked Systems Design and Implementation . USENIX Association, 26–26
Karthik Nagaraj, Charles Killian, and Jennifer Neville. 2012 · 2012
Cited alongside, same era.
Improved duplicate bug report identification. In 2012 16th European Conference on Software Maintenance and Reengineering . IEEE, 385–390
Yuan Tian, Chengnian Sun, and David Lo. 2012 · 2012
Cited alongside, same era.
Efficient Bug Triaging Using Text Mining
Mamdouh Alenezi, Kenneth Magel, and Shadi Banitaan. 2013 · 2013
Cited alongside, same era.
A scalable random forest algorithm based on mapreduce. In 2013 IEEE 4th International Conference on Software Engineering and Service Science . IEEE, 849–852
Jiawei Han, Yanheng Liu, and Xin Sun. 2013 · 2013
Cited alongside, same era.
Experience report: Anomaly detection of cloud application operations using log and cloud metric correlation analysis. In 2015 IEEE 26th International Symposium on Software Reliability Engineering (ISSRE) . IEEE, 24–34
Mostafa Farshchi, Jean-Guy Schneider, Ingo Weber, and John Grundy. 2015 · 2015
Cited alongside, same era.
Jianguo Chen, Kenli Li, Zhuo Tang, Kashif Bilal, Shui Yu, Chuliang Weng, and Keqin Li. 2017 · 2017
Later among the works it cites.
Deeplog: Anomaly detection and diagnosis from system logs through deep learning. In Proceedings of the 2017 ACM SIGSAC Conference on Computer and Communications Security . ACM, 1285–1298
Min Du, Feifei Li, Guineng Zheng, and Vivek Srikumar. 2017 · 2017
Later among the works it cites.
Improving automated bug triaging with specialized topic model
Xin Xia, David Lo, Ying Ding, Jafar M Al-Kofahi, Tien N Nguyen, and Xinyu Wang. 2017 · 2017
Later among the works it cites.
Troubleshooting Transiently-Recurring Errors in Production Systems with Blame-Proportional Logging. In 2018 USENIX Annual Technical Conference (USENIX ATC 18) . 321–334
Liang Luo, Suman Nath, Lenin Ravindranath Sivalingam, Madan Musuvathi, and Luis Ceze. 2018 · 2018
Later among the works it cites.
Amazon AWA SLA
Amazon.com. 2019 · 2019
Closest in time.
Power BI
PowerBI.com. 2019 · 2019
Closest in time.
Learning a hierarchical monitoring system for detecting and diagnosing service issues. In Proceedings of the 21th ACM SIGKDD International Conference on Knowledge Discovery and Data Mining . ACM, 2029–2038
Vinod Nair, Ameya Raul, Shwetabh Khanduja, Vikas Bahirwani, Qihong Shao, Sundararajan Sellamanickam, Sathiya Keerthi, Steve Herbert, and Sudheer Dhulipalla. 2015 · 2038
Closest in time.