Fetching the paper…
Reading the bibliography…
UniLog: Deploy One Model and Specialize it for All Log Analysis Tasks
Arithmetic coding for data compression
Ian H Witten, Radford M Neal, and John G Cleary · 1987
Earlier work this paper cites.
Textrank: Bringing order into text
Rada Mihalcea and Paul Tarau · 2004
Earlier work this paper cites.
What supercomputers say: A study of five system logs
Adam Oliner and Jon Stearley · 2007
Earlier work this paper cites.
Using hidden semi-markov models for effective online failure prediction
Felix Salfner and Miroslaw Malek · 2007
Earlier work this paper cites.
Predicting computer system failures using support vector machines
Errin W Fulp, Glenn A Fink, and Jereme N Haack · 2008
Earlier work this paper cites.
Largescale system problem detection by mining console logs
Wei Xu, Ling Huang, Armando Fox, David Patterson, and Michael Jordan · 2009
Earlier work this paper cites.
Mining invariants from console logs for system problem detection
Jian-Guang Lou, Qiang Fu, Shengqi Yang, Ye Xu, and Jiang Li · 2010
Earlier work this paper cites.
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov · 2014
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Jimmy Lei Ba, Jamie Ryan Kiros, and Geoffrey E Hinton · 2016
Earlier work this paper cites.
Experience report: System log analysis for anomaly detection
Shilin He, Jieming Zhu, et al · 2016
Earlier work this paper cites.
Log clustering based problem identification for online service systems
Qingwei Lin, Hongyu Zhang, Jian-Guang Lou, et al · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, et al · 2017
Earlier work this paper cites.
Deeplog: Anomaly detection and diagnosis from system logs through deep learning
Min Du, Feifei Li, Guineng Zheng, and Vivek Srikumar · 2017
Earlier work this paper cites.
Language modeling with gated convolutional networks
Yann N Dauphin, Angela Fan, Michael Auli, and David Grangier · 2017
Cited alongside, same era.
Searching for activation functions
Prajit Ramachandran, Barret Zoph, and Quoc V Le · 2017
Cited alongside, same era.
Prefix: Switch failure prediction in datacenter networks
Shenglin Zhang, Ying Liu, Weibin Meng, Zhiling Luo, Jiahao Bu, Sen Yang, Peixian Liang, Dan Pei, Jun Xu, Yuzhi Zhang, et al · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
Universal language model fine-tuning for text classification
Jeremy Howard and Sebastian Ruder · 2018
Cited alongside, same era.
A survey on automated log analysis for reliability engineering
Shilin He, Pinjia He, Zhuangbin Chen, Tianyi Yang, Yuxin Su, and Michael R Lyu · 2020
Later among the works it cites.
Summarizing unstructured logs in online services
Weibin Meng, Federico Zaiter, Yuheng Huang, Ying Liu, Shenglin Zhang, Yuzhe Zhang, Yichen Zhu, Tianke Zhang, En Wang, Zuomin Ren, et al · 2020
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Later among the works it cites.
Language models are few-shot learners
Tom B Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Fixing weight decay regularization in adam
Ilya Loshchilov and Frank Hutter · 2018
Cited alongside, same era.
The cost of downtime at the world’s biggest online retailer
UpGuard · 2019
Cited alongside, same era.
Loganomaly: Unsupervised detection of sequential and quantitative anomalies in unstructured logs
Weibin Meng, Ying Liu, Yichen Zhu, Shenglin Zhang, et al · 2019
Cited alongside, same era.
Logzip: extracting hidden structures via iterative clustering for log compression
Jinyang Liu, Jieming Zhu, Shilin He, Pinjia He, Zibin Zheng, and Michael R Lyu · 2019
Cited alongside, same era.
Fine-tune bert for extractive summarization
Yang Liu · 2019
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, et al · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Cited alongside, same era.
Logparse: Making log parsing adaptive through word classification
Weibin Meng, Ying Liu, Federico Zaiter, Shenglin Zhang, Yihao Chen, Yuzhe Zhang, Yichen Zhu, En Wang, Ruizhi Zhang, Shimin Tao, et al · 2020
Later among the works it cites.
On layer normalization in the transformer architecture
Ruibin Xiong, Yunchang Yang, Di He, Kai Zheng, Shuxin Zheng, Chen Xing, Huishuai Zhang, Yanyan Lan, Liwei Wang, and Tieyan Liu · 2020
Later among the works it cites.
Loghub: a large collection of system log datasets towards automated log analytics
Shilin He, Jieming Zhu, Pinjia He, and Michael R Lyu · 2020
Later among the works it cites.
A simple framework for contrastive learning of visual representations
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton · 2020
Later among the works it cites.
Logtransfer: Cross-system log anomaly detection for software systems with transfer learning
Rui Chen, Shenglin Zhang, Dongwen Li, Yuzhe Zhang, Fangrui Guo, Weibin Meng, Dan Pei, Yuzhi Zhang, Xu Chen, and Yuqing Liu · 2020
Later among the works it cites.
Do transformers really perform bad for graph representation?
Chengxuan Ying, Tianle Cai, Shengjie Luo, Shuxin Zheng, Guolin Ke, Di He, Yanming Shen, and Tie-Yan Liu · 2021
Closest in time.
Student customized knowledge distillation: Bridging the gap between student and teacher
Yichen Zhu and Yi Wang · 2021
Closest in time.
Make a long image short: Adaptive token length for vision transformers, 2021
Yichen Zhu, Yuqin Zhu, Jie Du, Yi Wang, Zhicai Ou, Feifei Feng, and Jian Tang · 2021
Closest in time.
Logbert: Log anomaly detection via bert
Haixuan Guo, Shuhan Yuan, and Xintao Wu · 2021
Closest in time.
Neural architecture search without training
Joe Mellor, Jack Turner, Amos Storkey, and Elliot J Crowley · 2021
Closest in time.