Fetching the paper…
Reading the bibliography…
Binary malware summarization aims to automatically generate human-readable descriptions of malware behaviors from executable files, facilitating tasks like malware cracking and detection.
K. Papineni, S. Roukos, T. Ward, and W.-J. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in Proc. of ACL , 2002
2002
Earlier work this paper cites.
C.-Y. Lin, “ROUGE: A package for automatic evaluation of summaries,” in Proc. of ACL , 2004
2004
Earlier work this paper cites.
B. Cornelissen, A. Zaidman, A. Van Deursen, L. Moonen, and R. Koschke, “A systematic survey of program comprehension through dynamic analysis,” IEEE Transactions on Software Engineering , vol. 35, no. 5, pp. 684–702, 2009
2009
Earlier work this paper cites.
A. Lavie and M. Denkowski, “The meteor metric for automatic evaluation of machine translation,” Machine Translation , vol. 23, pp. 105–115, 2009
2009
Earlier work this paper cites.
G. Sridhara, E. Hill, D. Muppaneni, L. Pollock, and K. Vijay-Shanker, “Towards automatically generating summary comments for java methods,” in Proceedings of the 25th IEEE/ACM international conference on Automated software engineering , 2010, pp. 43–52
2010
Earlier work this paper cites.
A. Jain, S. Soner, and A. Gadwal, “Reverse engineering: Journey from code to design,” in Proc. of ICECT , 2011
2011
Earlier work this paper cites.
2013
Earlier work this paper cites.
P. W. McBurney and C. McMillan, “Automatic documentation generation via source code summarization of method context,” in Proceedings of the 22nd International Conference on Program Comprehension , 2014, pp. 279–290
2014
Earlier work this paper cites.
E. C. R. Shin, D. Song, and R. Moazzezi, “Recognizing functions in binaries with neural networks,” in Proc. of USENIX Security , 2015
2015
Earlier work this paper cites.
M. J. Kusner, Y. Sun, N. I. Kolkin, and K. Q. Weinberger, “From word embeddings to document distances,” in Proc. of ICML , 2015
2015
Earlier work this paper cites.
K. Yakdan, S. Dechand, E. Gerhards-Padilla, and M. Smith, “Helping johnny to analyze malware: A usability-optimized decompiler and malware analysis user study,” in Proc. of SP , 2016
2016
Earlier work this paper cites.
D. M. Berris, A. Veitch, N. Heintze, E. Anderson, and N. Wang, “Xray: A function call tracing system,” Technical report, 2016. A white paper on XRay, a function call tracing system developed at Google , 2016
2016
Earlier work this paper cites.
X. Deng and J. Mirkovic, “Malware analysis through high-level behavior,” in Proc. of CSET , 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
R. Labaca-Castro, B. Biggio, and G. Dreo Rodosek, “Poster: Attacking malware classifiers by crafting gradient-attacks that preserve functionality,” in Proc. of CCS , 2019
2019
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” in Proc. of NAACL-HLT , 2019
2019
Earlier work this paper cites.
W. Zhao, M. Peyrard, F. Liu, Y. Gao, C. M. Meyer, and S. Eger, “MoverScore: Text generation evaluating with contextualized embeddings and earth mover distance,” in Proc. of EMNLP-IJCNLP , 2019
2019
Earlier work this paper cites.
D. Gibert, C. Mateu, and J. Planes, “Hydra: A multimodal deep learning framework for malware classification,” Computers & Security , vol. 95, p. 101873, 2020
2020
Cited alongside, same era.
Z. Yu, R. Cao, Q. Tang, S. Nie, J. Huang, and S. Wu, “Order matters: Semantic-aware neural networks for binary code similarity detection,” in Proc. of AAAI , 2020
2020
Cited alongside, same era.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” Journal of Machine Learning Research , vol. 21, no. 140, pp. 1–67, 2020
2020
Cited alongside, same era.
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang, and M. Zhou, “CodeBERT: A pre-trained model for programming and natural languages,” in Proc. of EMNLP , 2020
2020
Cited alongside, same era.
A. Al-Kaswan, T. Ahmed, M. Izadi, A. A. Sawant, P. Devanbu, and A. van Deursen, “Extending source code pre-trained language models to summarise decompiled binarie,” in Proc. of SANER , 2023
2023
Later among the works it cites.
J. Xiong, G. Chen, K. Chen, H. Gao, S. Cheng, and W. Zhang, “Hext5: Unified pre-training for stripped binary code information inference,” in Proc. of ASE , 2023
2023
Later among the works it cites.
E. Nijkamp, B. Pang, H. Hayashi, L. Tu, H. Wang, Y. Zhou, S. Savarese, and C. Xiong, “Codegen: An open large language model for code with multi-turn program synthesis,” in Proc. of ICLR , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Sellam, D. Das, and A. Parikh, “BLEURT: Learning robust metrics for text generation,” in Proc. of ACL , 2020
2020
Cited alongside, same era.
M. O. F. Rokon, R. Islam, A. Darki, E. E. Papalexakis, and M. Faloutsos, “SourceFinder: Finding malware Source-Code from publicly available repositories in GitHub,” in Proc. of RAID , 2020
2020
Cited alongside, same era.
C. Beaman, A. Barkworth, T. D. Akande, S. Hakak, and M. K. Khan, “Ransomware: Recent advances, analysis, challenges and future research directions,” Computers & Security , vol. 111, p. 102490, 2021
2021
Cited alongside, same era.
Z. Liu, “Binary code similarity detection,” in Proc. of ASE , 2021
2021
Cited alongside, same era.
Y. Wang, W. Wang, S. Joty, and S. C. Hoi, “CodeT5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and generation,” in Proc. of EMNLP , 2021
2021
Cited alongside, same era.
L. Yang, A. Ciptadi, I. Laziuk, A. Ahmadzadeh, and G. Wang, “Bodmas: An open dataset for learning based temporal analysis of pe malware,” in Proc. of DLS , 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
F. A. Aboaoja, A. Zainal, F. A. Ghaleb, B. A. S. Al-rimy, T. A. E. Eisa, and A. A. H. Elnour, “Malware detection issues, challenges, and future directions: A survey,” Applied Sciences , vol. 12, no. 17, 2022
2022
Cited alongside, same era.
Y. Wang, H. Le, A. Gotmare, N. Bui, J. Li, and S. Hoi, “CodeT5+: Open code large language models for code understanding and generation,” in Proc. of EMNLP , 2023
2023
Later among the works it cites.
W. Zhu, Z. Feng, Z. Zhang, J. Chen, Z. Ou, M. Yang, and C. Zhang, “Callee: Recovering call graphs for binaries with transfer and contrastive learning,” in Proc. of SP , 2023
2023
Later among the works it cites.
——. The technology behind github’s new code search. [Online]. Available: https://github.blog/2023-02-06-the-technology-behind-githubs-new-code-search/
2023
Later among the works it cites.
T. Ye, L. Wu, T. Ma, X. Zhang, Y. Du, P. Liu, S. Ji, and W. Wang, “CP-BCS: Binary code summarization guided by control flow graph and pseudo code,” in Proc. of EMNLP , 2023
2023
Later among the works it cites.
Z. Ji, N. Lee, R. Frieske, T. Yu, D. Su, Y. Xu, E. Ishii, Y. J. Bang, A. Madotto, and P. Fung, “Survey of hallucination in natural language generation,” ACM Computing Surveys , vol. 55, no. 12, pp. 1–38, 2023
2023
Later among the works it cites.
“Malware statistics &trends report — av-test,” https://www.av-test.org/en/statistics/malware/
2024
Closest in time.
2024
Closest in time.
A. Mastropaolo, M. Ciniselli, M. Di Penta, and G. Bavota, “Evaluating code summarization techniques: A new metric and an empirical characterization,” in Proc. of ICSE , 2024
2024
Closest in time.
2024
Closest in time.
K. Pal, A. Bajaj, P. Banerjee, A. Dutcher, M. Nakamura, Z. Basque, H. Gupta, S. Sawant, U. Anantheswaran, Y. Shoshitaishvili, A. Doupe, C. Baral, and R. Wang, “;len or index or count, anything but v1": Predicting variable names in decompilation output with transfer learning,” in Proc. of SP , 2024
2024
Closest in time.
C. Xu, Q. Sun, K. Zheng, X. Geng, P. Zhao, J. Feng, C. Tao, Q. Lin, and D. Jiang, “WizardLM: Empowering large pre-trained language models to follow complex instructions,” in Proc. of ICLR , 2024
2024
Closest in time.
B. Rozière, J. Gehring, F. Gloeckle, S. Sootla, I. Gat, X. E. Tan, Y. Adi, J. Liu, R. Sauvestre, T. Remez, J. Rapin, A. Kozhevnikov, I. Evtimov, J. Bitton, M. Bhatt, C. C. Ferrer, A. Grattafiori, W. Xiong, A. Défossez, J. Copet, F. Azhar, H. Touvron, L. Martin, N. Usunier, T. Scialom, and G. Synnaeve, “Code llama: Open foundation models for code,” 2024
2024
Closest in time.