Fetching the paper…
Reading the bibliography…
A C decompiler converts an executable into source code.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al · 1901
Earlier work this paper cites.
Machine independence: its technology and economics
Mark I Halpern. 1965 · 1965
Earlier work this paper cites.
Reverse compilation techniques
Cristina Cifuentes. 1994 · 1994
Earlier work this paper cites.
Decompilation of Binary Programs
Cristina Cifuentes and K. John Gough. 1995 · 1995
Earlier work this paper cites.
An infrastructure for adaptive dynamic optimization. In International Symposium on Code Generation and Optimization, 2003. CGO 2003. IEEE, 265–275
Derek Bruening, Timothy Garnett, and Saman Amarasinghe. 2003 · 2003
Earlier work this paper cites.
Pin: building customized program analysis tools with dynamic instrumentation
Chi-Keung Luk, Robert Cohn, Robert Muth, Harish Patil, Artur Klauser, Geoff Lowney, Steven Wallace, Vijay Janapa Reddi, and Kim Hazelwood. 2005 · 2005
Earlier work this paper cites.
Valgrind: a framework for heavyweight dynamic binary instrumentation
Nicholas Nethercote and Julian Seward. 2007 · 2007
Earlier work this paper cites.
GenProg: A Generic Method for Automatic Software Repair
Claire Le Goues, Thanhvu Nguyen, Stephanie Forrest, and Westley Weimer. 2012 · 2012
Earlier work this paper cites.
A systematic study of automated program repair: Fixing 55 out of 105 bugs for $8 each. In ICSE
Claire Le Goues, Michael Dewey-Vogt, Stephanie Forrest, and Westley Weimer. 2012 · 2012
Earlier work this paper cites.
AddressSanitizer: A Fast Address Sanity Checker (USENIX ATC’12) . 28–28
Konstantin Serebryany, Derek Bruening, Alexander Potapenko, and Dmitry Vyukov. 2012 · 2012
Earlier work this paper cites.
Native x86 Decompilation Using Semantics-Preserving Structural Analysis and Iterative Control-Flow Structuring. In Presented as part of the 22nd USENIX Security Symposium (USENIX Security 13) . 353–368
David Brumley, JongHyup Lee, Edward J. Schwartz, and Maverick Woo. 2013 · 2013
Earlier work this paper cites.
Automatic patch generation learned from human-written patches. In 2013 35th International Conference on Software Engineering (ICSE) . IEEE, 802–811
Dongsun Kim, Jaechang Nam, Jaewoo Song, and Sunghun Kim. 2013 · 2013
Earlier work this paper cites.
ByteWeight: Learning to Recognize Functions in Binary Code (USENIX Security)
Tiffany Bao, Jonathan Burket, Maverick Woo, Rafael Turner, and David Brumley. 2014 · 2014
Earlier work this paper cites.
Tracelet-based Code Search in Executables. In PLDI
Yaniv David and Eran Yahav. 2014 · 2014
Earlier work this paper cites.
IDA Pro: a cross-platform multi-processor disassembler and debugger
SA Hex-Rays. 2014 · 2014
Earlier work this paper cites.
Staged program repair with condition synthesis. In Proceedings of the 2015 10th Joint Meeting on Foundations of Software Engineering . 166–178
Fan Long and Martin Rinard. 2015 · 2015
Earlier work this paper cites.
An analysis of patch plausibility and correctness for generate-and-validate patch generation systems
Zichao Qi, Fan Long, Sara Achour, and Martin C. Rinard. 2015 · 2015
Earlier work this paper cites.
Recognizing functions in binaries with neural networks. In 24th USENIX security symposium (USENIX Security 15) . 611–626
Eui Chul Richard Shin, Dawn Song, and Reza Moazzezi. 2015 · 2015
Earlier work this paper cites.
Reassembleable disassembling. In 24th USENIX Security Symposium (USENIX Security 15) . 627–642
Shuai Wang, Pei Wang, and Dinghao Wu. 2015 · 2015
Earlier work this paper cites.
An { \{ In-Depth } \} Analysis of Disassembly on { \{ Full-Scale } \} x86/x64 Binaries. In 25th USENIX security symposium (USENIX security 16) . 583–600
Dennis Andriesse, Xi Chen, Victor Van Der Veen, Asia Slowinska, and Herbert Bos. 2016 · 2016
Earlier work this paper cites.
Bingo: Cross-architecture cross-os binary search. In Proceedings of the 2016 24th ACM SIGSOFT International Symposium on Foundations of Software Engineering . 678–689
Mahinthan Chandramohan, Yinxing Xue, Zhengzi Xu, Yang Liu, Chia Yuan Cho, and Hee Beng Kuan Tan. 2016 · 2016
Earlier work this paper cites.
Automatic patch generation by learning correct code. In Proceedings of the 43rd Annual ACM SIGPLAN-SIGACT Symposium on Principles of Programming Languages . 298–312
Fan Long and Martin Rinard. 2016 · 2016
Earlier work this paper cites.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei. 2017 · 2017
Earlier work this paper cites.
Ramblr: Making Reassembly Great Again.. In NDSS
Ruoyu Wang, Yan Shoshitaishvili, Antonio Bianchi, Aravind Machiry, John Grosen, Paul Grosen, Christopher Kruegel, and Giovanni Vigna. 2017 · 2017
Earlier work this paper cites.
Superset Disassembly: Statically Rewriting x86 Binaries Without Heuristics.. In NDSS
Erick Bauman, Zhiqiang Lin, and Kevin W Hamlen. 2018 · 2018
Earlier work this paper cites.
FirmUp: Precise Static Detection of Common Vulnerabilities in Firmware. In ASPLOS
Yaniv David, Nimrod Partush, and Eran Yahav. 2018 · 2018
Earlier work this paper cites.
Debin: Predicting Debug Information in Stripped Binaries. In CCS ’18
Jingxuan He, Pesho Ivanov, Petar Tsankov, Veselin Raychev, and Martin Vechev. 2018 · 2018
Earlier work this paper cites.
Vuldeepecker: A deep learning-based system for vulnerability detection
Zhen Li, Deqing Zou, Shouhuai Xu, Xinyu Ou, Hai Jin, Sujuan Wang, Zhijun Deng, and Yuyi Zhong. 2018 · 2018
Earlier work this paper cites.
Software Protection on the Go: A Large-scale Empirical Study on Mobile App Obfuscation. In ICSE
Pei Wang, Qinkun Bao, Li Wang, Shuai Wang, Zhaofeng Chen, Tao Wei, and Dinghao Wu. 2018 · 2018
Earlier work this paper cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
Dire: A neural approach to decompiled identifier naming. In ASE
Jeremy Lacomis, Pengcheng Yin, Edward Schwartz, Miltiadis Allamanis, Claire Le Goues, Graham Neubig, and Bogdan Vasilescu. 2019 · 2019
Cited alongside, same era.
Probabilistic disassembly. In 2019 IEEE/ACM 41st International Conference on Software Engineering (ICSE) . IEEE, 1187–1198
Kenneth Miller, Yonghwi Kwon, Yi Sun, Zhuo Zhang, Xiangyu Zhang, and Zhiqiang Lin. 2019 · 2019
Cited alongside, same era.
The NSA makes Ghidra, a powerful cybersecurity tool, open source
Lily Hay Newman. 2019 · 2019
Cited alongside, same era.
Survey of machine learning techniques for malware analysis
Daniele Ucci, Leonardo Aniello, and Roberto Baldoni. 2019 · 2019
Cited alongside, same era.
Classifying malware represented as control flow graphs using deep graph convolutional neural network. In 2019 49th annual IEEE/IFIP international conference on dependable systems and networks (DSN) . IEEE, 52–63
Jiaqi Yan, Guanhua Yan, and Dong Jin. 2019 · 2019
Cited alongside, same era.
Explanations from large language models make small reasoners better
Shiyang Li, Jianshu Chen, Yelong Shen, Zhiyu Chen, Xinlu Zhang, Zekun Li, Hong Wang, Jing Qian, Baolin Peng, Yi Mao, et al · 2022
Later among the works it cites.
Competition-level code generation with alphacode
Yujia Li, David Choi, Junyoung Chung, Nate Kushman, Julian Schrittwieser, Rémi Leblond, Tom Eccles, James Keeling, Felix Gimeno, Agustin Dal Lago, et al · 2022
Later among the works it cites.
On the advance of making language models better reasoners
Yifei Li, Zeqi Lin, Shizhuo Zhang, Qiang Fu, Bei Chen, Jian-Guang Lou, and Weizhu Chen. 2022c · 2022
Later among the works it cites.
Sok: Demystifying binary lifters through the lens of downstream applications. In 2022 IEEE Symposium on Security and Privacy (SP) . IEEE, 1100–1119
Zhibo Liu, Yuanyuan Yuan, Shuai Wang, and Yuyan Bao. 2022 · 2022
Later among the works it cites.
The Convergence of Source Code and Binary Vulnerability Discovery–A Case Study (AsiaCCS)
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
BinRec: dynamic binary lifting and recompilation. In Proceedings of the Fifteenth European Conference on Computer Systems . 1–16
Anil Altinay, Joseph Nash, Taddeus Kroes, Prabhu Rajasekaran, Dixin Zhou, Adrian Dabrowski, David Gens, Yeoul Na, Stijn Volckaert, Cristiano Giuffrida, et al · 2020
Cited alongside, same era.
Cati: Context-assisted type inference from stripped binaries. In 2020 50th Annual IEEE/IFIP International Conference on Dependable Systems and Networks (DSN) . IEEE, 88–98
Ligeng Chen, Zhongling He, and Bing Mao. 2020 · 2020
Cited alongside, same era.
Scalable Validation for Binary Lifters
Sandeep Dasgupta, Sushant Dinesh, Deepan Venkatesh, Vikram S Adve, and Christopher W Fletcher. 2020 · 2020
Cited alongside, same era.
Neural reverse engineering of stripped binaries using augmented control flow graphs
Yaniv David, Uri Alon, and Eran Yahav. 2020 · 2020
Cited alongside, same era.
Retrowrite: Statically instrumenting cots binaries for fuzzing and sanitization. In 2020 IEEE Symposium on Security and Privacy (SP) . IEEE, 1497–1511
Sushant Dinesh, Nathan Burow, Dongyan Xu, and Mathias Payer. 2020 · 2020
Cited alongside, same era.
Binary rewriting without control flow recovery. In Proceedings of the 41st ACM SIGPLAN conference on programming language design and implementation . 151–163
Gregory J Duck, Xiang Gao, and Abhik Roychoudhury. 2020 · 2020
Cited alongside, same era.
Shortcut learning in deep neural networks
Robert Geirhos, Jörn-Henrik Jacobsen, Claudio Michaelis, Richard Zemel, Wieland Brendel, Matthias Bethge, and Felix A Wichmann. 2020 · 2020
Cited alongside, same era.
Alessandro Mantovani, Luca Compagna, Yan Shoshitaishvili, and Davide Balzarotti. 2022 · 2022
Later among the works it cites.
Codegen: An open large language model for code with multi-turn program synthesis
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2022 · 2022
Later among the works it cites.
Ground truth for binary disassembly is not easy. In 31st USENIX Security Symposium (USENIX Security 22) . 2479–2495
Chengbin Pang, Tiantai Zhang, Ruotong Yu, Bing Mao, and Jun Xu. 2022 · 2022
Later among the works it cites.
Pemma Reiter, Hui Jun Tay, Westley Weimer, Adam Doup’e, Ruoyu Wang, and Stephanie Forrest. 2022 · 2022
Later among the works it cites.
Rationale-augmented ensembles in language models
Xuezhi Wang, Jason Wei, Dale Schuurmans, Quoc Le, Ed Chi, and Denny Zhou. 2022 · 2022
Later among the works it cites.
Emergent abilities of large language models
Jason Wei, Yi Tay, Rishi Bommasani, Colin Raffel, Barret Zoph, Sebastian Borgeaud, Dani Yogatama, Maarten Bosma, Denny Zhou, Donald Metzler, et al · 2022
Later among the works it cites.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Chi, Quoc Le, and Denny Zhou. 2022b · 2022
Later among the works it cites.
How Long Can Open-Source LLMs Truly Promise on Context Length?
[n.d.] · 2023
Closest in time.
IDA-Pro Lumina
2023 · 2023
Closest in time.
Large Language Models for Compiler Optimization
Chris Cummins, Volker Seeker, Dejan Grubisic, Mostafa Elhoushi, Youwei Liang, Baptiste Roziere, Jonas Gehring, Fabian Gloeckle, Kim Hazelwood, Gabriel Synnaeve, et al · 2023
Closest in time.
A too-good-to-be-true prior to reduce shortcut reliance
Nikolay Dagaev, Brett D Roads, Xiaoliang Luo, Daniel N Barry, Kaustubh R Patil, and Bradley C Love. 2023 · 2023
Closest in time.
Automated repair of programs from large language models. In ICSE
Zhiyu Fan, Xiang Gao, Martin Mirchev, Abhik Roychoudhury, and Shin Hwei Tan. 2023 · 2023
Closest in time.
Survey of Hallucination in Natural Language Generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Ye Jin Bang, Andrea Madotto, and Pascale Fung. 2023 · 2023
Closest in time.
Impact of code language models on automated program repair. In International conference on software engineering (ICSE)
Nan Jiang, Kevin Liu, Thibaud Lutellier, and Lin Tan. 2023 · 2023
Closest in time.
Challenges and applications of large language models
Jean Kaddour, Joshua Harris, Maximilian Mozes, Herbie Bradley, Roberta Raileanu, and Robert McHardy. 2023 · 2023
Closest in time.
CODAMOSA: Escaping Coverage Plateaus in Test Generation with Pre-trained Large Language Models. In ICSE
Caroline Lemieux, Jeevana Priya Inala, Shuvendu K Lahiri, and Siddhartha Sen. 2023 · 2023
Closest in time.
StarCoder: may the source be with you!
Raymond Li, Loubna Ben Allal, Yangtian Zi, Niklas Muennighoff, Denis Kocetkov, Chenghao Mou, Marc Marone, Christopher Akiki, Jia Li, Jenny Chim, et al · 2023
Closest in time.
Lost in the middle: How language models use long contexts
Nelson F Liu, Kevin Lin, John Hewitt, Ashwin Paranjape, Michele Bevilacqua, Fabio Petroni, and Percy Liang. 2023a · 2023
Closest in time.
Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing
Pengfei Liu, Weizhe Yuan, Jinlan Fu, Zhengbao Jiang, Hiroaki Hayashi, and Graham Neubig. 2023b · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Direct preference optimization: Your language model is secretly a reward model
Rafael Rafailov, Archit Sharma, Eric Mitchell, Stefano Ermon, Christopher D Manning, and Chelsea Finn. 2023 · 2023
Closest in time.
Code llama: Open foundation models for code
Baptiste Rozière, Jonas Gehring, Fabian Gloeckle, Sten Sootla, Itai Gat, Xiaoqing Ellen Tan, Yossi Adi, Jingyu Liu, Tal Remez, Jérémy Rapin, et al · 2023
Closest in time.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, et al · 2023
Closest in time.
Conversational Automated Program Repair
Chun Xia and Lingming Zhang. 2023 · 2023
Closest in time.
Automated program repair in the era of large pre-trained language models. In Proceedings of the 45th International Conference on Software Engineering (ICSE 2023). Association for Computing Machinery
Chunqiu Steven Xia, Yuxiang Wei, and Lingming Zhang. 2023 · 2023
Closest in time.
LmPa: Improving Decompilation by Synergy of Large Language Model and Program Analysis
Xiangzhe Xu, Zhuo Zhang, Shiwei Feng, Yapeng Ye, Zian Su, Nan Jiang, Siyuan Cheng, Lin Tan, and Xiangyu Zhang. 2023 · 2023
Closest in time.