Fetching the paper…
Reading the bibliography…
Formally verifying software properties is a highly desirable but labor-intensive task.
Fast Transformer Decoding: One Write-Head is All You Need
Noam Shazeer. 2019 · 1911
Earlier work this paper cites.
The Current State and Future of Search Based Software Engineering. In ACM/IEEE International Conference on Software Engineering (ICSE) . 342–357
Mark Harman. 2007 · 2007
Earlier work this paper cites.
SeL4: Formal Verification of an OS Kernel. In Proceedings of the ACM SIGOPS 22nd Symposium on Operating Systems Principles (Big Sky, Montana, USA) (SOSP ’09) . Association for Computing Machinery, New York, NY, USA, 207–220
Gerwin Klein, Kevin Elphinstone, Gernot Heiser, June Andronick, David Cock, Philip Derrin, Dhammika Elkaduwe, Kai Engelhardt, Rafal Kolanski, Michael Norrish, Thomas Sewell, Harvey Tuch, and Simon Winwood. 2009 · 2009
Earlier work this paper cites.
Formal Verification of a Realistic Compiler
Xavier Leroy. 2009 · 2009
Earlier work this paper cites.
Generative Language Modeling for Automated Theorem Proving
Stanislas Polu and Ilya Sutskever. 2020 · 2009
Earlier work this paper cites.
GenProg: A Generic Method for Automatic Software Repair
Claire Le Goues, ThanhVu Nguyen, Stephanie Forrest, and Westley Weimer. 2012 · 2011
Earlier work this paper cites.
Finding and understanding bugs in C compilers. In ACM SIGPLAN Conference on Programming Language Design and Implementation (PLDI) . San Jose, CA, USA, 283–294
Xuejun Yang, Yang Chen, Eric Eide, and John Regehr. 2011 · 2011
Earlier work this paper cites.
Automatic patch generation learned from human-written patches. In ACM/IEEE International Conference on Software Engineering (ICSE) . San Francisco, CA, USA, 802–811
Dongsun Kim, Jaechang Nam, Jaewoo Song, and Sunghun Kim. 2013 · 2013
Earlier work this paper cites.
Leveraging Program Equivalence for Adaptive Program Repair: Models and First Results. In IEEE/ACM International Conference on Automated Software Engineering (ASE) . Palo Alto, CA, USA, 356–366
Westley Weimer, Zachary P. Fry, and Stephanie Forrest. 2013 · 2013
Earlier work this paper cites.
ACL2(ml): Machine-learning for ACL2
Jónathan Heras and Ekaterina Komendantskaya. 2014 · 2014
Earlier work this paper cites.
Repairing Programs with Semantic Code Search. In ASE (9–13). 295–306
Yalin Ke, Kathryn T. Stolee, Claire Le Goues, and Yuriy Brun. 2015 · 2015
Earlier work this paper cites.
An Analysis of Patch Plausibility and Correctness for Generate-and-validate Patch Generation Systems. In ISSTA . 24–36
Zichao Qi, Fan Long, Sara Achour, and Martin Rinard. 2015 · 2015
Earlier work this paper cites.
Is the Cure Worse than the Disease? Overfitting in Automated Program Repair. In ESEC/FSE . 532–543
Edward K. Smith, Earl Barr, Claire Le Goues, and Yuriy Brun. 2015 · 2015
Earlier work this paper cites.
Qlose: Program Repair with Quantitative Objectives. In International Conference on Computer Aided Verification (CAV) . Toronto, ON, Canada, 383–401
Loris D’Antoni, Roopsha Samanta, and Rishabh Singh. 2016 · 2016
Earlier work this paper cites.
History Driven Program Repair. In Intl. Conf. on Software Analysis, Evolution, and Reengineering , Vol. 1. 213–224
Xuan Bach D. Le, David Lo, and Claire Le Goues. 2016 · 2016
Earlier work this paper cites.
Automatic Patch Generation by Learning Correct Code. In ACM SIGPLAN-SIGACT Symposium on Principles of Programming Languages (POPL) . St. Petersburg, FL, USA, 298–312
Fan Long and Martin Rinard. 2016 · 2016
Earlier work this paper cites.
WaveNet: A Generative Model for Raw Audio
Aäron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew W. Senior, and Koray Kavukcuoglu. 2016 · 2016
Earlier work this paper cites.
Fault Localization for Automated Program Repair: Effectiveness, Performance, Repair Correctness
Fatmah Yousef Assiri and James M Bieman. 2017 · 2017
Earlier work this paper cites.
ICoq: Regression proof selection for large-scale verification projects. In IEEE/ACM International Conference on Automated Software Engineering (ASE) . Urbana-Champaign, IL, USA, 171–182
Ahmet Celik, Karl Palmskog, and Milos Gligoric. 2017 · 2017
Earlier work this paper cites.
Contract-based program repair without the contracts. In IEEE/ACM International Conference on Automated Software Engineering (ASE) . Urbana, IL, USA, 637–647
Liushan Chen, Yu Pei, and Carlo A. Furia. 2017 · 2017
Earlier work this paper cites.
DeepFix: Fixing Common C Language Errors by Deep Learning. In AAAI
Rahul Gupta, Soham Pal, Aditya Kanade, and Shirish K. Shevade. 2017 · 2017
Earlier work this paper cites.
Generating Good Generators for Inductive Relations
Leonidas Lampropoulos, Zoe Paraskevopoulou, and Benjamin C. Pierce. 2017 · 2017
Earlier work this paper cites.
ELIXIR: Effective object oriented program repair. In ASE . 648–659
Ripon K. Saha, Yingjun Lyu, Hiroaki Yoshida, and Mukul R. Prasad. 2017 · 2017
Earlier work this paper cites.
Coq, v.8.7
The Coq Development Team. 2017 · 2017
Earlier work this paper cites.
Automatically diagnosing and repairing error handling bugs in C. In European Software Engineering Conference and ACM SIGSOFT International Symposium on Foundations of Software Engineering (ESEC/FSE) . Paderborn, Germany, 752–762
Yuchi Tian and Baishakhi Ray. 2017 · 2017
Earlier work this paper cites.
Attention is all you need. In NeurIPS
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Identifying Test-suite-overfitted Patches through Test Case Generation. In ISSTA . 226–236
Qi Xin and Steven P. Reiss. 2017 · 2017
Earlier work this paper cites.
Better test cases for better automated program repair. In ESEC/FSE . 831–841
Jinqiu Yang, Alexey Zhikhartsev, Yuefei Liu, and Lin Tan. 2017 · 2017
Earlier work this paper cites.
A Regression Proof Selection Tool for Coq. In International Conference on Software Engineering Demonstrations Track (ICSE DEMO) . Gothenburg, Sweden, 117–120
Ahmet Celik, Karl Palmskog, and Milos Gligoric. 2018 · 2018
Earlier work this paper cites.
Hammer for Coq: Automation for Dependent Type Theory
Łukasz Czajka and Cezary Kaliszyk. 2018 · 2018
Earlier work this paper cites.
Hierarchical Neural Story Generation. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Melbourne, Australia, 889–898
Angela Fan, Mike Lewis, and Yann Dauphin. 2018 · 2018
Earlier work this paper cites.
Automated Clustering and Program Repair for Introductory Programming Assignments. In PLDI . 465–480
Sumit Gulwani, Ivan Radiček, and Florian Zuleger. 2018 · 2018
Earlier work this paper cites.
Shaping Program Repair Space with Existing Patches and Similar Code. In ISSTA
Jiajun Jiang, Yingfei Xiong, Hongyu Zhang, Qing Gao, and Xiangqun Chen. 2018 · 2018
Earlier work this paper cites.
Semantic Program Repair Using a Reference Implementation. In ICSE . 129–139
Sergey Mechtaev, Manh-Dung Nguyen, Yannic Noller, Lars Grunske, and Abhik Roychoudhury. 2018 · 2018
Cited alongside, same era.
PaMpeR: Proof Method Recommendation System for Isabelle/HOL. In International Conference on Automated Software Engineering (ASE) . Montpellier, France, 362–372
Yutaka Nagashima and Yilun He. 2018 · 2018
Cited alongside, same era.
PiCoq: Parallel Regression Proving for Large-Scale Verification Projects. In ACM SIGSOFT International Symposium on Software Testing and Analysis (ISSTA) . Amsterdam, Netherlands, 344–355
Karl Palmskog, Ahmet Celik, and Milos Gligoric. 2018 · 2018
Cited alongside, same era.
Refining Fitness Functions in Test-Based Program Repair. In APR . 13–14
Justyna Petke and Aymeric Blot. 2018 · 2018
Cited alongside, same era.
Search-Based Efficient Automated Program Repair Using Mutation and Fault Localization. In COMPSAC , Vol. 1. 174–183
Shuyao Sun, Junxia Guo, Ruilian Zhao, and Zheng Li. 2018 · 2018
Automated Patch Correctness Assessment: How Far Are We?. In ASE . Association for Computing Machinery, 968–980
Shangwen Wang, Ming Wen, Bo Lin, Hongjun Wu, Yihao Qin, Deqing Zou, Xiaoguang Mao, and Hai Jin. 2020 · 2020
Later among the works it cites.
Evaluating Large Language Models Trained on Code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harrison Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Joshua Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, and Wojciech Zaremba. 2021 · 2021
Later among the works it cites.
Training Verifiers to Solve Math Word Problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Jacob Hilton, Reiichiro Nakano, Christopher Hesse, and John Schulman. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Search, align, and repair: Data-driven feedback generation for introductory programming exercises. In PLDI . 481–495
Ke Wang, Rishabh Singh, and Zhendong Su. 2018 · 2018
Cited alongside, same era.
Context-Aware Patch Generation for Better Automated Program Repair. In ACM/IEEE International Conference on Software Engineering (ICSE) . Gothenburg, Sweden, 1–11
Ming Wen, Junjie Chen, Rongxin Wu, Dan Hao, and Shing-Chi Cheung. 2018 · 2018
Cited alongside, same era.
Evaluating the Strategies of Statement Selection in Automated Program Repair. In SATE . Springer
Deheng Yang, Yuhua Qi, and Xiaoguang Mao. 2018 · 2018
Cited alongside, same era.
SOSRepair: Expressive Semantic Search for Real-World Program Repair
Afsoon Afzal, Manish Motwani, Kathryn T. Stolee, Yuriy Brun, and Claire Le Goues. 2021 · 2019
Cited alongside, same era.
HOList: An Environment for Machine Learning of Higher Order Logic Theorem Proving. In Proceedings of the 36th International Conference on Machine Learning, ICML 2019, 9-15 June 2019, Long Beach, California, USA (Proceedings of Machine Learning Research, Vol. 97) , Kamalika Chaudhuri and Ruslan Salakhutdinov (Eds.). PMLR, 454–463
Kshitij Bansal, Sarah M. Loos, Markus N. Rabe, Christian Szegedy, and Stewart Wilcox. 2019 · 2019
Cited alongside, same era.
Mutation Analysis for Coq. In IEEE/ACM International Conference on Automated Software Engineering (ASE) . San Diego, California, 539–551
Ahmet Celik, Karl Palmskog, Marinela Parovic, Emilio Jesús Gallego Arias, and Milos Gligoric. 2019 · 2019
Cited alongside, same era.
Sequencer: Sequence-to-sequence learning for end-to-end program repair
Zimin Chen, Steve James Kommrusch, Michele Tufano, Louis-Noël Pouchet, Denys Poshyvanyk, and Martin Monperrus. 2019 · 2019
Cited alongside, same era.
TacticToe: Learning to Prove with Tactics
Thibault Gauthier, Cezary Kaliszyk, Josef Urban, Ramana Kumar, and Michael Norrish. 2021 · 2021
Later among the works it cites.
Measuring Mathematical Problem Solving With the MATH Dataset
Dan Hendrycks, Collin Burns, Saurav Kadavath, Akul Arora, Steven Basart, Eric Tang, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Later among the works it cites.
LISA: Language models of ISAbelle proofs. In Conference on Artificial Intelligence and Theorem Proving (AITP . Aussois, France, 17.1–17.3
Albert Qiaochu Jiang, Wenda Li, Jesse Michael Han, and Yuhuai Wu. 2021 · 2021
Later among the works it cites.
Roosterize: Suggesting Lemma Names for Coq Verification Projects Using Deep Learning. In International Conference on Software Engineering Demonstrations Track (ICSE DEMO) . Madrid, Spain, 21–24
Pengyu Nie, Karl Palmskog, Junyi Jessy Li, and Milos Gligoric. 2021 · 2021
Later among the works it cites.
Pretrained Language Models are Symbolic Mathematics Solvers too!
Kimia Noorbakhsh, Modar Sulaiman, Mahdi Sharifi, Kallol Roy, and Pooyan Jamshidi. 2021 · 2021
Later among the works it cites.
Show Your Work: Scratchpads for Intermediate Computation with Language Models
Maxwell I. Nye, Anders Johan Andreassen, Guy Gur-Ari, Henryk Michalewski, Jacob Austin, David Bieber, David Dohan, Aitor Lewkowycz, Maarten Bosma, David Luan, Charles Sutton, and Augustus Odena. 2021 · 2021
Later among the works it cites.
Proof Repair
Talia Ringer. 2021 · 2021
Later among the works it cites.
RoFormer: Enhanced Transformer with Rotary Position Embedding
Jianlin Su, Yu Lu, Shengfeng Pan, Bo Wen, and Yunfeng Liu. 2021 · 2021
Later among the works it cites.
Minchao Wu, Michael Norrish, Christian Walder, and Amir Dezfouli. 2021 · 2021
Later among the works it cites.
Break-it-fix-it: Unsupervised learning for program repair. In International Conference on Machine Learning (ICML) . PMLR, 11941–11952
Michihiro Yasunaga and Percy Liang. 2021 · 2021
Later among the works it cites.
Automated patch assessment for program repair at scale
He Ye, Matias Martinez, and Martin Monperrus. 2021 · 2021
Later among the works it cites.
A syntax-guided edit decoder for neural program repair. In ESEC/FSE . 341–353
Qihao Zhu, Zeyu Sun, Yuan an Xiao, Wenjie Zhang, Kang Yuan, Yingfei Xiong, and Lu Zhang. 2021 · 2021
Later among the works it cites.
ProofNet: A benchmark for autoformalizing and formally proving undergraduate-level mathematics problems. In Workshop MATH-AI: Toward Human-Level Mathematical Reasoning . New Orleans, Louisiana, USA
Zhangir Azerbayev, Bartosz Piotrowski, and Jeremy Avigad. 2022 · 2022
Later among the works it cites.
GPT-NeoX-20B: An Open-Source Autoregressive Language Model. In Proceedings of BigScience Episode #5 – Workshop on Challenges & Perspectives in Creating Large Language Models . Association for Computational Linguistics, virtual+Dublin, 95–136
Sidney Black, Stella Biderman, Eric Hallahan, Quentin Anthony, Leo Gao, Laurence Golding, Horace He, Connor Leahy, Kyle McDonell, Jason Phang, Michael Pieler, Usvsn Sai Prashanth, Shivanshu Purohit, Laria Reynolds, Jonathan Tow, Ben Wang, and Samuel Weinbach. 2022 · 2022
Later among the works it cites.
PaLM: Scaling Language Modeling with Pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, Parker Schuh, Kensen Shi, Sasha Tsvyashchenko, Joshua Maynez, Abhishek Rao, Parker Barnes, Yi Tay, Noam Shazeer, Vinodkumar Prabhakaran, Emily Reif, Nan Du, Ben Hutchinson, Reiner Pope, James Bradbury, Jacob Austin, Michael Isard, Guy Gur-Ari, Pengcheng Yin, Toju Duke, Anselm Levskaya, Sanjay Ghemawat, Sunipa Dev, Henryk Michalewski, Xavier Garcia, Vedant Misra, Kevin Robinson, Liam Fedus, Denny Zhou, Daphne Ippolito, David Luan, Hyeontaek Lim, Barret Zoph, Alexander Spiridonov, Ryan Sepassi, David Dohan, Shivani Agrawal, Mark Omernick, Andrew M. Dai, Thanumalayan Sankaranarayana Pillai, Marie Pellat, Aitor Lewkowycz, Erica Moreira, Rewon Child, Oleksandr Polozov, Katherine Lee, Zongwei Zhou, Xuezhi Wang, Brennan Saeta, Mark Diaz, Orhan Firat, Michele Catasta, Jason Wei, Kathy Meier-Hellstern, Douglas Eck, Jeff Dean, Slav Petrov, and Noah Fiedel. 2022 · 2022
Later among the works it cites.
Formal Specifications from Natural Language
Christopher Hahn, Frederik Schmitt, Julia J. Tillman, Niklas Metzger, Julian Siber, and Bernd Finkbeiner. 2022 · 2022
Later among the works it cites.
Proof Artifact Co-Training for Theorem Proving with Language Models. In The Tenth International Conference on Learning Representations, ICLR 2022, Virtual Event, April 25-29, 2022 . OpenReview.net
Jesse Michael Han, Jason Rute, Yuhuai Wu, Edward W. Ayers, and Stanislas Polu. 2022 · 2022
Later among the works it cites.
Draft, Sketch, and Prove: Guiding Formal Theorem Provers with Informal Proofs
Albert Q. Jiang, Sean Welleck, Jin Peng Zhou, Wenda Li, Jiacheng Liu, Mateja Jamnik, Timothée Lacroix, Yuhuai Wu, and Guillaume Lample. 2022b · 2022
Later among the works it cites.
HyperTree Proof Search for Neural Theorem Proving
Guillaume Lample, Marie-Anne Lachaux, Thibaut Lavril, Xavier Martinet, Amaury Hayat, Gabriel Ebner, Aurélien Rodriguez, and Timothée Lacroix. 2022 · 2022
Later among the works it cites.
Solving Quantitative Reasoning Problems with Language Models
Aitor Lewkowycz, Anders Andreassen, David Dohan, Ethan Dyer, Henryk Michalewski, Vinay V. Ramasesh, Ambrose Slone, Cem Anil, Imanol Schlag, Theo Gutman-Solo, Yuhuai Wu, Behnam Neyshabur, Guy Gur-Ari, and Vedant Misra. 2022 · 2022
Later among the works it cites.
Proof Mate: An Interactive Proof Helper for PVS (Tool Paper). In NASA Formal Methods Symposium . Springer, 809–815
Paolo Masci and Aaron Dutle. 2022 · 2022
Later among the works it cites.
Efficiently Scaling Transformer Inference
Reiner Pope, Sholto Douglas, Aakanksha Chowdhery, Jacob Devlin, James Bradbury, Anselm Levskaya, Jonathan Heek, Kefan Xiao, Shivani Agrawal, and Jeff Dean. 2022 · 2022
Later among the works it cites.
Chain of Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed H. Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Later among the works it cites.
Autoformalization with Large Language Models
Yuhuai Wu, Albert Q. Jiang, Wenda Li, Markus N. Rabe, Charles Staats, Mateja Jamnik, and Christian Szegedy. 2022 · 2022
Later among the works it cites.
miniF2F: A cross-system benchmark for formal Olympiad-level mathematics. In ICLR
Kunhao Zheng, Jesse Michael Han, and Stanislas Polu. 2022 · 2022
Later among the works it cites.
Towards Autoformalization of Mathematics and Code Correctness: Experiments with Elementary Proofs
Garett Cunningham, Razvan C. Bunescu, and David Juedes. 2023 · 2023
Closest in time.
Better Automatic Program Repair by Using Bug Reports and Tests Together. In International Conference on Software Engineering (ICSE) (14–20). Melbourne, Australia
Manish Motwani and Yuriy Brun. 2023 · 2023
Closest in time.
The Sledgehammer: Let Automatic Theorem Provers write your Isabelle scripts!
Larry Paulson and Tobias Nipkow. 2023 · 2023
Closest in time.
Passport: Improving Automated Formal Verification Using Identifiers
Alex Sanchez-Stern, Emily First, Timothy Zhou, Zhanna Kaufman, Yuriy Brun, and Talia Ringer. 2023 · 2023
Closest in time.