Fetching the paper…
Reading the bibliography…
Many data extraction tasks of practical relevance require not only syntactic pattern matching but also semantic reasoning about the content of the underlying text.
Language Models are Few-Shot Learners. In Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M.F. Balcan, and H. Lin (Eds.), Vol. 33. Curran Associates, Inc., 1877–1901
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 1901
Earlier work this paper cites.
Lambda-calculus models of programming languages
James Hiram Morris. 1968 · 1968
Earlier work this paper cites.
Complexity of automaton identification from given data
E Mark Gold. 1978 · 1978
Earlier work this paper cites.
Web Data Extraction Using Hybrid Program Synthesis: A Combination of Top-down and Bottom-up Inference. In Proceedings of the 2020 ACM SIGMOD International Conference on Management of Data (Portland, OR, USA) (SIGMOD ’20) . Association for Computing Machinery, New York, NY, USA, 1967–1978
Mohammad Raza and Sumit Gulwani. 2020 · 1978
Earlier work this paper cites.
Learning Regular Sets from Queries and Counterexamples
Dana Angluin. 1987 · 1987
Earlier work this paper cites.
Inference of Finite Automata Using Homing Sequences. In Proceedings of the Twenty-first Annual ACM Symposium on Theory of Computing (STOC ’89) . ACM, 411–420
R. L. Rivest and R. E. Schapire. 1989 · 1989
Earlier work this paper cites.
Incremental Grammatical Inference From Positive And Negative Data Using Unbiased Finite State Automata. In In Proceedings of the ACL’02 Workshop on Unsupervised Lexical Acquisition . 291–300
R. Alquezar and A. Sanfeliu. 1994 · 1994
Earlier work this paper cites.
An incremental interactive algorithm for regular grammar inference. In Grammatical Interference: Learning Syntax from Sentences , Laurent Miclet and Colin de la Higuera (Eds.). Springer Berlin Heidelberg, Berlin, Heidelberg, 238–249
Rajesh Parekh and Vasant Honavar. 1996 · 1996
Earlier work this paper cites.
Kleene Algebra with Tests
Dexter Kozen. 1997 · 1997
Earlier work this paper cites.
Learning Regular Languages from Positive Evidence. In Proceedings of the Twentieth Annual Conference of the Cognitive Science Society . 350–355
Laura Firoiu, Tim Oates, and Paul R. Cohen. 1998 · 1998
Earlier work this paper cites.
Learning DFA from Simple Examples
Rajesh Parekh and Vasant Honavar. 2001 · 2001
Earlier work this paper cites.
DBpedia - A crystallization point for the Web of Data
Christian Bizer, Jens Lehmann, Georgi Kobilarov, Sören Auer, Christian Becker, Richard Cyganiak, and Sebastian Hellmann. 2009 · 2009
Earlier work this paper cites.
Automating String Processing in Spreadsheets Using Input-Output Examples. In Proceedings of the 38th Annual ACM SIGPLAN-SIGACT Symposium on Principles of Programming Languages (Austin, Texas, USA) (POPL ’11) . Association for Computing Machinery, New York, NY, USA, 317–330
Sumit Gulwani. 2011 · 2011
Earlier work this paper cites.
FlashExtract: A Framework for Data Extraction by Examples. In Proceedings of the 35th ACM SIGPLAN Conference on Programming Language Design and Implementation (PLDI ’14) . ACM, 542–553
Vu Le and Sumit Gulwani. 2014 · 2014
Earlier work this paper cites.
Zero-shot Entity Extraction from Web Pages. In Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, Baltimore, Maryland, 391–401
Panupong Pasupat and Percy Liang. 2014 · 2014
Earlier work this paper cites.
Synthesizing Data Structure Transformations from Input-Output Examples. In Proceedings of the 36th ACM SIGPLAN Conference on Programming Language Design and Implementation (Portland, OR, USA) (PLDI ’15) . Association for Computing Machinery, New York, NY, USA, 229–239
John K. Feser, Swarat Chaudhuri, and Isil Dillig. 2015 · 2015
Earlier work this paper cites.
Type-and-Example-Directed Program Synthesis. In Proceedings of the 36th ACM SIGPLAN Conference on Programming Language Design and Implementation (Portland, OR, USA) (PLDI ’15) . Association for Computing Machinery, New York, NY, USA, 619–630
Peter-Michael Osera and Steve Zdancewic. 2015 · 2015
Earlier work this paper cites.
Compositional Semantic Parsing on Semi-Structured Tables. In Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Association for Computational Linguistics, Beijing, China, 1470–1480
Panupong Pasupat and Percy Liang. 2015 · 2015
Earlier work this paper cites.
FlashMeta: A Framework for Inductive Program Synthesis. In Proceedings of the 2015 ACM SIGPLAN International Conference on Object-Oriented Programming, Systems, Languages, and Applications (Pittsburgh, PA, USA) (OOPSLA 2015) . Association for Computing Machinery, New York, NY, USA, 107–126
Oleksandr Polozov and Sumit Gulwani. 2015 · 2015
Cited alongside, same era.
Compositional Program Synthesis from Natural Language and Examples. In Proceedings of the 24th International Conference on Artificial Intelligence (Buenos Aires, Argentina) (IJCAI’15) . AAAI Press, 792–800
Mohammad Raza, Sumit Gulwani, and Natasa Milic-Frayling. 2015 · 2015
Cited alongside, same era.
Learning to Compose Neural Networks for Question Answering. In Proceedings of the 2016 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, San Diego, California, 1545–1554
Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Dan Klein. 2016a · 2016
Cited alongside, same era.
Neural Module Networks. In 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . 39–48
Neuralizing Regular Expressions for Slot Filling. In Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Online and Punta Cana, Dominican Republic, 9481–9498
Chengyue Jiang, Zijian Jin, and Kewei Tu. 2021 · 2021
Later among the works it cites.
Multi-Modal Program Inference: A Marriage of Pre-Trained Language Models and Component-Based Synthesis
Kia Rahmani, Mohammad Raza, Sumit Gulwani, Vu Le, Daniel Morris, Arjun Radhakrishna, Gustavo Soares, and Ashish Tiwari. 2021 · 2021
Later among the works it cites.
Semantic Programming by Example with Pre-Trained Models
Gust Verbruggen, Vu Le, and Sumit Gulwani. 2021 · 2021
Later among the works it cites.
Optimal Neural Program Synthesis from Multimodal Specifications. In Findings of the Association for Computational Linguistics: EMNLP 2021 . Association for Computational Linguistics, Punta Cana, Dominican Republic, 1691–1704
Xi Ye, Qiaochu Chen, Isil Dillig, and Greg Durrett. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jacob Andreas, Marcus Rohrbach, Trevor Darrell, and Dan Klein. 2016b · 2016
Cited alongside, same era.
Example-Directed Synthesis: A Type-Theoretic Interpretation. In Proceedings of the 43rd Annual ACM SIGPLAN-SIGACT Symposium on Principles of Programming Languages (St. Petersburg, FL, USA) (POPL ’16) . Association for Computing Machinery, New York, NY, USA, 802–815
Jonathan Frankle, Peter-Michael Osera, David Walker, and Steve Zdancewic. 2016 · 2016
Cited alongside, same era.
Synthesizing Regular Expressions from Examples for Introductory Automata Assignments. In Proceedings of the 2016 ACM SIGPLAN International Conference on Generative Programming: Concepts and Experiences (Amsterdam, Netherlands) (GPCE 2016) . Association for Computing Machinery, New York, NY, USA, 70–80
Mina Lee, Sunbeom So, and Hakjoo Oh. 2016 · 2016
Cited alongside, same era.
Program Synthesis from Polymorphic Refinement Types. In Proceedings of the 37th ACM SIGPLAN Conference on Programming Language Design and Implementation (Santa Barbara, CA, USA) (PLDI ’16) . Association for Computing Machinery, New York, NY, USA, 522–538
Nadia Polikarpova, Ivan Kuraj, and Armando Solar-Lezama. 2016 · 2016
Cited alongside, same era.
Component-Based Synthesis of Table Consolidation and Transformation Tasks from Examples. In Proceedings of the 38th ACM SIGPLAN Conference on Programming Language Design and Implementation (Barcelona, Spain) (PLDI 2017) . Association for Computing Machinery, New York, NY, USA, 422–436
Yu Feng, Ruben Martins, Jacob Van Geffen, Isil Dillig, and Swarat Chaudhuri. 2017 · 2017
Cited alongside, same era.
Differentiable Programs with Neural Libraries. In Proceedings of the 34th International Conference on Machine Learning - Volume 70 (Sydney, NSW, Australia) (ICML’17) . JMLR.org, 1213–1222
Alexander L. Gaunt, Marc Brockschmidt, Nate Kushman, and Daniel Tarlow. 2017 · 2017
Cited alongside, same era.
Program Synthesis Using Conflict-Driven Learning. In Proceedings of the 39th ACM SIGPLAN Conference on Programming Language Design and Implementation (Philadelphia, PA, USA) (PLDI 2018) . Association for Computing Machinery, New York, NY, USA, 420–435
Yu Feng, Ruben Martins, Osbert Bastani, and Isil Dillig. 2018 · 2018
Cited alongside, same era.
HOUDINI: Lifelong Learning as Program Synthesis. In Advances in Neural Information Processing Systems , S. Bengio, H. Wallach, H. Larochelle, K. Grauman, N. Cesa-Bianchi, and R. Garnett (Eds.), Vol. 31. Curran Associates, Inc
Lazar Valkov, Dipak Chaudhari, Akash Srivastava, Charles Sutton, and Swarat Chaudhuri. 2018 · 2018
Cited alongside, same era.
Fonduer: Knowledge Base Construction from Richly Formatted Data. In Proceedings of the 2018 International Conference on Management of Data (Houston, TX, USA) (SIGMOD ’18) . Association for Computing Machinery, New York, NY, USA, 1301–1316
Sen Wu, Luke Hsiao, Xiao Cheng, Braden Hancock, Theodoros Rekatsinas, Philip Levis, and Christopher Ré. 2018 · 2018
Cited alongside, same era.
UDF to SQL Translation through Compositional Lazy Inductive Synthesis
Guoqiang Zhang, Yuanchao Xu, Xipeng Shen, and Işıl Dillig. 2021 · 2021
Later among the works it cites.
Compositional Safety LTL Synthesis. In Verified Software. Theories, Tools and Experiments.: 14th International Conference, VSTTE 2022, Trento, Italy, October 17–18, 2022, Revised Selected Papers (Trento, Italy). Springer-Verlag, Berlin, Heidelberg, 1–19
Suguman Bansal, Giuseppe De Giacomo, Antonio Di Stasio, Yong Li, Moshe Y. Vardi, and Shufang Zhu. 2023 · 2022
Later among the works it cites.
Interpretable, Verifiable, and Robust Reinforcement Learning via Program Synthesis
Osbert Bastani, Jeevana Priya Inala, and Armando Solar-Lezama. 2022 · 2022
Later among the works it cites.
PaLM: Scaling Language Modeling with Pathways
Aakanksha Chowdhery, Sharan Narang, Jacob Devlin, Maarten Bosma, Gaurav Mishra, Adam Roberts, Paul Barham, Hyung Won Chung, Charles Sutton, Sebastian Gehrmann, Parker Schuh, Kensen Shi, Sasha Tsvyashchenko, Joshua Maynez, Abhishek Rao, Parker Barnes, Yi Tay, Noam M. Shazeer, Vinodkumar Prabhakaran, Emily Reif, Nan Du, Benton C. Hutchinson, Reiner Pope, James Bradbury, Jacob Austin, Michael Isard, Guy Gur-Ari, Pengcheng Yin, Toju Duke, Anselm Levskaya, Sanjay Ghemawat, Sunipa Dev, Henryk Michalewski, Xavier García, Vedant Misra, Kevin Robinson, Liam Fedus, Denny Zhou, Daphne Ippolito, David Luan, Hyeontaek Lim, Barret Zoph, Alexander Spiridonov, Ryan Sepassi, David Dohan, Shivani Agrawal, Mark Omernick, Andrew M. Dai, Thanumalayan Sankaranarayana Pillai, Marie Pellat, Aitor Lewkowycz, Erica Moreira, Rewon Child, Oleksandr Polozov, Katherine Lee, Zongwei Zhou, Xuezhi Wang, Brennan Saeta, Mark Díaz, Orhan Firat, Michele Catasta, Jason Wei, Kathleen S. Meier-Hellstern, Douglas Eck, Jeff Dean, Slav Petrov, and Noah Fiedel. 2022 · 2022
Later among the works it cites.
Structured information extraction from complex scientific text with fine-tuned large language models
Alexander Dunn, John Dagdelen, Nicholas Walker, Sanghoon Lee, Andrew S. Rosen, Gerbrand Ceder, Kristin Persson, and Anubhav Jain. 2022 · 2022
Later among the works it cites.
Kleene Algebra modulo Theories: A Framework for Concrete KATs. In Proceedings of the 43rd ACM SIGPLAN International Conference on Programming Language Design and Implementation (San Diego, CA, USA) (PLDI 2022) . Association for Computing Machinery, New York, NY, USA, 594–608
Michael Greenberg, Ryan Beckett, and Eric Campbell. 2022 · 2022
Later among the works it cites.
Jigsaw: Large Language Models Meet Program Synthesis. In Proceedings of the 44th International Conference on Software Engineering (Pittsburgh, Pennsylvania) (ICSE ’22) . Association for Computing Machinery, New York, NY, USA, 1219–1231
Naman Jain, Skanda Vaidyanath, Arun Iyer, Nagarajan Natarajan, Suresh Parthasarathy, Sriram Rajamani, and Rahul Sharma. 2022 · 2022
Later among the works it cites.
Synchromesh: Reliable Code Generation from Pre-trained Language Models. In International Conference on Learning Representations
Gabriel Poesia, Alex Polozov, Vu Le, Ashish Tiwari, Gustavo Soares, Christopher Meek, and Sumit Gulwani. 2022 · 2022
Later among the works it cites.
Few-Shot Semantic Parsing with Language Models Trained on Code. In Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Association for Computational Linguistics, Seattle, United States, 5417–5425
Richard Shin and Benjamin Van Durme. 2022 · 2022
Later among the works it cites.
Binding Language Models in Symbolic Languages. In The Eleventh International Conference on Learning Representations
Zhoujun Cheng, Tianbao Xie, Peng Shi, Chengzu Li, Rahul Nadkarni, Yushi Hu, Caiming Xiong, Dragomir Radev, Mari Ostendorf, Luke Zettlemoyer, Noah A. Smith, and Tao Yu. 2023 · 2023
Closest in time.
CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis. In The Eleventh International Conference on Learning Representations
Erik Nijkamp, Bo Pang, Hiroaki Hayashi, Lifu Tu, Huan Wang, Yingbo Zhou, Silvio Savarese, and Caiming Xiong. 2023 · 2023
Closest in time.
Introducing ChatGPT
OpenAI. 2022 · 2023
Closest in time.
DocPrompting: Generating Code by Retrieving the Docs. In The Eleventh International Conference on Learning Representations
Shuyan Zhou, Uri Alon, Frank F. Xu, Zhengbao Jiang, and Graham Neubig. 2023 · 2023
Closest in time.
On Robustness of Prompt-based Semantic Parsing with Large Pre-trained Language Model: An Empirical Study on Codex
Terry Yue Zhuo, Zhuang Li, Yujin Huang, Fatemeh Shiri, Weiqing Wang, Gholamreza Haffari, and Yuan-Fang Li. 2023 · 2023
Closest in time.