Fetching the paper…
Reading the bibliography…
While self-supervised learning has made rapid advances in natural language processing, it remains unclear when researchers should engage in resource-intensive domain-specific pretraining (domain pretraining).
ERNIE: Enhanced Representation through Knowledge Integration
Yu Sun, Shuohuan Wang, Yukun Li, Shikun Feng, Xuyi Chen, Han Zhang, Xin Tian, Danxiang Zhu, Hao Tian, and Hua Wu. 2019 · 1904
Earlier work this paper cites.
Pre-Training with Whole Word Masking for Chinese BERT
Yiming Cui, Wanxiang Che, Ting Liu, Bing Qin, Ziqing Yang, Shijin Wang, and Guoping Hu. 2019 · 1906
Earlier work this paper cites.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Emad Elwany, Dave Moore, and Gaurav Oberoi. 2019 · 1911
Earlier work this paper cites.
Council Directive 93/13/EEC of 5 April 1993 on unfair terms in consumer contracts
European Union 1993 · 1993
Earlier work this paper cites.
Legal Language
P.M. Tiersma. 1999 · 1999
Earlier work this paper cites.
How judges overrule: Speech act theory and the doctrine of stare decisis
Pintip Hompluem Dunn. 2003 · 2003
Earlier work this paper cites.
The language of the law
David Mellinkoff. 2004 · 2004
Earlier work this paper cites.
The Cost of Training NLP Models: A Concise Overview
Or Sharir, Barak Peleg, and Yoav Shoham. 2020 · 2004
Earlier work this paper cites.
The Language of Law School: Learning to “Think Like a Lawyer”
Elizabeth Mertz. 2007 · 2007
Earlier work this paper cites.
Citators: Past, Present, and Future
Laura C. Dabney. 2008 · 2008
Earlier work this paper cites.
Aligning AI With Shared Human Values
Dan Hendrycks, Collin Burns, Steven Basart, Andrew Critch, Jerry Li, Dawn Song, and Jacob Steinhardt. 2021a · 2008
Earlier work this paper cites.
Measuring Massive Multitask Language Understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2021b · 2009
Earlier work this paper cites.
How Good is Your Tokenizer? On the Monolingual Performance of Multilingual Language Models
Phillip Rust, Jonas Pfeiffer, Ivan Vulić, Sebastian Ruder, and Iryna Gurevych. 2020 · 2012
Earlier work this paper cites.
Efficient Estimation of Word Representations in Vector Space
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
The Bluebook: A Uniform System of Citation (21st ed.)
Columbia Law Review Ass’n, Harvard Law Review Ass’n, and Yale Law Journal. 2015 · 2015
Earlier work this paper cites.
An overview of the BIOASQ large-scale biomedical semantic indexing and question answering competition
George Tsatsaronis, Georgios Balikas, Prodromos Malakasiotis, Ioannis Partalas, Matthias Zschunke, Michael R. Alvers, Dirk Weissenborn, Anastasia Krithara, Sergios Petridis, Dimitris Polychronopoulos, Yannis Almirantis, John Pavlopoulos, Nicolas Baskiotis, Patrick Gallinari, Thierry Artiéres, Axel-Cyrille Ngonga Ngomo, Norman Heino, Eric Gaussier, Liliana Barrio-Alvers, Michael Schroeder, Ion Androutsopoulos, and Georgios Paliouras. 2015 · 2015
Cited alongside, same era.
Predicting judicial decisions of the European Court of Human Rights: A natural language processing perspective
Nikolaos Aletras, Dimitrios Tsarapatsanis, Daniel Preoţiuc-Pietro, and Vasileios Lampos. 2016 · 2016
Cited alongside, same era.
SQuAD: 100,000+ Questions for Machine Comprehension of Text. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Austin, Texas, 2383–2392
Pranav Rajpurkar, Jian Zhang, Konstantin Lopyrev, and Percy Liang. 2016 · 2016
Cited alongside, same era.
Google’s Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Łukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Greg Corrado, Macduff Hughes, and Jeffrey Dean. 2016 · 2016
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Association for Computational Linguistics, Minneapolis, Minnesota, 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
BioBERT: a pre-trained biomedical language representation model for biomedical text mining
Jinhyuk Lee, Wonjin Yoon, Sungdong Kim, Donghyeon Kim, Sunkyu Kim, Chan Ho So, and Jaewoo Kang. 2019 · 2019
Later among the works it cites.
CLAUDETTE: an automated detector of potentially unfair clauses in online terms of service
Marco Lippi, Przemysław Pałka, Giuseppe Contissa, Francesca Lagioia, Hans-Wolfgang Micklitz, Giovanni Sartor, and Paolo Torroni. 2019 · 2019
Later among the works it cites.
Coqa: A conversational question answering challenge
Siva Reddy, Danqi Chen, and Christopher D Manning. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Harvesting and Utilizing Explanatory Parentheticals
Pablo D Arredondo. 2017 · 2017
Cited alongside, same era.
Extracting Contract Elements. In Proceedings of the 16th Edition of the International Conference on Articial Intelligence and Law (London, United Kingdom) (ICAIL ’17) . Association for Computing Machinery, New York, NY, USA, 19–28
Ilias Chalkidis, Ion Androutsopoulos, and Achilleas Michos. 2017 · 2017
Cited alongside, same era.
Exploring the Use of Text Classification in the Legal Domain
Octavia-Maria, Marcos Zampieri, Shervin Malmasi, Mihaela Vela, Liviu P. Dinu, and Josef van Genabith. 2017 · 2017
Cited alongside, same era.
Sentence boundary detection in adjudicatory decisions in the United States
Jaromir Savelka, Vern R Walker, Matthias Grabmair, and Kevin D Ashley. 2017 · 2017
Cited alongside, same era.
LexNLP: Natural language processing and information extraction for legal and regulatory texts
Michael J. Bommarito II, Daniel Martin Katz, and Eric M. Detterman. 2018 · 2018
Cited alongside, same era.
Measuring the Evolution of a Scientific Field through Citation Frames
David Jurgens, Srijan Kumar, Raine Hoover, Dan McFarland, and Dan Jurafsky. 2018 · 2018
Cited alongside, same era.
Taku Kudo and John Richardson. 2018 · 2018
Cited alongside, same era.
emrQA: A Large Corpus for Question Answering on Electronic Medical Records. In Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Brussels, Belgium, 2357–2368
Anusri Pampari, Preethi Raghavan, Jennifer Liang, and Jian Peng. 2018 · 2018
Cited alongside, same era.
AI Goes to Court: The Growing Landscape of AI for Access to Justice
Jonah Wu. 2019 · 2019
Later among the works it cites.
LEGAL-BERT: The Muppets straight out of Law School. In Findings of the Association for Computational Linguistics: EMNLP 2020 . Association for Computational Linguistics, Online, 2898–2904
Ilias Chalkidis, Manos Fergadiotis, Prodromos Malakasiotis, Nikolaos Aletras, and Ion Androutsopoulos. 2020 · 2020
Later among the works it cites.
Algorithmic accountability in the administrative state
David Freeman Engstrom and Daniel E Ho. 2020 · 2020
Later among the works it cites.
Government by Algorithm: Artificial Intelligence in Federal Administrative Agencies
David Freeman Engstrom, Daniel E. Ho, Catherine Sharkey, and Mariano-Florentino Cuéllar. 2020 · 2020
Later among the works it cites.
SpanBERT: Improving Pre-training by Representing and Predicting Spans
Mandar Joshi, Danqi Chen, Yinhan Liu, Daniel S. Weld, Luke Zettlemoyer, and Omer Levy. 2020 · 2020
Later among the works it cites.
Neural Mask Generator: Learning to Generate Adaptive Word Maskings for Language Model Adaptation. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) . Association for Computational Linguistics, Online, 6102–6120
Minki Kang, Moonsu Han, and Sung Ju Hwang. 2020 · 2020
Later among the works it cites.
How to build a more open justice system
Adam R. Pah, David L. Schwartz, Sarath Sanga, Zachary D. Clopton, Peter DiCola, Rachel Davis Mersey, Charlotte S. Alexander, Kristian J. Hammond, and Luís A. Nunes Amaral. 2020 · 2020
Later among the works it cites.
Improving Access to Justice with Legal Chatbots
Marc Queudot, Éric Charton, and Marie-Jean Meurs. 2020 · 2020
Later among the works it cites.
Koustuv Sinha, Prasanna Parthasarathi, Joelle Pineau, and Adina Williams. 2020 · 2020
Later among the works it cites.
How Does NLP Benefit Legal System: A Summary of Legal Artificial Intelligence. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Online, 5218–5230
Haoxi Zhong, Chaojun Xiao, Cunchao Tu, Tianyang Zhang, Zhiyuan Liu, and Maosong Sun. 2020a · 2020
Later among the works it cites.
On the Dangers of Stochastic Parrots: Can Language Models Be Too Big?
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Closest in time.
A primer in bertology: What we know about how bert works
Anna Rogers, Olga Kovaleva, and Anna Rumshisky. 2021 · 2021
Closest in time.