Fetching the paper…
Reading the bibliography…
While research on scientific claim verification has led to the development of powerful systems that appear to approach human performance, these approaches have yet to be tested in a realistic setting against large corpora of scientific literature.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Report on the need for and provision of an "ideal" information retrieval test collection
Karen Sparck Jones and C. J. van Rijsbergen. 1975 · 1975
Earlier work this paper cites.
Efficient construction of large test collections
Gordon V. Cormack, Christopher R. Palmer, and Charles L. A. Clarke. 1998 · 1998
Earlier work this paper cites.
Longformer: The Long-Document Transformer
Iz Beltagy, Matthew E. Peters, and Arman Cohan. 2020 · 2004
Earlier work this paper cites.
TREC: Experiment and Evaluation in Information Retrieval
Ellen M. Voorhees and Donna Harman. 2005 · 2005
Earlier work this paper cites.
The Probabilistic Relevance Framework: BM25 and Beyond
S. Robertson and H. Zaragoza. 2009 · 2009
Earlier work this paper cites.
An Empirical Investigation of Statistical Significance in NLP
Taylor Berg-Kirkpatrick, David Burkett, and Dan Klein. 2012 · 2012
Earlier work this paper cites.
MS MARCO: A Human Generated MAchine Reading COmprehension Dataset
Daniel Fernando Campos, T. Nguyen, M. Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, L. Deng, and Bhaskar Mitra. 2016 · 2016
Earlier work this paper cites.
Overview of the TREC 2016 clinical decision support track
Kirk Roberts, Dina Demner-Fushman, Ellen M. Voorhees, and William R. Hersh. 2016 · 2016
Earlier work this paper cites.
The Hitchhiker’s Guide to Testing Statistical Significance in Natural Language Processing
Rotem Dror, Gili Baumer, Segev Shlomov, and Roi Reichart. 2018 · 2018
Earlier work this paper cites.
FEVER: a large-scale dataset for Fact Extraction and VERification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
CLIMATE-FEVER: A Dataset for Verification of Real-World Climate Claims
T. Diggelmann, Jordan L. Boyd-Graber, Jannis Bulian, Massimiliano Ciaramita, and Markus Leippold. 2020 · 2020
Earlier work this paper cites.
Dense Passage Retrieval for Open-Domain Question Answering
V. Karpukhin, Barlas Oğuz, Sewon Min, Patrick Lewis, Ledell Yu Wu, Sergey Edunov, Danqi Chen, and Wen tau Yih. 2020 · 2020
Cited alongside, same era.
Explainable Automated Fact-Checking for Public Health Claims
Neema Kotonya and F. Toni. 2020 · 2020
Cited alongside, same era.
S2ORC: The Semantic Scholar Open Research Corpus
Kyle Lo, Lucy Lu Wang, Mark Neumann, Rodney Michael Kinney, and Daniel S. Weld. 2020 · 2020
Cited alongside, same era.
On Faithfulness and Factuality in Abstractive Summarization
Joshua Maynez, Shashi Narayan, Bernd Bohnet, and Ryan T. McDonald. 2020 · 2020
Cited alongside, same era.
AmbigQA: Answering Ambiguous Open-domain Questions
Sewon Min, Julian Michael, Hannaneh Hajishirzi, and Luke Zettlemoyer. 2020 · 2020
Cited alongside, same era.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam M. Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, W. Li, and Peter J. Liu. 2020 · 2020
COVID-Fact: Fact Extraction and Verification of Real-World Claims on COVID-19 Pandemic
Arkadiy Saakyan, Tuhin Chakrabarty, and Smaranda Muresan. 2021 · 2021
Later among the works it cites.
Evidence-based Fact-Checking of Health-related Claims
Mourad Sarrouti, Asma Ben Abacha, Yassine Mrabet, and Dina Demner-Fushman. 2021 · 2021
Later among the works it cites.
Get Your Vitamin C! Robust Fact Verification with Contrastive Evidence
Tal Schuster, Adam Fisch, and R. Barzilay. 2021 · 2021
Later among the works it cites.
The Choice of Knowledge Base in Automated Claim Checking
Dominik Stammbach, Boyang Zhang, and Elliott Ash. 2021 · 2021
Later among the works it cites.
BEIR: A Heterogenous Benchmark for Zero-shot Evaluation of Information Retrieval Models
Nandan Thakur, N. Reimers, Andreas Ruckl’e, Abhishek Srivastava, and Iryna Gurevych. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Overview of the TREC 2020 Precision Medicine Track
Kirk Roberts, Dina Demner-Fushman, Ellen M. Voorhees, William R. Hersh, Steven Bedrick, Alexander J. Lazar, and Shubham Pant. 2020b · 2020
Cited alongside, same era.
Automatic Fact-guided Sentence Modification
Darsh J. Shah, Tal Schuster, and R. Barzilay. 2020 · 2020
Cited alongside, same era.
Fact or Fiction: Verifying Scientific Claims
David Wadden, Kyle Lo, Lucy Lu Wang, Shanchuan Lin, Madeleine van Zuylen, Arman Cohan, and Hannaneh Hajishirzi. 2020 · 2020
Cited alongside, same era.
How reliable are the results of large-scale information retrieval experiments?
Justin Zobel. 1998 · 2020
Cited alongside, same era.
MSˆ2: Multi-Document Summarization of Medical Studies
Jay DeYoung, Iz Beltagy, Madeleine van Zuylen, Bailey Kuehl, and Lucy Lu Wang. 2021 · 2021
Cited alongside, same era.
A Paragraph-level Multi-task Learning Model for Scientific Fact-Verification
Xiangci Li, G. Burns, and Nanyun Peng. 2021 · 2021
Cited alongside, same era.
Evidence-based factual error correction
James Thorne and Andreas Vlachos. 2021 · 2021
Later among the works it cites.
Overview and Insights from the SciVer Shared Task on Scientific Claim Verification
David Wadden and Kyle Lo. 2021 · 2021
Later among the works it cites.
Generating (Factual?) Narrative Summaries of RCTs: Experiments with Neural Multi-Document Summarization
Byron C. Wallace, Sayantani Saha, Frank Soboczenski, and Iain James Marshall. 2021 · 2021
Later among the works it cites.
Abstract, Rationale, Stance: A Joint Model for Scientific Claim Verification
Zhiwei Zhang, Jiyi Li, Fumiyo Fukumoto, and Yanming Ye. 2021 · 2021
Later among the works it cites.
BERN2: an advanced neural biomedical named entity recognition and normalization tool
Mujeen Sung, Minbyul Jeong, Yonghwa Choi, Donghyeon Kim, Jinhyuk Lee, and Jaewoo Kang. 2022 · 2022
Closest in time.
MultiVerS: Improving scientific claim verification with weak supervision and full-document context
David Wadden, Kyle Lo, Lucy Lu Wang, Arman Cohan, Iz Beltagy, and Hannaneh Hajishirzi. 2022 · 2022
Closest in time.
Generating Scientific Claims for Zero-Shot Scientific Fact Checking
Dustin Wright, David Wadden, Kyle Lo, Bailey Kuehl, Arman Cohan, Isabelle Augenstein, and Lucy Lu Wang. 2022 · 2022
Closest in time.