Fetching the paper…
Reading the bibliography…
Automated fact-checking systems verify claims against evidence to predict their veracity.
Stop Measuring Calibration When Humans Disagree
Joris Baan, Wilker Aziz, Barbara Plank, and Raquel Fernandez. 2022 · 1915
Earlier work this paper cites.
Presupposition and Linguistic Context
Lauri Karttunen. 1974 · 1974
Earlier work this paper cites.
Logic and conversation
HP Grice. 1975 · 1975
Earlier work this paper cites.
Maximum Likelihood Estimation of Observer Error-Rates Using the EM Algorithm
Alexander Philip Dawid and Allan M. Skene. 1979 · 1979
Earlier work this paper cites.
Coherence and coreference
Jerry R Hobbs. 1979 · 1979
Earlier work this paper cites.
Vagueness: A Reader
Rosanna Kenney and Peter Smith. 1997 · 1997
Earlier work this paper cites.
The PASCAL Recognising Textual Entailment Challenge
Ido Dagan, Oren Glickman, and Bernardo Magnini. 2006 · 2005
Earlier work this paper cites.
Generic Statements Require Little Evidence for Acceptance but Have Powerful Implications
Andrei Cimpian, Amanda C. Brandone, and Susan A. Gelman. 2010 · 2010
Earlier work this paper cites.
Distilling the Knowledge in a Neural Network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2015 · 2015
Earlier work this paper cites.
On Calibration of Modern Neural Networks
Chuan Guo, Geoff Pleiss, Yu Sun, and Kilian Q. Weinberger. 2017 · 2017
Earlier work this paper cites.
Argumentation Mining in User-Generated Web Discourse
Ivan Habernal and Iryna Gurevych. 2017 · 2017
Earlier work this paper cites.
Checking how fact-checkers check
Chloe Lim. 2018 · 2018
Earlier work this paper cites.
FEVER: a Large-scale Dataset for Fact Extraction and VERification
James Thorne, Andreas Vlachos, Christos Christodoulopoulos, and Arpit Mittal. 2018 · 2018
Earlier work this paper cites.
A Broad-Coverage Challenge Corpus for Sentence Understanding through Inference
Adina Williams, Nikita Nangia, and Samuel Bowman. 2018 · 2018
Earlier work this paper cites.
MultiFC: A Real-World Multi-Domain Dataset for Evidence-Based Fact Checking of Claims
Isabelle Augenstein, Christina Lioma, Dongsheng Wang, Lucas Chaves Lima, Casper Hansen, Christian Hansen, and Jakob Grue Simonsen. 2019 · 2019
Earlier work this paper cites.
Seeing Things from a Different Angle: Discovering Diverse Perspectives about Claims
Sihao Chen, Daniel Khashabi, Wenpeng Yin, Chris Callison-Burch, and Dan Roth. 2019 · 2019
Earlier work this paper cites.
BoolQ: Exploring the surprising difficulty of natural yes/no questions
Christopher Clark, Kenton Lee, Ming-Wei Chang, Tom Kwiatkowski, Michael Collins, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
A Richly Annotated Corpus for Different Tasks in Automated Fact-Checking
Andreas Hanselowski, Christian Stab, Claudia Schulz, Zile Li, and Iryna Gurevych. 2019 · 2019
Earlier work this paper cites.
Inherent Disagreements in Human Textual Inferences
Ellie Pavlick and Tom Kwiatkowski. 2019 · 2019
Cited alongside, same era.
Human Uncertainty Makes Classification More Robust
Joshua C Peterson, Ruairidh M Battleday, Thomas L Griffiths, and Olga Russakovsky. 2019 · 2019
Cited alongside, same era.
Towards Debiasing Fact Verification Models
Tal Schuster, Darsh Shah, Yun Jie Serene Yeo, Daniel Roberto Filizzola Ortiz, Enrico Santus, and Regina Barzilay. 2019 · 2019
Cited alongside, same era.
Calibration of Pre-trained Transformers
Shrey Desai and Greg Durrett. 2020 · 2020
Cited alongside, same era.
CLIMATE-FEVER: A Dataset for Verification of Real-World Climate Claims
Thomas Diggelmann, Jordan Boyd-Graber, Jannis Bulian, Massimiliano Ciaramita, and Markus Leippold. 2020 · 2020
Cited alongside, same era.
HoVer: A Dataset for Many-Hop Fact Extraction And Claim Verification
Evidence-based Fact-Checking of Health-related Claims
Mourad Sarrouti, Asma Ben Abacha, Yassine Mrabet, and Dina Demner-Fushman. 2021 · 2021
Closest in time.
Get Your Vitamin C! Robust Fact Verification with Contrastive Evidence
Tal Schuster, Adam Fisch, and Regina Barzilay. 2021 · 2021
Closest in time.
A General-Purpose Crowdsourcing Computational Quality Control Toolkit for Python
Dmitry Ustalov, Nikita Pavlichenko, Vladimir Losev, Iulian Giliazev, and Evgeny Tulin. 2021 · 2021
Closest in time.
FaithDial: A Faithful Benchmark for Information-Seeking Dialogue
Nouha Dziri, Ehsan Kamalloo, Sivan Milton, Osmar Zaiane, Mo Yu, Edoardo M. Ponti, and Siva Reddy. 2022 · 2022
Closest in time.
Missing Counter-Evidence Renders NLP Fact-Checking Unrealistic for Misinformation
Max Glockner, Yufang Hou, and Iryna Gurevych. 2022 · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yichen Jiang, Shikha Bordia, Zheng Zhong, Charles Dognin, Maneesh Singh, and Mohit Bansal. 2020 · 2020
Cited alongside, same era.
AmbigQA: Answering Ambiguous Open-domain Questions
Sewon Min, Julian Michael, Hannaneh Hajishirzi, and Luke Zettlemoyer. 2020 · 2020
Cited alongside, same era.
What Can We Learn from Collective Human Opinions on Natural Language Inference Data?
Yixin Nie, Xiang Zhou, and Mohit Bansal. 2020 · 2020
Cited alongside, same era.
Automated Fact-Checking of Claims from Wikipedia
Aalok Sathe, Salar Ather, Tuan Manh Le, Nathan Perry, and Joonsuk Park. 2020 · 2020
Cited alongside, same era.
A Case for Soft Loss Functions
Alexandra Uma, Tommaso Fornaciari, Dirk Hovy, Silviu Paun, Barbara Plank, and Massimo Poesio. 2020 · 2020
Cited alongside, same era.
Fact or Fiction: Verifying Scientific Claims
David Wadden, Shanchuan Lin, Kyle Lo, Lucy Lu Wang, Madeleine van Zuylen, Arman Cohan, and Hannaneh Hajishirzi. 2020 · 2020
Cited alongside, same era.
Transformers: State-of-the-Art Natural Language Processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush. 2020 · 2020
Cited alongside, same era.
Zhijiang Guo, Michael Schlichtkrull, and Andreas Vlachos. 2022 · 2022
Closest in time.
Investigating Reasons for Disagreement in Natural Language Inference
Nan-Jiang Jiang and Marie-Catherine de Marneffe. 2022 · 2022
Closest in time.
WatClaimCheck: A new Dataset for Claim Entailment and Inference
Kashif Khan, Ruizhe Wang, and Pascal Poupart. 2022 · 2022
Closest in time.
FaVIQ: FAct Verification from Information-seeking Questions
Jungsoo Park, Sewon Min, Jaewoo Kang, Luke Zettlemoyer, and Hannaneh Hajishirzi. 2022 · 2022
Closest in time.
The “Problem” of Human Label Variation: On Ground Truth in Data, Modeling and Evaluation
Barbara Plank. 2022 · 2022
Closest in time.
Learning from Disagreement: A Survey
Alexandra Uma, Tommaso Fornaciari, Dirk Hovy, Silviu Paun, Barbara Plank, and Massimo Poesio. 2022 · 2022
Closest in time.
When the Majority is Wrong: Leveraging Annotator Disagreement for Subjective Tasks
Eve Fleisig, Rediet Abebe, and Dan Klein. 2023 · 2023
Closest in time.
Survey of Hallucination in Natural Language Generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Ye Jin Bang, Andrea Madotto, and Pascale Fung. 2023 · 2023
Closest in time.
WiCE: Real-World Entailment for Claims in Wikipedia
Ryo Kamoi, Tanya Goyal, Juan Diego Rodriguez, and Greg Durrett. 2023 · 2023
Closest in time.
FactKG: Fact Verification via Reasoning on Knowledge Graphs
Jiho Kim, Sungjin Park, Yeonsu Kwon, Yohan Jo, James Thorne, and Edward Choi. 2023 · 2023
Closest in time.
SemEval-2023 Task 11: Learning With Disagreements (LeWiDi)
Elisa Leonardelli, Alexandra Uma, Gavin Abercrombie, Dina Almanea, Valerio Basile, Tommaso Fornaciari, Barbara Plank, Verena Rieser, and Massimo Poesio. 2023 · 2023
Closest in time.
AVeriTeC: A Dataset for Real-world Claim Verification with Evidence from the Web
Michael Schlichtkrull, Zhijiang Guo, and Andreas Vlachos. 2023 · 2023
Closest in time.
Multi2Claim: Generating Scientific Claims from Multi-Choice Questions for Scientific Fact-Checking
Neset Tan, Trung Nguyen, Josh Bensemann, Alex Peng, Qiming Bao, Yang Chen, Mark Gahegan, and Michael Witbrock. 2023 · 2023
Closest in time.