Fetching the paper…
Reading the bibliography…
In this paper, we investigate which questions are challenging for retrieval-based Question Answering (QA).
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
“is this document relevant?…probably”: A survey of probabilistic models in information retrieval
Fabio Crestani, Mounia Lalmas, Cornelis J. Van Rijsbergen, and Iain Campbell. 1998 · 1998
Earlier work this paper cites.
BLEU: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2001 · 2001
Earlier work this paper cites.
Query performance prediction in web search environments
Yun Zhou and W. Bruce Croft. 2007 · 2007
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset
Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng. 2016 · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Modeling diverse relevance patterns in ad-hoc retrieval
Yixing Fan, Jiafeng Guo, Yanyan Lan, Jun Xu, Chengxiang Zhai, and Xueqi Cheng. 2018 · 2018
Earlier work this paper cites.
Retrieve and re-rank: A simple and effective IR approach to simple question answering over knowledge graphs
Vishal Gupta, Manoj Chinnakotla, and Manish Shrivastava. 2018 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever. 2018 · 2018
Earlier work this paper cites.
The web as a knowledge-base for answering complex questions
A. Talmor and J. Berant. 2018 · 2018
Earlier work this paper cites.
HotpotQA: A dataset for diverse, explainable multi-hop question answering
Zhilin Yang, Peng Qi, Saizheng Zhang, Yoshua Bengio, William Cohen, Ruslan Salakhutdinov, and Christopher D. Manning. 2018a · 2018
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
DROP: A reading comprehension benchmark requiring discrete reasoning over paragraphs
Dheeru Dua, Yizhong Wang, Pradeep Dasigi, Gabriel Stanovsky, Sameer Singh, and Matt Gardner. 2019 · 2019
Earlier work this paper cites.
Natural questions: a benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Matthew Kelcey, Jacob Devlin, Kenton Lee, Kristina N. Toutanova, Llion Jones, Ming-Wei Chang, Andrew Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov. 2019 · 2019
Earlier work this paper cites.
Electra: Pre-training text encoders as discriminators rather than generators
Kevin Clark, Minh-Thang Luong, Quoc V. Le, and Christopher D. Manning. 2020 · 2020
Earlier work this paper cites.
Tanda: Transfer and adapt pre-trained transformer models for answer sentence selection
Siddhant Garg, Thuy Vu, and Alessandro Moschitti. 2020 · 2020
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020 · 2020
Earlier work this paper cites.
Colbert: Efficient and effective passage search via contextualized late interaction over bert
Omar Khattab and Matei Zaharia. 2020 · 2020
Earlier work this paper cites.
Unsupervised question decomposition for question answering
Ethan Perez, Patrick Lewis, Wen tau Yih, Kyunghyun Cho, and Douwe Kiela. 2020 · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Cited alongside, same era.
Match²: A matching over matching model for similar question identification
Zizhen Wang, Yixing Fan, Jiafeng Guo, Liu Yang, Ruqing Zhang, Yanyan Lan, Xueqi Cheng, Hui Jiang, and Xiaozhao Wang. 2020 · 2020
Cited alongside, same era.
A dataset for answering time-sensitive questions
Wenhu Chen, Xinyi Wang, and William Yang Wang. 2021 · 2021
Cited alongside, same era.
Did Aristotle Use a Laptop? A Question Answering Benchmark with Implicit Reasoning Strategies
Mor Geva, Daniel Khashabi, Elad Segal, Tushar Khot, Dan Roth, and Jonathan Berant. 2021 · 2021
Understand before answer: Improve temporal reading comprehension via precise question understanding
Hao Huang, Xiubo Geng, Guodong Long, and Daxin Jiang. 2022 · 2022
Later among the works it cites.
A survey on multi-hop question answering and generation
Vaibhav Mavi, Anubhav Jangra, and Adam Jatowt. 2022 · 2022
Later among the works it cites.
Improving time sensitivity for question answering over temporal knowledge graphs
Chao Shang, Guangtao Wang, Peng Qi, and Jing Huang. 2022 · 2022
Later among the works it cites.
RoBERTa-based traditional Chinese medicine named entity recognition model
Ming-Hsiang Su, Chin-Wei Lee, Chi-Lun Hsu, and Ruei-Cyuan Su. 2022 · 2022
Later among the works it cites.
BERTScore is unfair: On social bias in language model-based metrics for text generation
Tianxiang Sun, Junliang He, Xipeng Qiu, and Xuanjing Huang. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Pengcheng He, Jianfeng Gao, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
Answer generation for retrieval-based question answering systems
Chao-Chun Hsu, Eric Lind, Luca Soldaini, and Alessandro Moschitti. 2021 · 2021
Cited alongside, same era.
Jimmy Lin and Xueguang Ma. 2021 · 2021
Cited alongside, same era.
Learning passage impacts for inverted indexes
Antonio Mallia, Omar Khattab, Nicola Tonellotto, and Torsten Suel. 2021 · 2021
Cited alongside, same era.
KILT: a benchmark for knowledge intensive language tasks
Fabio Petroni, Aleksandra Piktus, Angela Fan, Patrick Lewis, Majid Yazdani, Nicola De Cao, James Thorne, Yacine Jernite, Vladimir Karpukhin, Jean Maillard, Vassilis Plachouras, Tim Rocktäschel, and Sebastian Riedel. 2021 · 2021
Cited alongside, same era.
Question answering over temporal knowledge graphs
Apoorv Saxena, Soumen Chakrabarti, and Partha Talukdar. 2021 · 2021
Cited alongside, same era.
AVA: an automatic eValuation approach for question answering systems
Thuy Vu and Alessandro Moschitti. 2021 · 2021
Cited alongside, same era.
MuSiQue: Multihop questions via single-hop question composition
Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal. 2022 · 2022
Later among the works it cites.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Later among the works it cites.
Xuan-Quy Dao. 2023 · 2023
Later among the works it cites.
Square: Automatic question answering evaluation using multiple positive and negative references
Matteo Gabburo, Siddhant Garg, Rik Koncel Kedziorski, and Alessandro Moschitti. 2023 · 2023
Later among the works it cites.
Lei Huang, Weijiang Yu, Weitao Ma, Weihong Zhong, Zhangyin Feng, Haotian Wang, Qianglong Chen, Weihua Peng, Xiaocheng Feng, Bing Qin, and Ting Liu. 2023 · 2023
Later among the works it cites.
Albert Qiaochu Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de Las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, Lucile Saulnier, L’elio Renard Lavaud, Marie-Anne Lachaux, Pierre Stock, Teven Le Scao, Thibaut Lavril, Thomas Wang, Timothée Lacroix, and William El Sayed. 2023 · 2023
Later among the works it cites.
Haoran Luo, E. Haihong, Zichen Tang, Shiyao Peng, Yikai Guo, Wentai Zhang, Chenghao Ma, Guanting Dong, Meina Song, and Wei Lin. 2023 · 2023
Later among the works it cites.
Qa dataset explosion: A taxonomy of nlp resources for question answering and reading comprehension
Anna Rogers, Matt Gardner, and Isabelle Augenstein. 2023 · 2023
Later among the works it cites.
Freshllms: Refreshing large language models with search engine augmentation
Tu Vu, Mohit Iyyer, Xuezhi Wang, Noah Constant, Jerry Wei, Jason Wei, Chris Tar, Yun-Hsuan Sung, Denny Zhou, Quoc Le, and Thang Luong. 2023 · 2023
Later among the works it cites.
BLEURT has universal translations: An analysis of automatic metrics by minimum risk training
Yiming Yan, Tao Wang, Chengqi Zhao, Shujian Huang, Jiajun Chen, and Mingxuan Wang. 2023 · 2023
Later among the works it cites.
Answering questions by meta-reasoning over multiple chains of thought
Ori Yoran, Tomer Wolfson, Ben Bogin, Uri Katz, Daniel Deutch, and Jonathan Berant. 2023 · 2023
Later among the works it cites.
TwiRGCN: Temporally weighted graph convolution for question answering over temporal knowledge graphs
Aditya Sharma, Apoorv Saxena, Chitrank Gupta, Mehran Kazemi, Partha Talukdar, and Soumen Chakrabarti. 2023 · 2060
Closest in time.