Fetching the paper…
Reading the bibliography…
Pretrained contextualized language models such as BERT and T5 have established a new state-of-the-art for ad-hoc search.
Rodrigo Nogueira and Kyunghyun Cho. 2019 · 1901
Earlier work this paper cites.
Document expansion by query prediction
Rodrigo Nogueira, Wei Yang, Jimmy Lin, and Kyunghyun Cho. 2019 · 1904
Earlier work this paper cites.
RoBERTa: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, M. Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019b · 1907
Earlier work this paper cites.
Context-aware sentence/passage term importance estimation for first stage retrieval
Zhuyun Dai and J. Callan. 2019a · 1910
Earlier work this paper cites.
HuggingFace’s Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, R’emi Louf, Morgan Funtowicz, and Jamie Brew. 2019 · 1910
Earlier work this paper cites.
Document ranking with a pretrained sequence-to-sequence model
Rodrigo Nogueira, Zhiying Jiang, and Jimmy Lin. 2020 · 2003
Earlier work this paper cites.
A formal study of information retrieval heuristics
Hui Fang, T. Tao, and ChengXiang Zhai. 2004 · 2004
Earlier work this paper cites.
Semantic term matching in axiomatic approaches to information retrieval
Hui Fang and ChengXiang Zhai. 2006 · 2006
Earlier work this paper cites.
Terrier: A high performance and scalable information retrieval platform
I. Ounis, G. Amati, V. Plachouras, B. He, C. Macdonald, and C. Lioma. 2006 · 2006
Earlier work this paper cites.
An exploration of proximity measures in information retrieval
T. Tao and ChengXiang Zhai. 2007 · 2007
Earlier work this paper cites.
Approximate nearest neighbor negative contrastive learning for dense text retrieval
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul Bennett, Junaid Ahmed, and Arnold Overwijk. 2021 · 2007
Earlier work this paper cites.
Parade: Passage representation aggregation for document reranking
Canjia Li, Andrew Yates, Sean MacAvaney, Ben He, and Yingfei Sun. 2020 · 2008
Earlier work this paper cites.
Language models and word sense disambiguation: An overview and analysis
D. Loureiro, Kiamehr Rezaee, Mohammad Taher Pilehvar, and José Camacho-Collados. 2020 · 2008
Earlier work this paper cites.
Pretrained transformers for text ranking: BERT and beyond
Jimmy Lin, Rodrigo Nogueira, and A. Yates. 2020 · 2010
Earlier work this paper cites.
Diagnostic evaluation of information retrieval models
Hui Fang, T. Tao, and ChengXiang Zhai. 2011 · 2011
Earlier work this paper cites.
Gensim–Python framework for vector space modelling
Radim Rehurek and Petr Sojka. 2011 · 2011
Earlier work this paper cites.
Scalable modified Kneser-Ney language model estimation
Kenneth Heafield, Ivan Pouzyrevsky, Jonathan H. Clark, and Philipp Koehn. 2013 · 2013
Earlier work this paper cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D. Manning. 2014 · 2014
Earlier work this paper cites.
From word embeddings to document distances
Matt J. Kusner, Yu Sun, Nicholas I. Kolkin, and Kilian Q. Weinberger. 2015 · 2015
Earlier work this paper cites.
MS MARCO: A human generated machine reading comprehension dataset
Daniel Fernando Campos, T. Nguyen, M. Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, L. Deng, and Bhaskar Mitra. 2016 · 2016
Cited alongside, same era.
Optimizing statistical machine translation for text simplification
Wei Xu, Courtney Napoles, Ellie Pavlick, Quanze Chen, and Chris Callison-Burch. 2016 · 2016
Cited alongside, same era.
spaCy 2: Natural language understanding with Bloom embeddings, convolutional neural networks and incremental parsing
Matthew Honnibal and Ines Montani. 2017 · 2017
Cited alongside, same era.
LightGBM: A highly efficient gradient boosting decision tree
Guolin Ke, Q. Meng, Thomas Finley, Taifeng Wang, Wei Chen, Weidong Ma, Qiwei Ye, and T. Liu. 2017 · 2017
Cited alongside, same era.
JFLEG: A fluency corpus and benchmark for grammatical error correction
Courtney Napoles, Keisuke Sakaguchi, and Joel Tetreault. 2017 · 2017
Cited alongside, same era.
BERT rediscovers the classical NLP pipeline
Ian Tenney, Dipanjan Das, and Ellie Pavlick. 2019 · 2019
Later among the works it cites.
Diagnosing BERT with retrieval heuristics
Arthur Câmara and Claudia Hauff. 2020 · 2020
Closest in time.
Building a better search engine for semantic scholar
Sergey Feldman. 2020 · 2020
Closest in time.
ANTIQUE: A non-factoid question answering benchmark
Helia Hashemi, Mohammad Aliannejadi, Hamed Zamani, and Bruce Croft. 2020 · 2020
Closest in time.
Interpretable and time-budget-constrained contextualization for re-ranking
Sebastian Hofstätter, Markus Zlabinger, and A. Hanbury. 2020 · 2020
Closest in time.
ColBERT: Efficient and effective passage search via contextualized late interaction over BERT
O. Khattab and M. Zaharia. 2020 · 2020
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Get to the point: Summarization with pointer-generator networks
A. See, Peter J. Liu, and Christopher D. Manning. 2017 · 2017
Cited alongside, same era.
Don’t give me the details, just the summary! Topic-aware convolutional neural networks for extreme summarization
Shashi Narayan, Shay B. Cohen, and Mirella Lapata. 2018 · 2018
Cited alongside, same era.
Deep contextualized word representations
Matthew Peters, Mark Neumann, Mohit Iyyer, Matt Gardner, Christopher Clark, Kenton Lee, and Luke Zettlemoyer. 2018 · 2018
Cited alongside, same era.
Dear sir or madam, may I introduce the YAFC corpus: Corpus, benchmarks and metrics for formality style transfer
Sudha Rao and J. Tetreault. 2018 · 2018
Cited alongside, same era.
Anserini: Reproducible ranking baselines using Lucene
Peilin Yang, Hui Fang, and Jimmy Lin. 2018 · 2018
Cited alongside, same era.
Overview of the TREC 2019 deep learning track
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, Daniel Campos, and Ellen M Voorhees. 2019 · 2019
Cited alongside, same era.
CAsT 2019: The conversational assistance track overview
Jeffrey Dalton, Chenyan Xiong, and Jamie Callan. 2019 · 2019
Cited alongside, same era.
OpenNIR: A complete neural ad-hoc ranking pipeline
Sean MacAvaney. 2020 · 2020
Closest in time.
Expansion via prediction of importance with contextualization
Sean MacAvaney, Franco Maria Nardini, Raffaele Perego, Nicola Tonellotto, Nazli Goharian, and Ophir Frieder. 2020 · 2020
Closest in time.
Automatically neutralizing subjective bias in text
Reid Pryzant, Richard Diehl Martinez, Nathan Dass, S. Kurohashi, Dan Jurafsky, and Diyi Yang. 2020 · 2020
Closest in time.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Closest in time.
Beyond accuracy: Behavioral testing of NLP models with checklist
Marco Túlio Ribeiro, Tongshuang Wu, Carlos Guestrin, and Sameer Singh. 2020 · 2020
Closest in time.
A primer in BERTology: What we know about how BERT works
Anna Rogers, O. Kovaleva, and Anna Rumshisky. 2020 · 2020
Closest in time.
Which *BERT? a survey organizing contextualized encoders
Patrick Xia, Shijie Wu, and B. Van Durme. 2020 · 2020
Closest in time.
Matteo Alleman, J. Mamou, M. D. Rio, Hanlin Tang, Yoon Kim, and SueYeon Chung. 2021 · 2021
Closest in time.
Evaluating multilingual text encoders for unsupervised cross-lingual retrieval
Robert Litschko, Ivan Vuli’c, Simone Paolo Ponzetto, and Goran Glavavs. 2021 · 2021
Closest in time.
Simplified data wrangling with ir_datasets
Sean MacAvaney, Andrew Yates, Sergey Feldman, Doug Downey, Arman Cohan, and Nazli Goharian. 2021 · 2021
Closest in time.
PyTerrier: Declarative experimentation in python from BM25 to dense retrieval
Craig Macdonald, Nicola Tonellotto, Sean MacAvaney, and Iadh Ounis. 2021 · 2021
Closest in time.
Koustuv Sinha, Robin Jia, Dieuwke Hupkes, J. Pineau, Adina Williams, and Douwe Kiela. 2021 · 2021
Closest in time.
Towards axiomatic explanations for neural ranking models
Michael Völske, A. Bondarenko, Maik Fröbe, Matthias Hagen, Benno Stein, Jaspreet Singh, and Avishek Anand. 2021 · 2021
Closest in time.