Fetching the paper…
Reading the bibliography…
The goal of text ranking is to generate an ordered list of texts retrieved from a corpus in response to a query.
An updated Duet model for passage re-ranking
B. Mitra and N. Craswell · 1903
Earlier work this paper cites.
Simple applications of BERT for ad hoc document retrieval
W. Yang, H. Zhang, and J. Lin · 1903
Earlier work this paper cites.
Document expansion by query prediction
R. Nogueira, W. Yang, J. Lin, and K. Cho · 1904
Earlier work this paper cites.
ERNIE: Enhanced representation through knowledge integration
Y. Sun, S. Wang, Y. Li, S. Feng, X. Chen, H. Zhang, X. Tian, D. Zhu, H. Tian, and H. Wu · 1904
Earlier work this paper cites.
Data augmentation for BERT fine-tuning in open-domain question answering
W. Yang, Y. Xie, L. Tan, K. Xiong, M. Li, and J. Lin · 1904
Earlier work this paper cites.
RoBERTa: A robustly optimized BERT pretraining approach
Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer, and V. Stoyanov · 1907
Earlier work this paper cites.
Context-aware sentence/passage term importance estimation for first stage retrieval
Z. Dai and J. Callan · 1910
Earlier work this paper cites.
Multi-stage document ranking with BERT
R. Nogueira, W. Yang, K. Cho, and J. Lin · 1910
Earlier work this paper cites.
What would Elsa do? Freezing layers during transformer fine-tuning
J. Lee, R. Tang, and J. Lin · 1911
Earlier work this paper cites.
Attentive student meets multi-task teacher: Improved knowledge distillation for pretrained models
L. Liu, H. Wang, J. Lin, R. Socher, and C. Xiong · 1911
Earlier work this paper cites.
An analysis of BERT in document ranking
J. Zhan, J. Mao, Y. Liu, M. Zhang, and S. Ma · 1944
Earlier work this paper cites.
As we may think
V. Bush · 1945
Earlier work this paper cites.
Section III. Opening plenary session
J. E. Holmstrom · 1948
Earlier work this paper cites.
Cloze procedure: A new tool for measuring readability
W. L. Taylor · 1953
Earlier work this paper cites.
The automatic creation of literature abstracts
H. P. Luhn · 1958
Earlier work this paper cites.
On relevance, probabilistic indexing and information retrieval
M. E. Maron and J. L. Kuhns · 1960
Earlier work this paper cites.
The Structure of Scientific Revolutions
T. S. Kuhn · 1962
Earlier work this paper cites.
The process of asking questions
R. S. Taylor · 1962
Earlier work this paper cites.
Probability of error of some adaptive pattern-recognition machines
H. Scudder · 1965
Earlier work this paper cites.
Answering English questions by computer: A survey
R. F. Simmons · 1965
Earlier work this paper cites.
Document indexing based on relevance feedback
T. L. Brauen, R. C. Holt, and T. R. Wilcox · 1968
Earlier work this paper cites.
Relevance assessments and retrieval system evaluation
M. E. Lesk and G. Salton · 1968
Earlier work this paper cites.
Automatic content analysis in information retrieval
G. Salton · 1968
Earlier work this paper cites.
Computer evaluation of indexing and text processing
G. Salton and M. E. Lesk · 1968
Earlier work this paper cites.
State of the art in selective dissemination of information
E. M. Housman and E. D. Kaskela · 1970
Earlier work this paper cites.
Relevance feedback in information retrieval
J. J. Rocchio · 1971
Earlier work this paper cites.
A new comparison between conventional indexing (MEDLARS) and automatic text processing (SMART)
G. Salton · 1972
Earlier work this paper cites.
The use of title and cited titles as document representation for automatic classification
K. L. Kwok · 1975
Earlier work this paper cites.
A vector space model for automatic indexing
G. Salton, Y. Wong, and C.-S. Yang · 1975
Earlier work this paper cites.
Relevance: A review of and a framework for thinking on the notion in information science
T. Saracevic · 1975
Earlier work this paper cites.
Report on the need for and provision of an “ideal” information retrieval test collection
K. Sparck Jones and C. J. van Rijsbergen · 1975
Earlier work this paper cites.
Relevance weighting of search terms
S. Robertson and K. Spark Jones · 1976
Earlier work this paper cites.
The probability ranking principle in IR
S. Robertson · 1977
Earlier work this paper cites.
Probabilistic models of document retrieval with relevance information
W. B. Croft and D. J. Harper · 1979
Earlier work this paper cites.
The “file drawer problem” and tolerance for null results
R. Rosenthal · 1979
Earlier work this paper cites.
Anomalous states of knowledge as a basis for information retrieval
N. J. Belkin · 1980
Earlier work this paper cites.
Implementation of the SMART information retrieval system
C. Buckley · 1985
Earlier work this paper cites.
The vocabulary problem in human-system communication
G. W. Furnas, T. K. Landauer, L. M. Gomez, and S. T. Dumais · 1987
Earlier work this paper cites.
On the use of spreading activation methods in automatic information
G. Salton and C. Buckley · 1988
Earlier work this paper cites.
Optimum polynomial retrieval functions based on the probability ranking principle
N. Fuhr · 1989
Earlier work this paper cites.
A statistical approach to machine translation
P. F. Brown, J. Cocke, S. D. Pietra, V. J. D. Pietra, F. Jelinek, J. D. Lafferty, R. L. Mercer, and P. S. Roossin · 1990
Earlier work this paper cites.
Indexing by latent semantic analysis
S. Deerwester, S. T. Dumais, G. W. Furnas, T. K. Landauer, and R. Harshman · 1990
Earlier work this paper cites.
Information filtering and information retrieval: Two sides of the same coin?
N. J. Belkin and W. B. Croft · 1992
Earlier work this paper cites.
Subtopic structuring for full-length document access
M. A. Hearst and C. Plaunt · 1993
Earlier work this paper cites.
Approaches to passage retrieval in full text information systems
G. Salton, J. Allan, and C. Buckley · 1993
Earlier work this paper cites.
Vector expansion in a large collection
E. M. Voorhees and Y.-W. Hou · 1993
Earlier work this paper cites.
Computation of term association by a neural network
S. K. M. Wong, Y. J. Cai, and Y. Y. Yao · 1993
Earlier work this paper cites.
Automatic combination of multiple ranked retrieval systems
B. T. Bartell, G. W. Cottrell, and R. K. Belew · 1994
Earlier work this paper cites.
Passage-level evidence in document retrieval
J. P. Callan · 1994
Earlier work this paper cites.
Inferring probability of relevance using the method of logistic regression
F. C. Gey · 1994
Earlier work this paper cites.
Okapi at TREC-3
S. Robertson, S. Walker, S. Jones, M. Hancock-Beaulieu, and M. Gatford · 1994
Earlier work this paper cites.
Query expansion using lexical-semantic relations
E. M. Voorhees · 1994
Earlier work this paper cites.
Effective retrieval of structured documents
R. Wilkinson · 1994
Earlier work this paper cites.
Query by humming: Musical information retrieval in an audio database
A. Ghias, J. Logan, D. Chamberlin, and B. C. Smith · 1995
Earlier work this paper cites.
“The whisky was invisible”, or persistent myths of MT
J. Hutchins · 1995
Earlier work this paper cites.
The TREC-4 filtering track
D. D. Lewis · 1995
Earlier work this paper cites.
WordNet: A lexical database for English
G. Miller · 1995
Earlier work this paper cites.
Unsupervised word sense disambiguation rivaling supervised methods
D. Yarowsky · 1995
Earlier work this paper cites.
An empirical study of smoothing techniques for language modeling
S. F. Chen and J. Goodman · 1996
Earlier work this paper cites.
Automatically generating extraction patterns from untagged text
E. Riloff · 1996
Earlier work this paper cites.
Pivoted document length normalization
A. Singhal, C. Buckley, and M. Mitra · 1996
Earlier work this paper cites.
A comparison of head transducers and transfer for a limited domain translation application
H. Alshawi, A. L. Buchsbaum, and F. Xia · 1997
Earlier work this paper cites.
Question answering from frequently-asked question files: Experiences with the FAQ Finder system
R. D. Burke, K. J. Hammond, V. A. Kulyukin, S. L. Lytinen, N. Tomuro, and S. Schoenberg · 1997
Earlier work this paper cites.
Coupling information retrieval and information extraction: A new text technology for gathering information from the web
R. Gaizauskas and A. M. Robertson · 1997
Earlier work this paper cites.
Passage retrieval revisited
M. Kaszkiel and J. Zobel · 1997
Earlier work this paper cites.
Okapi at TREC-6 automatic ad hoc, VLC, routing, filtering and QSDR
S. Walker, S. E. Robertson, M. Boughanem, G. J. Jones, and K. S. Jones · 1997
Earlier work this paper cites.
Extracting patterns and relations from the World Wide Web
S. Brin · 1998
Earlier work this paper cites.
Approximate nearest neighbors: Towards removing the curse of dimensionality
P. Indyk and R. Motwani · 1998
Earlier work this paper cites.
A language modeling approach to information retrieval
J. M. Ponte and W. B. Croft · 1998
Earlier work this paper cites.
Overview of the Seventh Text REtrieval Conference (TREC-7)
E. M. Voorhees and D. Harman · 1998
Earlier work this paper cites.
How reliable are the results of large-scale information retrieval experiments?
J. Zobel · 1998
Earlier work this paper cites.
Information retrieval as statistical translation
A. Berger and J. Lafferty · 1999
Earlier work this paper cites.
“Is this document relevant?… probably”: A survey of probabilistic models in information retrieval
F. Crestani, M. Lalmas, C. J. van Rijsbergen, and I. Campbell · 1999
Earlier work this paper cites.
Similarity search in high dimensions via hashing
A. Gionis, P. Indyk, and R. Motwani · 1999
Earlier work this paper cites.
Document expansion for speech retrieval
A. Singhal and F. Pereira · 1999
Earlier work this paper cites.
Fusion via a linear combination of scores
C. C. Vogt and G. W. Cottrell · 1999
Earlier work this paper cites.
Snowball: Extracting relations from large plain-text collections
E. Agichtein and L. Gravano · 2000
Earlier work this paper cites.
Relevance ranking for one to three term queries
C. L. A. Clarke, G. Cormack, and E. Tudhope · 2000
Earlier work this paper cites.
Do batch and user evaluations give the same results?
W. R. Hersh, A. Turpin, S. Price, B. Chan, D. Kramer, L. Sacherek, and D. Olson · 2000
Earlier work this paper cites.
Variations in relevance judgments and the measurement of retrieval effectiveness
E. M. Voorhees · 2000
Earlier work this paper cites.
Improving the effectiveness of information retrieval with local context analysis
J. Xu and W. B. Croft · 2000
Earlier work this paper cites.
Vector-space ranking with effective early termination
V. N. Anh, O. de Kretser, and A. Moffat · 2001
Earlier work this paper cites.
Scaling to very very large corpora for natural language disambiguation
M. Banko and E. Brill · 2001
Earlier work this paper cites.
Automatic labeling of semantic roles
D. Gildea and D. Jurafsky · 2001
Earlier work this paper cites.
Relevance-based language models
V. Lavrenko and W. B. Croft · 2001
Earlier work this paper cites.
Regions and levels: Mapping and measuring users’ relevance judgments
A. Spink and H. Greisdorf · 2001
Earlier work this paper cites.
Overview of the TREC 2001 question answering track
E. M. Voorhees · 2001
Earlier work this paper cites.
Global term weights for document retrieval learned from TREC data
W. J. Wilbur · 2001
Earlier work this paper cites.
Topic Detection and Tracking: Event-Based Information Organization
J. Allan · 2002
Earlier work this paper cites.
Probabilistic models of information retrieval based on measuring the divergence from randomness
G. Amati and C. J. van Rijsbergen · 2002
Earlier work this paper cites.
Impact transformation: Effective and efficient web retrieval
V. N. Anh and A. Moffat · 2002
Earlier work this paper cites.
Title language model for information retrieval
R. Jin, A. G. Hauptmann, and C. X. Zhai · 2002
Earlier work this paper cites.
Optimizing search engines using clickthrough data
T. Joachims · 2002
Earlier work this paper cites.
SUMMAC: A text summarization evaluation
I. Mani, G. Klein, D. House, and L. Hirschman · 2002
Earlier work this paper cites.
Condorcet fusion for improved retrieval
M. Montague and J. A. Aslam · 2002
Earlier work this paper cites.
The TREC 2002 filtering track report
S. Robertson and I. Soboroff · 2002
Earlier work this paper cites.
Liberal relevance criteria of TREC—counting on negligible documents?
E. Sormunen · 2002
Earlier work this paper cites.
Probabilistic Models of Information Retrieval Based on Divergence from Randomness
G. Amati · 2003
Earlier work this paper cites.
The Google File System
S. Ghemawat, H. Gobioff, and S.-T. Leung · 2003
Earlier work this paper cites.
The web as a parallel corpus
P. Resnik and N. A. Smith · 2003
Earlier work this paper cites.
UMass at TREC 2004: Novelty and HARD
N. Abdul-Jaleel, J. Allan, W. B. Croft, F. Diaz, L. Larkey, X. Li, D. Metzler, M. D. Smucker, T. Strohman, H. Turtle, and C. Wade · 2004
Earlier work this paper cites.
Ask Jeeves bets on smart search, 2004
BBC · 2004
Earlier work this paper cites.
Retrieval evaluation with incomplete information
C. Buckley and E. M. Voorhees · 2004
Earlier work this paper cites.
MapReduce: Simplified data processing on large clusters
J. Dean and S. Ghemawat · 2004
Earlier work this paper cites.
A formal study of information retrieval heuristics
H. Fang, T. Tao, and C. Zhai · 2004
Earlier work this paper cites.
EARL: Speedup transformer-based rankers with pre-computed representation
L. Gao, Z. Dai, and J. Callan · 2004
Earlier work this paper cites.
Complementing lexical retrieval with semantic residual embedding
L. Gao, Z. Dai, Z. Fan, and J. Callan · 2004
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering
V. Karpukhin, B. Oğuz, S. Min, P. Lewis, L. Wu, S. Edunov, D. Chen, and W.-t. Yih · 2004
Earlier work this paper cites.
Combining the language model and inference network approaches to retrieval
D. Metzler and W. B. Croft · 2004
Earlier work this paper cites.
Indri at TREC 2004: Terabyte track
D. Metzler, T. Strohman, H. Turtle, and W. B. Croft · 2004
Earlier work this paper cites.
Direct maximization of average precision by hill-climbing, with a comparison to a maximum entropy approach
W. Morgan, W. Greiff, and J. Henderson · 2004
Earlier work this paper cites.
University of Glasgow at TREC2004: Experiments in web, robust and terabyte tracks with Terrier
V. Plachouras, B. He, and I. Ounis · 2004
Earlier work this paper cites.
Overview of the TREC 2004 robust track
E. M. Voorhees · 2004
Earlier work this paper cites.
CORD-19: The COVID-19 Open Research Dataset
L. L. Wang, K. Lo, Y. Chandrasekhar, R. Reas, J. Yang, D. Burdick, D. Eide, K. Funk, Y. Katsis, R. Kinney, Y. Li, Z. Liu, W. Merrill, P. Mooney, D. Murdick, D. Rishi, J. Sheehan, Z. Shen, B. Stilson, A. Wade, K. Wang, N. X. R. Wang, C. Wilhelm, B. Xie, D. Raymond, D. S. Weld, O. Etzioni, and S. Kohlmeier · 2004
Earlier work this paper cites.
Beyond 512 tokens: Siamese multi-depth transformer-based hierarchical encoder for document matching
L. Yang, M. Zhang, C. Li, M. Bendersky, and M. Najork · 2004
Earlier work this paper cites.
When will information retrieval be “good enough”? User effectiveness as a function of retrieval accuracy
J. Allan, B. Carterette, and J. Lewis · 2005
Earlier work this paper cites.
Paraphrasing with bilingual parallel corpora
C. Bannard and C. Callison-Burch · 2005
Earlier work this paper cites.
LSH forest: Self tuning indexes for similarity search
M. Bawa, T. Condie, and P. Ganesan · 2005
Earlier work this paper cites.
MG4J at TREC 2005
P. Boldi and S. Vigna · 2005
Earlier work this paper cites.
Learning to rank using gradient descent
C. J. C. Burges, T. Shaked, E. Renshaw, A. Lazier, M. Deeds, N. Hamilton, and G. Hullender · 2005
Earlier work this paper cites.
Overview of DUC 2005
H. T. Dang · 2005
Earlier work this paper cites.
SLEDGE: A simple yet effective baseline for coronavirus scientific knowledge search
S. MacAvaney, A. Cohan, and N. Goharian · 2005
Earlier work this paper cites.
A Markov random field model for term dependencies
D. Metzler and W. B. Croft · 2005
Earlier work this paper cites.
Improving machine translation performance by exploiting non-parallel corpora
D. S. Munteanu and D. Marcu · 2005
Earlier work this paper cites.
Query chains: Learning to rank from implicit feedback
F. Radlinski and T. Joachims · 2005
Earlier work this paper cites.
Information retrieval system evaluation: Effort, sensitivity, and reliability
M. Sanderson and J. Zobel · 2005
Earlier work this paper cites.
TREC: Experiment and Evaluation in Information Retrieval
E. M. Voorhees and D. K. Harman · 2005
Earlier work this paper cites.
Semantic term matching in axiomatic approaches to information retrieval
H. Fang and C. Zhai · 2006
Earlier work this paper cites.
High accuracy retrieval with multiple nested ranker
I. Matveeva, C. Burges, T. Burkard, A. Laucius, and L. Wong · 2006
Earlier work this paper cites.
Terrier: A high performance and scalable information retrieval platform
I. Ounis, G. Amati, V. Plachouras, B. He, C. Macdonald, and C. Lioma · 2006
Earlier work this paper cites.
A picture of search
G. Pass, A. Chowdhury, and C. Torgeson · 2006
Earlier work this paper cites.
Find-Similar: Similarity browsing as a search tool
M. D. Smucker and J. Allan · 2006
Earlier work this paper cites.
LDA-based document models for ad-hoc retrieval
X. Wei and W. B. Croft · 2006
Earlier work this paper cites.
Supervised probabilistic principal component analysis
S. Yu, K. Yu, V. Tresp, H.-P. Kriegel, and M. Wu · 2006
Earlier work this paper cites.
RepBERT: Contextualized text embeddings for first-stage retrieval
J. Zhan, J. Mao, Y. Liu, M. Zhang, and S. Ma · 2006
Earlier work this paper cites.
Revisiting few-sample BERT fine-tuning
T. Zhang, F. Wu, A. Katiyar, K. Q. Weinberger, and Y. Artzi · 2006
Earlier work this paper cites.
Inverted files for text search engines
J. Zobel and A. Moffat · 2006
Earlier work this paper cites.
FUB, IASI-CNR and University of Tor Vergata at TREC 2007 blog track
G. Amati, E. Ambrosi, M. Bianchi, C. Gaibisso, and G. Gambosi · 2007
Earlier work this paper cites.
Large language models in machine translation
T. Brants, A. C. Popat, P. Xu, F. J. Och, and J. Dean · 2007
Earlier work this paper cites.
Bias and the limits of pooling for large collections
C. Buckley, D. Dimmick, I. Soboroff, and E. Voorhees · 2007
Earlier work this paper cites.
Learning to rank: From pairwise approach to listwise approach
Z. Cao, T. Qin, T.-Y. Liu, M.-F. Tsai, and H. Li · 2007
Earlier work this paper cites.
The third PASCAL recognizing textual entailment challenge
D. Giampiccolo, B. Magnini, I. Dagan, and B. Dolan · 2007
Earlier work this paper cites.
Evaluating the accuracy of implicit feedback from clicks and query reformulations in Web search
T. Joachims, L. Granka, B. Pan, H. Hembrooke, F. Radlinski, and G. Gay · 2007
Earlier work this paper cites.
Practical guide to controlled experiments on the web: Listen to your customers not to the HiPPO
R. Kohavi, R. M. Henne, and D. Sommerfield · 2007
Earlier work this paper cites.
PubMed related articles: A probabilistic topic-based model for content similarity
J. Lin and W. J. Wilbur · 2007
Earlier work this paper cites.
Alternatives to bpref
T. Sakai · 2007
Earlier work this paper cites.
IR evaluation using multiple assessors per topic
A. Trotman and D. Jenkinson · 2007
Earlier work this paper cites.
Deep reinforced query reformulation for information retrieval
X. Wang, C. Macdonald, and I. Ounis · 2007
Earlier work this paper cites.
Approximate nearest neighbor negative contrastive learning for dense text retrieval
L. Xiong, C. Xiong, Y. Li, K.-F. Tang, J. Liu, P. Bennett, J. Ahmed, and A. Overwijk · 2007
Earlier work this paper cites.
Relevance assessment: Are judges exchangeable and does it matter?
P. Bailey, N. Craswell, I. Soboroff, P. Thomas, A. P. de Vries, and E. Yilmaz · 2008
Earlier work this paper cites.
Discovering key concepts in verbose queries
M. Bendersky and W. B. Croft · 2008
Earlier work this paper cites.
PARADE: Passage representation aggregation for document reranking
C. Li, A. Yates, S. MacAvaney, B. He, and Y. Sun · 2008
Earlier work this paper cites.
Rank-biased precision for measurement of retrieval effectiveness
A. Moffat and J. Zobel · 2008
Earlier work this paper cites.
Taking notes on the fly helps BERT pre-training
Q. Wu, C. Xing, Y. Li, G. Ke, D. He, and T.-Y. Liu · 2008
Earlier work this paper cites.
Statistical Language Models for Information Retrieval
C. Zhai · 2008
Earlier work this paper cites.
Improvements that don’t add up: Ad-hoc retrieval results since 1998
T. G. Armstrong, A. Moffat, W. Webber, and J. Zobel · 2009
Earlier work this paper cites.
Curriculum learning
Y. Bengio, J. Louradour, R. Collobert, and J. Weston · 2009
Earlier work this paper cites.
Machine learning for information retrieval: TREC 2009 web, relevance feedback and legal tracks
G. V. Cormack and M. Mojdeh · 2009
Earlier work this paper cites.
Reciprocal rank fusion outperforms Condorcet and individual rank learning methods
G. V. Cormack, C. L. A. Clarke, and S. Büttcher · 2009
Earlier work this paper cites.
The difficulty of training deep architectures and the effect of unsupervised pre-training
D. Erhan, P.-A. Manzagol, Y. Bengio, S. Bengio, and P. Vincent · 2009
Earlier work this paper cites.
Methods for evaluating interactive information retrieval systems with users
D. Kelly · 2009
Earlier work this paper cites.
Is searching full text more effective than searching abstracts?
J. Lin · 2009
Earlier work this paper cites.
Of Ivory and Smurfs: Loxodontan MapReduce experiments for web search
J. Lin, D. Metzler, T. Elsayed, and L. Wang · 2009
Earlier work this paper cites.
Learning to rank for information retrieval
T.-Y. Liu · 2009
Earlier work this paper cites.
Score aggregation techniques in retrieval experimentation
S. D. Ravana and A. Moffat · 2009
Earlier work this paper cites.
The probabilistic relevance framework: BM25 and beyond
S. Robertson and H. Zaragoza · 2009
Earlier work this paper cites.
Project Blacklight: A next generation library catalog at a first generation university
E. B. Sadler · 2009
Earlier work this paper cites.
From RankNet to LambdaRank to LambdaMART: An overview
C. J. C. Burges · 2010
Earlier work this paper cites.
Tie-breaking bias: Effect of an uncontrolled parameter on information retrieval evaluation
G. Cabanac, G. Hubert, M. Boughanem, and C. Chrisment · 2010
Earlier work this paper cites.
Early exit optimizations for additive machine learned ranking systems
B. B. Cambazoglu, H. Zaragoza, O. Chapelle, J. Chen, C. Liao, Z. Zheng, and J. Degenhardt · 2010
Earlier work this paper cites.
Distilling dense representations for ranking using tightly-coupled teachers
S.-C. Lin, J.-H. Yang, and J. Lin · 2010
Earlier work this paper cites.
Semantic Role Labeling
M. Palmer, D. Gildea, and N. Xue · 2010
Earlier work this paper cites.
Google gets smarter & says there’s more to come, 2010
ReadWrite · 2010
Cited alongside, same era.
Economic impact assessment of NIST’s Text REtrieval Conference (TREC) program: Final report
B. R. Rowe, D. W. Wood, A. N. Link, and D. A. Simoni · 2010
Cited alongside, same era.
Extracting parallel sentences from comparable corpora using document level alignment
J. R. Smith, C. Quirk, and K. Toutanova · 2010
Cited alongside, same era.
Large scale parallel document mining for machine translation
J. Uszkoreit, J. Ponte, A. Popat, and M. Dubiner · 2010
Cited alongside, same era.
Learning to efficiently rank
L. Wang, J. Lin, and D. Metzler · 2010
Cited alongside, same era.
Learning to retrieve: How to train a dense retrieval model effectively and efficiently
Natural Questions: A benchmark for question answering research
T. Kwiatkowski, J. Palomaki, O. Redfield, M. Collins, A. Parikh, C. Alberti, D. Epstein, I. Polosukhin, J. Devlin, K. Lee, K. Toutanova, L. Jones, M. Kelcey, M.-W. Chang, A. M. Dai, J. Uszkoreit, Q. Le, and S. Petrov · 2019
Later among the works it cites.
The neural hype, justified! A recantation
J. Lin · 2019
Later among the works it cites.
The impact of score ties on repeatability in document ranking
J. Lin and P. Yang · 2019
Later among the works it cites.
Hierarchical transformers for multi-document summarization
Y. Liu and M. Lapata · 2019
Later among the works it cites.
CEDR: Contextualized embeddings for document ranking
S. MacAvaney, A. Yates, A. Cohan, and N. Goharian · 2019
Later among the works it cites.
Content-based weak supervision for ad-hoc re-ranking
S. MacAvaney, A. Yates, K. Hui, and O. Frieder · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Zhan, J. Mao, Y. Liu, M. Zhang, and S. Ma · 2010
Cited alongside, same era.
Domain adaptation via pseudo in-domain data selection
A. Axelrod, X. He, and J. Gao · 2011
Cited alongside, same era.
Efficient and effective spam filtering and re-ranking for large web datasets
G. V. Cormack, M. D. Smucker, and C. L. A. Clarke · 2011
Cited alongside, same era.
Diagnostic evaluation of information retrieval models
H. Fang, T. Tao, and C. Zhai · 2011
Cited alongside, same era.
Bagging gradient-boosted trees for high precision, low variance ranking models
Y. Ganjisaffar, R. Caruana, and C. V. Lopes · 2011
Cited alongside, same era.
Information Retrieval Evaluation
D. Harman · 2011
Cited alongside, same era.
Automatic management of partitioned, replicated search services
F. Leibert, J. Mannix, J. Lin, and B. Hamadani · 2011
Cited alongside, same era.
Later among the works it cites.
PISA: Performant indexes and search for academia
A. Mallia, M. Siedlaczek, J. Mackenzie, and T. Suel · 2019
Later among the works it cites.
B. Mitra, C. Rosset, D. Hawking, N. Craswell, F. Diaz, and E. Yilmaz · 2019
Later among the works it cites.
An anatomy for neural search engines
T. A. Nakamura, P. H. Calais, D. de Castro Reis, and A. P. Lemos · 2019
Later among the works it cites.
R. Nogueira and K. Cho · 2019
Later among the works it cites.
From doc2query to docTTTTTquery, 2019
R. Nogueira and J. Lin · 2019
Later among the works it cites.
Investigating the successes and failures of BERT for passage re-ranking
H. Padigela, H. Zamani, and W. B. Croft · 2019
Later among the works it cites.
TF-Ranking: Scalable TensorFlow library for learning-to-rank
R. K. Pasumarthi, S. Bruch, X. Wang, C. Li, M. Bendersky, M. Najork, J. Pfeifer, N. Golbandi, R. Anil, and S. Wolf · 2019
Later among the works it cites.
PyTorch: An imperative style, high-performance deep learning library
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Köpf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala · 2019
Later among the works it cites.
Language models as knowledge bases?
F. Petroni, T. Rocktäschel, S. Riedel, P. Lewis, A. Bakhtin, Y. Wu, and A. Miller · 2019
Later among the works it cites.
How multilingual is multilingual BERT?
T. Pires, E. Schlinger, and D. Garrette · 2019
Later among the works it cites.
Zero-shot text classification with generative language models
R. Puri and B. Catanzaro · 2019
Later among the works it cites.
Understanding the behaviors of BERT in ranking
Y. Qiao, C. Xiong, Z. Liu, and Z. Liu · 2019
Later among the works it cites.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
N. Reimers and I. Gurevych · 2019
Later among the works it cites.
DistilBERT, a distilled version of BERT: Smaller, faster, cheaper and lighter
V. Sanh, L. Debut, J. Chaumond, and T. Wolf · 2019
Later among the works it cites.
Cross-lingual relevance transfer for document retrieval
P. Shi and J. Lin · 2019
Later among the works it cites.
Megatron-LM: Training multi-billion parameter language models using model parallelism
M. Shoeybi, M. Patwary, R. Puri, P. LeGresley, J. Casper, and B. Catanzaro · 2019
Later among the works it cites.
On extractive and abstractive neural document summarization with transformer language models
S. Subramanian, R. Li, J. Pilault, and C. Pal · 2019
Later among the works it cites.
Patient knowledge distillation for BERT model compression
S. Sun, Y. Cheng, Z. Gan, and J. Liu · 2019
Later among the works it cites.
Distilling task-specific knowledge from BERT into simple neural networks
R. Tang, Y. Lu, L. Liu, L. Mou, O. Vechtomova, and J. Lin · 2019
Later among the works it cites.
BERT rediscovers the classical NLP pipeline
I. Tenney, D. Das, and E. Pavlick · 2019
Later among the works it cites.
Micro- and macro-optimizations of SaaT
A. Trotman and M. Crane · 2019
Later among the works it cites.
Well-read students learn better: On the importance of pre-training compact models
I. Turc, M.-W. Chang, K. Lee, and K. Toutanova · 2019
Later among the works it cites.
Analyzing multi-head self-attention: Specialized heads do the heavy lifting, the rest can be pruned
E. Voita, D. Talbot, F. Moiseev, R. Sennrich, and I. Titov · 2019
Later among the works it cites.
Beto, bentz, becas: The surprising cross-lingual effectiveness of BERT
S. Wu and M. Dredze · 2019
Later among the works it cites.
IDST at TREC 2019 deep learning track: Deep cascade ranking with generation-based document expansion and pre-trained language modeling
M. Yan, C. Li, C. Wu, B. Bi, W. Wang, J. Xia, and L. Si · 2019
Later among the works it cites.
Aligning cross-lingual entities with multi-aspect information
H.-W. Yang, Y. Zou, P. Shi, W. Lu, J. Lin, and X. Sun · 2019
Later among the works it cites.
Reproducing and generalizing semantic term matching in axiomatic information retrieval
P. Yang and J. Lin · 2019
Later among the works it cites.
Critically examining the “neural hype”: Weak baselines and the additivity of effectiveness gains from neural ranking models
W. Yang, K. Lu, P. Yang, and J. Lin · 2019
Later among the works it cites.
End-to-end open-domain question answering with BERTserini
W. Yang, Y. Xie, A. Lin, X. Li, L. Tan, K. Xiong, M. Li, and J. Lin · 2019
Later among the works it cites.
XLNet: Generalized autoregressive pretraining for language understanding
Z. Yang, Z. Dai, Y. Yang, J. Carbonell, R. Salakhutdinov, and Q. V. Le · 2019
Later among the works it cites.
Dialog System Technology Challenge 7
K. Yoshino, C. Hori, J. Perez, L. F. D’Haro, L. Polymenakos, C. Gunasekara, W. S. Lasecki, J. K. Kummerfeld, M. Galley, C. Brockett, J. Gao, B. Dolan, X. Gao, H. Alamari, T. K. Marks, D. Parikh, and D. Batra · 2019
Later among the works it cites.
Simple techniques for cross-collection relevance feedback
R. Yu, Y. Xie, and J. Lin · 2019
Later among the works it cites.
HIBERT: Document level pre-training of hierarchical bidirectional transformers for document summarization
X. Zhang, F. Wei, and M. Zhou · 2019
Later among the works it cites.
Transformer-XH: Multi-evidence reasoning with extra hop attention
C. Zhao, C. Xiong, C. Rosset, X. Song, P. Bennett, and S. Tiwary · 2019
Later among the works it cites.
Structural language models of code
U. Alon, R. Sadaka, O. Levy, and E. Yahav · 2020
Closest in time.
On the cross-lingual transferability of monolingual representations
M. Artetxe, S. Ruder, and D. Yogatama · 2020
Closest in time.
SparTerm: Learning term-based sparse representation for fast text retrieval
Y. Bai, X. Li, G. Wang, C. Zhang, L. Shang, J. Xu, Z. Wang, F. Wang, and Q. Liu · 2020
Closest in time.
Scalable attentive sentence-pair modeling via distilled sentence embedding
O. Barkan, N. Razin, I. Malkiel, O. Katz, A. Caciularu, and N. Koenigstein · 2020
Closest in time.
Longformer: The long-document transformer
I. Beltagy, M. E. Peters, and A. Cohan · 2020
Closest in time.
RRF102: Meeting the TREC-COVID challenge with a 100+ runs ensemble
M. Bendersky, H. Zhuang, J. Ma, S. Han, K. Hall, and R. McDonald · 2020
Closest in time.
MarkedBERT: Integrating traditional IR cues in pre-trained language models for passage retrieval
L. Boualili, J. G. Moreno, and M. Boughanem · 2020
Closest in time.
Keyphrase generation for scientific document retrieval
F. Boudin, Y. Gallina, and A. Aizawa · 2020
Closest in time.
Language models are few-shot learners
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei · 2020
Closest in time.
Diagnosing BERT with retrieval heuristics
A. Câmara and C. Hauff · 2020
Closest in time.
Pre-training tasks for embedding-based large-scale retrieval
W.-C. Chang, F. X. Yu, Y.-W. Chang, Y. Yang, and S. Kumar · 2020
Closest in time.
Open-domain question answering
D. Chen and W.-t. Yih · 2020
Closest in time.
DiPair: Fast and accurate distillation for trillion-scale text matching and pair modeling
J. Chen, L. Yang, K. Raman, M. Bendersky, J.-J. Yeh, Y. Zhou, M. Najork, D. Cai, and E. Emadzadeh · 2020
Closest in time.
ELECTRA: Pre-training text encoders as discriminators rather than generators
K. Clark, M.-T. Luong, Q. V. Le, and C. D. Manning · 2020
Closest in time.
Overview of the TREC 2019 deep learning track
N. Craswell, B. Mitra, E. Yilmaz, D. Campos, and E. M. Voorhees · 2020
Closest in time.
Overview of the TREC 2020 deep learning track
N. Craswell, B. Mitra, E. Yilmaz, and D. Campos · 2020
Closest in time.
Context-aware document term weighting for ad-hoc search
Z. Dai and J. Callan · 2020
Closest in time.
Beyond [CLS] through ranking by generation
C. dos Santos, X. Ma, R. Nallapati, Z. Huang, and B. Xiang · 2020
Closest in time.
What BERT is not: Lessons from a new suite of psycholinguistic diagnostics for language models
A. Ettinger · 2020
Closest in time.
CodeBERT: A pre-trained model for programming and natural languages
Z. Feng, D. Guo, D. Tang, N. Duan, X. Feng, M. Gong, L. Shou, B. Qin, T. Liu, D. Jiang, and M. Zhou · 2020
Closest in time.
Modularized transfomer-based ranking framework
L. Gao, Z. Dai, and J. Callan · 2020
Closest in time.
Understanding BERT rankers under distillation
L. Gao, Z. Dai, and J. Callan · 2020
Closest in time.
TANDA: Transfer and adapt pre-trained transformer models for answer sentence selection
S. Garg, T. Vu, and A. Moschitti · 2020
Closest in time.
Domain-specific language model pretraining for biomedical natural language processing
Y. Gu, R. Tinn, H. Cheng, M. Lucas, N. Usuyama, X. Liu, T. Naumann, J. Gao, and H. Poon · 2020
Closest in time.
Don’t stop pretraining: Adapt language models to domains and tasks
S. Gururangan, A. Marasović, S. Swayamdipta, K. Lo, I. Beltagy, D. Downey, and N. A. Smith · 2020
Closest in time.
REALM: Retrieval-augmented language model pre-training
K. Guu, K. Lee, Z. Tung, P. Pasupat, and M.-W. Chang · 2020
Closest in time.
DeBERTa: Decoding-enhanced BERT with disentangled attention
P. He, X. Liu, J. Gao, and W. Chen · 2020
Closest in time.
The unstoppable rise of computational linguistics in deep learning
J. Henderson · 2020
Closest in time.
Improving efficient neural ranking models with cross-architecture knowledge distillation
S. Hofstätter, S. Althammer, M. Schröder, M. Sertkan, and A. Hanbury · 2020
Closest in time.
Local self-attention over long text for efficient document retrieval
S. Hofstätter, H. Zamani, B. Mitra, N. Craswell, and A. Hanbury · 2020
Closest in time.
Interpretable & time-budget-constrained contextualization for re-ranking
S. Hofstätter, M. Zlabinger, and A. Hanbury · 2020
Closest in time.
Embedding-based retrieval in Facebook search
J.-T. Huang, A. Sharma, S. Sun, L. Xia, D. Zhang, P. Pronin, J. Padmanabhan, G. Ottaviano, and L. Yang · 2020
Closest in time.
Poly-encoders: Architectures and pre-training strategies for fast and accurate multi-sentence scoring
S. Humeau, K. Shuster, M.-A. Lachaux, and J. Weston · 2020
Closest in time.
Leveraging passage retrieval with generative models for open domain question answering
G. Izacard and E. Grave · 2020
Closest in time.
A memory efficient baseline for open domain question answering
G. Izacard, F. Petroni, L. Hosseini, N. D. Cao, S. Riedel, and E. Grave · 2020
Closest in time.
Long document ranking with query-directed sparse transformer
J.-Y. Jiang, C. Xiong, C.-J. Lee, and W. Wang · 2020
Closest in time.
Which BM25 do you mean? A large-scale reproducibility study of scoring variants
C. Kamphuis, A. de Vries, L. Boytsov, and J. Lin · 2020
Closest in time.
Scaling laws for neural language models
J. Kaplan, S. McCandlish, T. Henighan, T. B. Brown, B. Chess, R. Child, S. Gray, A. Radford, J. Wu, and D. Amodei · 2020
Closest in time.
Dense passage retrieval for open-domain question answering
V. Karpukhin, B. Oğuz, S. Min, P. Lewis, L. Wu, S. Edunov, D. Chen, and W.-t. Yih · 2020
Closest in time.
ColBERT: Efficient and effective passage search via contextualized late interaction over BERT
O. Khattab and M. Zaharia · 2020
Closest in time.
Reformer: The efficient transformer
N. Kitaev, Ł. Kaiser, and A. Levskaya · 2020
Closest in time.
ALBERT: A lite BERT for self-supervised learning of language representations
Z. Lan, M. Chen, S. Goodman, K. Gimpel, P. Sharma, and R. Soricut · 2020
Closest in time.
Mixout: Effective regularization to finetune large-scale pretrained language models
C. Lee, K. Cho, and W. Kang · 2020
Closest in time.
Pre-training via paraphrasing
M. Lewis, M. Ghazvininejad, G. Ghosh, A. Aghajanyan, S. Wang, and L. Zettlemoyer · 2020
Closest in time.
Train large, then compress: Rethinking model size for efficient training and inference of transformers
Z. Li, E. Wallace, S. Shen, K. Lin, K. Keutzer, D. Klein, and J. E. Gonzalez · 2020
Closest in time.
Reproducibility is a process, not an achievement: The replicability of IR reproducibility experiments
J. Lin and Q. Zhang · 2020
Closest in time.
Supporting interoperability between open-source search engines with the Common Index File Format
J. Lin, J. Mackenzie, C. Kamphuis, C. Macdonald, A. Mallia, M. Siedlaczek, A. Trotman, and A. de Vries · 2020
Closest in time.
FastBERT: A self-distilling BERT with adaptive inference time
W. Liu, P. Zhou, Z. Wang, Z. Zhao, H. Deng, and Q. Ju · 2020
Closest in time.
TwinBERT: Distilling knowledge to twin-structured BERT models for efficient retrieval
W. Lu, J. Jiao, and R. Zhang · 2020
Closest in time.
Sparse, dense, and attentional representations for text retrieval
Y. Luan, J. Eisenstein, K. Toutanova, and M. Collins · 2020
Closest in time.
Efficient document re-ranking for transformers by precomputing term representations
S. MacAvaney, F. M. Nardini, R. Perego, N. Tonellotto, N. Goharian, and O. Frieder · 2020
Closest in time.
Expansion via prediction of importance with contextualization
S. MacAvaney, F. M. Nardini, R. Perego, N. Tonellotto, N. Goharian, and O. Frieder · 2020
Closest in time.
Training curricula for open domain answer re-ranking
S. MacAvaney, F. M. Nardini, R. Perego, N. Tonellotto, N. Goharian, and O. Frieder · 2020
Closest in time.
Teaching a new dog old tricks: Resurrecting multilingual retrieval using zero-shot learning
S. MacAvaney, L. Soldaini, and N. Goharian · 2020
Closest in time.
Efficiency implications of term weighting for passage retrieval
J. Mackenzie, Z. Dai, L. Gallagher, and J. Callan · 2020
Closest in time.
Efficient and robust approximate nearest neighbor search using hierarchical navigable small world graphs
Y. A. Malkov and D. A. Yashunin · 2020
Closest in time.
Reranking for efficient transformer-based answer selection
Y. Matsubara, T. Vu, and A. Moschitti · 2020
Closest in time.
Conversational AI: Dialogue Systems, Conversational Agents, and Chatbots
M. McTear · 2020
Closest in time.
Conformer-kernel with query term independence for document retrieval
B. Mitra, S. Hofstätter, H. Zamani, and N. Craswell · 2020
Closest in time.
Document ranking with a pretrained sequence-to-sequence model
R. Nogueira, Z. Jiang, R. Pradeep, and J. Lin · 2020
Closest in time.
Rethinking query expansion for BERT reranking
R. Padaki, Z. Dai, and J. Callan · 2020
Closest in time.
English intermediate-task training improves zero-shot cross-lingual transfer too
J. Phang, I. Calixto, P. M. Htut, Y. Pruksachatkun, H. Liu, C. Vania, K. Kann, and S. R. Bowman · 2020
Closest in time.
Exploring the limits of transfer learning with a unified text-to-text transformer
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu · 2020
Closest in time.
TREC-COVID: Rationale and structure of an information retrieval shared task for COVID-19
K. Roberts, T. Alam, S. Bedrick, D. Demner-Fushman, K. Lo, I. Soboroff, E. Voorhees, L. L. Wang, and W. R. Hersh · 2020
Closest in time.
A primer in BERTology: What we know about how BERT works
A. Rogers, O. Kovaleva, and A. Rumshisky · 2020
Closest in time.
Recipes for building an open-domain chatbot
S. Roller, E. Dinan, N. Goyal, D. Ju, M. Williamson, Y. Liu, J. Xu, M. Ott, K. Shuster, E. M. Smith, Y.-L. Boureau, and J. Weston · 2020
Closest in time.
Masked language model scoring
J. Salazar, D. Liang, T. Q. Nguyen, and K. Kirchhoff · 2020
Closest in time.
The right tool for the job: Matching model and instance complexities
R. Schwartz, G. Stanovsky, S. Swayamdipta, J. Dodge, and N. A. Smith · 2020
Closest in time.
Longformer for MS MARCO document re-ranking task
I. Sekulić, A. Soleimani, M. Aliannejadi, and F. Crestani · 2020
Closest in time.
BISON: BM25-weighted self-attention framework for multi-fields document search
X. Shan, C. Liu, Y. Xia, Q. Chen, Y. Zhang, A. Luo, and Y. Luo · 2020
Closest in time.
Cross-lingual training of neural models for document ranking
P. Shi, H. Bai, and J. Lin · 2020
Closest in time.
The Cascade Transformer: An application for efficient answer sentence selection
L. Soldaini and A. Moschitti · 2020
Closest in time.
Distilling knowledge for fast retrieval-based chat-bots
A. V. Tahami, K. Ghajar, and A. Shakery · 2020
Closest in time.
Efficient transformers: A survey
Y. Tay, M. Dehghani, D. Bahri, and D. Metzler · 2020
Closest in time.
Approximate nearest neighbor search and lightweight dense vector reranking in multi-stage retrieval architectures
Z. Tu, W. Yang, Z. Fu, Y. Xie, L. Tan, K. Xiong, M. Li, and J. Lin · 2020
Closest in time.
TREC-COVID: Constructing a pandemic information retrieval test collection
E. M. Voorhees, T. Alam, S. Bedrick, D. Demner-Fushman, W. R. Hersh, K. Lo, K. Roberts, I. Soboroff, and L. L. Wang · 2020
Closest in time.
On-the-fly information retrieval augmentation for language models
H. Wang and D. McAllester · 2020
Closest in time.
Transformers: State-of-the-art natural language processing
T. Wolf, L. Debut, V. Sanh, J. Chaumond, C. Delangue, A. Moi, P. Cistac, T. Rault, R. Louf, M. Funtowicz, J. Davison, S. Shleifer, P. von Platen, C. Ma, Y. Jernite, J. Plu, C. Xu, T. Le Scao, S. Gugger, M. Drame, Q. Lhoest, and A. Rush · 2020
Closest in time.
Leveraging passage-level cumulative gain for document ranking
Z. Wu, J. Mao, Y. Liu, J. Zhan, Y. Zheng, M. Zhang, and S. Ma · 2020
Closest in time.
Distant supervision for multi-stage fine-tuning in retrieval-based question answering
Y. Xie, W. Yang, L. Tan, K. Xiong, N. J. Yuan, B. Huai, M. Li, and J. Lin · 2020
Closest in time.
DeeBERT: Dynamic early exiting for accelerating BERT inference
J. Xin, R. Tang, J. Lee, Y. Yu, and J. Lin · 2020
Closest in time.
Deep learning for matching in search and recommendation
J. Xu, X. He, and H. Li · 2020
Closest in time.
Beyond 512 tokens: Siamese multi-depth transformer-based hierarchical encoder for document matching
L. Yang, M. Zhang, C. Li, M. Bendersky, and M. Najork · 2020
Closest in time.
Is retriever merely an approximator of reader?
S. Yang and M. Seo · 2020
Closest in time.
Flexible IR pipelines with Capreolus
A. Yates, K. M. Jose, X. Zhang, and J. Lin · 2020
Closest in time.
Playing the lottery with rewards and multiple languages: Lottery tickets in RL and NLP
H. Yu, S. Edunov, Y. Tian, and A. S. Morcos · 2020
Closest in time.
A study of neural matching models for cross-lingual IR
P. Yu and J. Allan · 2020
Closest in time.
Big Bird: Transformers for longer sequences
M. Zaheer, G. Guruganesh, A. Dubey, J. Ainslie, C. Alberti, S. Ontanon, P. Pham, A. Ravula, Q. Wang, L. Yang, and A. Ahmed · 2020
Closest in time.
Selective weak supervision for neural information retrieval
K. Zhang, C. Xiong, Z. Liu, and Z. Liu · 2020
Closest in time.
BERT-QE: Contextualized Query Expansion for Document Re-ranking
Z. Zheng, K. Hui, B. He, X. Han, L. Sun, and A. Yates · 2020
Closest in time.
XOR QA: Cross-lingual open-retrieval question answering
A. Asai, J. Kasai, J. Clark, K. Lee, E. Choi, and H. Hajishirzi · 2021
Closest in time.
Exploring classic and neural lexical translation models for information retrieval: Interpretability, effectiveness, and efficiency benefits
L. Boytsov and Z. Kolter · 2021
Closest in time.
Simplified TinyBERT: Knowledge distillation for document retrieval
X. Chen, B. He, K. Hui, L. Sun, and Y. Sun · 2021
Closest in time.
MS MARCO: Benchmarking ranking models in the large-data regime
N. Craswell, B. Mitra, D. Campos, E. Yilmaz, and J. Lin · 2021
Closest in time.
Self-training improves pre-training for natural language understanding
J. Du, E. Grave, B. Gunel, V. Chaudhary, O. Celebi, M. Auli, V. Stoyanov, and A. Conneau · 2021
Closest in time.
Amnesic probing: Behavioral explanation with amnesic counterfactuals
Y. Elazar, S. Ravfogel, A. Jacovi, and Y. Goldberg · 2021
Closest in time.
SPLADE: Sparse lexical and expansion model for first stage ranking
T. Formal, B. Piwowarski, and S. Clinchant · 2021
Closest in time.
A white box analysis of ColBERT
T. Formal, B. Piwowarski, and S. Clinchant · 2021
Closest in time.
COIL: Revisit exact lexical match in information retrieval with contextualized inverted list
L. Gao, Z. Dai, and J. Callan · 2021
Closest in time.
Complementing lexical retrieval with semantic residual embedding
L. Gao, Z. Dai, T. Chen, Z. Fan, B. V. Durme, and J. Callan · 2021
Closest in time.
The simplest thing that can possibly work: (pseudo-)relevance feedback via text classification
X. Han, Y. Liu, and J. Lin · 2021
Closest in time.
BERTese: Learning to speak to BERT
A. Haviv, J. Berant, and A. Globerson · 2021
Closest in time.
Efficiently teaching an effective dense retriever with balanced topic aware sampling
S. Hofstätter, S.-C. Lin, J.-H. Yang, J. Lin, and A. Hanbury · 2021
Closest in time.
Answer generation for retrieval-based question answering systems
C.-C. Hsu, E. Lind, L. Soldaini, and A. Moschitti · 2021
Closest in time.
Composite code sparse autoencoders for first stage retrieval
C. Lassance, T. Formal, and S. Clinchant · 2021
Closest in time.
How many data points is a prompt worth?
T. Le Scao and A. Rush · 2021
Closest in time.
J. Lin and X. Ma · 2021
Closest in time.
Pyserini: A Python toolkit for reproducible information retrieval research with sparse and dense representations
J. Lin, X. Ma, S.-C. Lin, J.-H. Yang, R. Pradeep, and R. Nogueira · 2021
Closest in time.
In-batch negatives for knowledge distillation with tightly-coupled teachers for dense retrieval
S.-C. Lin, J.-H. Yang, and J. Lin · 2021
Closest in time.
H. Liu, Z. Dai, D. R. So, and Q. V. Le · 2021
Closest in time.
Studying catastrophic forgetting in neural ranking models
J. Lovón-Melgarejo, L. Soulier, K. Pinel-Sauvagnat, and L. Tamine · 2021
Closest in time.
Less is more: Pre-training a strong siamese encoder using a weak decoder
S. Lu, C. Xiong, D. He, G. Ke, W. Malik, Z. Dou, P. Bennett, T. Liu, and A. Overwijk · 2021
Closest in time.
Sparse, dense, and attentional representations for text retrieval
Y. Luan, J. Eisenstein, K. Toutanova, and M. Collins · 2021
Closest in time.
PROP: Pre-training with representative words prediction for ad-hoc retrieval
X. Ma, J. Guo, R. Zhang, Y. Fan, X. Ji, and X. Cheng · 2021
Closest in time.
How deep is your learning: The DL-HARD annotated deep learning dataset
I. Mackie, J. Dalton, and A. Yates · 2021
Closest in time.
Learning passage impacts for inverted indexes
A. Mallia, O. Khattab, T. Suel, and N. Tonellotto · 2021
Closest in time.
A systematic evaluation of transfer learning and pseudo-labeling with BERT-based ranking models
I. Mokrii, L. Boytsov, and P. Braslavski · 2021
Closest in time.
CEQE: Contextualized embeddings for query expansion
S. Naseri, J. Dalton, A. Yates, and J. Allan · 2021
Closest in time.
RocketQA: An optimized training approach to dense passage retrieval for open-domain question answering
Y. Qu, Y. Ding, J. Liu, K. Liu, R. Ren, W. X. Zhao, D. Dong, H. Wu, and H. Wang · 2021
Closest in time.
Exploiting cloze-questions for few-shot text classification and natural language inference
T. Schick and H. Schütze · 2021
Closest in time.
Are pretrained convolutions better than pretrained transformers?
Y. Tay, M. Dehghani, J. P. Gupta, V. Aribandi, D. Bahri, Z. Qin, and D. Metzler · 2021
Closest in time.
BEIR: A heterogenous benchmark for zero-shot evaluation of information retrieval models
N. Thakur, N. Reimers, A. Rücklé, A. Srivastava, and I. Gurevych · 2021
Closest in time.
Approximate nearest neighbor negative contrastive learning for dense text retrieval
L. Xiong, C. Xiong, Y. Li, K.-F. Tang, J. Liu, P. Bennett, J. Ahmed, and A. Overwijk · 2021
Closest in time.
Efficient passage retrieval with hashing for open-domain question answering
I. Yamada, A. Asai, and H. Hajishirzi · 2021
Closest in time.
A unified pretraining framework for passage ranking and expansion
M. Yan, C. Li, B. Bi, W. Wang, and S. Huang · 2021
Closest in time.
PGT: Pseudo relevance feedback using a graph-based transformer
H. Yu, Z. Dai, and J. Callan · 2021
Closest in time.
Optimizing dense retrieval model training with hard negatives
J. Zhan, J. Mao, Y. Liu, J. Guo, M. Zhang, and S. Ma · 2021
Closest in time.
Comparing score aggregation approaches for document retrieval with pretrained transformers
X. Zhang, A. Yates, and J. Lin · 2021
Closest in time.
SPARTA: Efficient open-domain question answering via sparse transformer matching retrieval
T. Zhao, X. Lu, and K. Lee · 2021
Closest in time.
Pre-trained language model based ranking in Baidu search
L. Zou, S. Zhang, H. Cai, D. Ma, S. Cheng, D. Shi, Z. Zhu, W. Su, S. Wang, Z. Cheng, and D. Yin · 2021
Closest in time.
PEGASUS: Pre-training with extracted gap-sentences for abstractive summarization
J. Zhang, Y. Zhao, M. Saleh, and P. J. Liu · 2032
Closest in time.