Fetching the paper…
Reading the bibliography…
Embedding-based retrieval methods construct vector indices to search for document representations that are most similar to the query representations.
Multidimensional binary search trees used for associative searching
Jon Louis Bentley · 1975
Earlier work this paper cites.
A k-means clustering algorithm
John A Hartigan and Manchek A Wong · 1979
Earlier work this paper cites.
The cluster hypothesis revisited
Ellen M. Voorhees · 1985
Earlier work this paper cites.
On relevance weights with little relevance information
Stephen E. Robertson and Steve Walker · 1997
Earlier work this paper cites.
Document language models, query models, and risk minimization for information retrieval
John D. Lafferty and ChengXiang Zhai · 2001
Earlier work this paper cites.
Locality-sensitive hashing scheme based on p-stable distributions
Mayur Datar, Nicole Immorlica, Piotr Indyk, and Vahab S. Mirrokni · 2004
Earlier work this paper cites.
Cluster-based retrieval using language models
Xiaoyong Liu and W. Bruce Croft · 2004
Earlier work this paper cites.
The probabilistic relevance framework: BM25 and beyond
Stephen E. Robertson and Hugo Zaragoza · 2009
Earlier work this paper cites.
Scikit-learn: Machine learning in python
Fabian Pedregosa, Gaël Varoquaux, Alexandre Gramfort, Vincent Michel, Bertrand Thirion, Olivier Grisel, Mathieu Blondel, Peter Prettenhofer, Ron Weiss, Vincent Dubourg, Jake VanderPlas, Alexandre Passos, David Cournapeau, Matthieu Brucher, Matthieu Perrot, and Edouard Duchesnay · 2011
Earlier work this paper cites.
Learning deep structured semantic models for web search using clickthrough data
Po-Sen Huang, Xiaodong He, Jianfeng Gao, Li Deng, Alex Acero, and Larry P. Heck · 2013
Earlier work this paper cites.
Stacked quantizers for compositional vector compression
Julieta Martinez, Holger H. Hoos, and James J. Little · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey E. Hinton, Oriol Vinyals, and Jeffrey Dean · 2015
Earlier work this paper cites.
A deep relevance matching model for ad-hoc retrieval
Jiafeng Guo, Yixing Fan, Qingyao Ai, and W. Bruce Croft · 2016
Earlier work this paper cites.
MS MARCO: A human generated machine reading comprehension dataset
Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, and Li Deng · 2016
Earlier work this paper cites.
A comparison between term-based and embedding-based methods for initial retrieval
Tonglei Guo, Jiafeng Guo, Yixing Fan, Yanyan Lan, Jun Xu, and Xueqi Cheng · 2018
Earlier work this paper cites.
Context-aware sentence/passage term importance estimation for first stage retrieval
Zhuyun Dai and Jamie Callan · 2019
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
Natural questions: a benchmark for question answering research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur P. Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, Kristina Toutanova, Llion Jones, Matthew Kelcey, Ming-Wei Chang, Andrew M. Dai, Jakob Uszkoreit, Quoc Le, and Slav Petrov · 2019
Earlier work this paper cites.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Cited alongside, same era.
From doc2query to doctttttquery
Rodrigo Nogueira, Jimmy Lin, and AI Epistemic · 2019
Cited alongside, same era.
Rand-nsg: Fast accurate billion-point nearest neighbor search on a single node
Suhas Jayaram Subramanya, Fnu Devvrit, Harsha Vardhan Simhadri, Ravishankar Krishnaswamy, and Rohan Kadekodi · 2019
Cited alongside, same era.
Simple applications of BERT for ad hoc document retrieval
Wei Yang, Haotian Zhang, and Jimmy Lin · 2019
Cited alongside, same era.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick S. H. Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih · 2020
Cited alongside, same era.
Matching-oriented embedding quantization for ad-hoc retrieval
Shitao Xiao, Zheng Liu, Yingxia Shao, Defu Lian, and Xing Xie · 2021
Later among the works it cites.
Approximate nearest neighbor negative contrastive learning for dense text retrieval
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul N. Bennett, Junaid Ahmed, and Arnold Overwijk · 2021
Later among the works it cites.
Joint learning of deep retrieval model and product quantization based embedding index
Han Zhang, Hongwei Shen, Yiming Qiu, Yunjiang Jiang, Songlin Wang, Sulong Xu, Yun Xiao, Bo Long, and Wen-Yun Yang · 2021
Later among the works it cites.
Scalable multi-grained cross-modal similarity query with interpretability
Mingdong Zhu, Derong Shen, Lixin Xu, and Xianfang Wang · 2021
Later among the works it cites.
Deep query likelihood model for information retrieval
Shengyao Zhuang, Hang Li, and Guido Zuccon · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Twinbert: Distilling knowledge to twin-structured compressed BERT models for large-scale retrieval
Wenhao Lu, Jian Jiao, and Ruofei Zhang · 2020
Cited alongside, same era.
Efficient and robust approximate nearest neighbor search using hierarchical navigable small world graphs
Yury A. Malkov and Dmitry A. Yashunin · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2020
Cited alongside, same era.
HM-ANN: efficient billion-point nearest neighbor search on heterogeneous memory
Jie Ren, Minjia Zhang, and Dong Li · 2020
Cited alongside, same era.
SPANN: highly-efficient billion-scale approximate nearest neighborhood search
Qi Chen, Bing Zhao, Haidong Wang, Mingqin Li, Chuanjie Liu, Zengzhong Li, Mao Yang, and Jingdong Wang · 2021
Cited alongside, same era.
Constructing an educational knowledge graph with concepts linked to wikipedia
Fu-Rong Dang, Jintao Tang, Kunyuan Pang, Ting Wang, Sha-Sha Li, and Xiao Li · 2021
Cited alongside, same era.
Onnx runtime
ONNX Runtime developers · 2021
Cited alongside, same era.
Michele Bevilacqua, Giuseppe Ottaviano, Patrick S. H. Lewis, Scott Yih, Sebastian Riedel, and Fabio Petroni · 2022
Later among the works it cites.
Disentangled graph recurrent network for document ranking
Qian Dong, Shuzi Niu, Tao Yuan, and Yucheng Li · 2022
Later among the works it cites.
Unsupervised corpus aware language model pre-training for dense passage retrieval
Luyu Gao and Jamie Callan · 2022
Later among the works it cites.
Heterogeneous memory enhanced graph reasoning network for cross-modal retrieval
Zhong Ji, Kexin Chen, Yuqing He, Yanwei Pang, and Xuelong Li · 2022
Later among the works it cites.
Deepwalk-aware graph convolutional networks
Taisong Jin, Huaqiang Dai, Liujuan Cao, Baochang Zhang, Feiyue Huang, Yue Gao, and Rongrong Ji · 2022
Later among the works it cites.
HET-GMP: A graph-based system approach to scaling large embedding model training
Xupeng Miao, Yining Shi, Hailin Zhang, Xin Zhang, Xiaonan Nie, Zhi Yang, and Bin Cui · 2022
Later among the works it cites.
Transformer memory as a differentiable search index
Yi Tay, Vinh Tran, Mostafa Dehghani, Jianmo Ni, Dara Bahri, Harsh Mehta, Zhen Qin, Kai Hui, Zhe Zhao, Jai Prakash Gupta, Tal Schuster, William W. Cohen, and Donald Metzler · 2022
Later among the works it cites.
A neural corpus indexer for document retrieval
Yujing Wang, Yingyan Hou, Haonan Wang, Ziming Miao, Shibin Wu, Qi Chen, Yuqing Xia, Chengmin Chi, Guoshuai Zhao, Zheng Liu, Xing Xie, Hao Sun, Weiwei Deng, Qi Zhang, and Mao Yang · 2022
Later among the works it cites.
Leveraging document-level and query-level passage cumulative gain for document ranking
Zhijing Wu, Yiqun Liu, Jiaxin Mao, Min Zhang, and Shaoping Ma · 2022
Later among the works it cites.
Distill-vq: Learning retrieval oriented vector quantization by distilling knowledge from dense embeddings
Shitao Xiao, Zheng Liu, Weihao Han, Jianjin Zhang, Defu Lian, Yeyun Gong, Qi Chen, Fan Yang, Hao Sun, Yingxia Shao, and Xing Xie · 2022
Later among the works it cites.
Adversarial retriever-ranker for dense text retrieval
Hang Zhang, Yeyun Gong, Yelong Shen, Jiancheng Lv, Nan Duan, and Weizhu Chen · 2022
Later among the works it cites.
A survey on non-autoregressive generation for neural machine translation and beyond
Yisheng Xiao, Lijun Wu, Junliang Guo, Juntao Li, Min Zhang, Tao Qin, and Tie-Yan Liu · 2023
Closest in time.