Fetching the paper…
Reading the bibliography…
Conventional document retrieval techniques are mainly based on the index-retrieve paradigm.
On Relevance Weights with Little Relevance Information. In SIGIR
Stephen E. Robertson and Steve Walker. 1997 · 1997
Earlier work this paper cites.
Constrained K-Means Clustering
Paul S. Bradley, Kristin P. Bennett, and Ayhan Demiriz. 2000 · 2000
Earlier work this paper cites.
Document Language Models, Query Models, and Risk Minimization for Information Retrieval. In SIGIR
John Lafferty and ChengXiang Zhai. 2001 · 2001
Earlier work this paper cites.
The Probabilistic Relevance Framework: BM25 and Beyond
Stephen E. Robertson and Hugo Zaragoza. 2009 · 2009
Earlier work this paper cites.
Sinkhorn Distances: Lightspeed Computation of Optimal Transport. In NIPS
Marco Cuturi. 2013 · 2013
Earlier work this paper cites.
Learning to Reweight Terms with Distributed Representations. In SIGIR
Guoqing Zheng and Jamie Callan. 2015 · 2015
Earlier work this paper cites.
MS MARCO: A Human Generated MAchine Reading COmprehension Dataset
Daniel Fernando Campos, Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, Li Deng, and Bhaskar Mitra. 2016 · 2016
Earlier work this paper cites.
A Deep Relevance Matching Model for Ad-hoc Retrieval. In CIKM
Jiafeng Guo, Yixing Fan, Qingyao Ai, and W. Bruce Croft. 2016 · 2016
Earlier work this paper cites.
Neural Ranking Models with Weak Supervision. In SIGIR
Mostafa Dehghani, Hamed Zamani, Aliaksei Severyn, Jaap Kamps, and W. Bruce Croft. 2017 · 2017
Earlier work this paper cites.
Billion-Scale Similarity Search with GPUs
Jeff Johnson, Matthijs Douze, and Herve Jegou. 2017 · 2017
Earlier work this paper cites.
Discrete Variational Autoencoders. In ICLR
Jason Tyler Rolfe. 2017 · 2017
Earlier work this paper cites.
Neural Discrete Representation Learning. In NIPS
Aäron van den Oord, Oriol Vinyals, and Koray Kavukcuoglu. 2017 · 2017
Earlier work this paper cites.
End-to-End Retrieval in Continuous Space
Daniel Gillick, Alessandro Presta, and Gaurav Singh Tomar. 2018 · 2018
Earlier work this paper cites.
Unsupervised Discrete Sentence Representation Learning for Interpretable Neural Dialog Generation. In ACL
Tiancheng Zhao, Kyusong Lee, and Maxine Eskénazi. 2018 · 2018
Earlier work this paper cites.
Fast Decoding in Sequence Models using Discrete Latent Variables. In ICML
Łukasz Kaiser, Aurko Roy, Ashish Vaswani, Niki Parmar, Samy Bengio, Jakob Uszkoreit, and Noam M. Shazeer. 2018 · 2018
Earlier work this paper cites.
From Doc2query to DocTTTTTquery
David R. Cheriton. 2019 · 2019
Earlier work this paper cites.
Context-Aware Sentence/Passage Term Importance Estimation For First Stage Retrieval
Zhuyun Dai and Jamie Callan. 2019 · 2019
Earlier work this paper cites.
Natural Questions: A Benchmark for Question Answering Research
Tom Kwiatkowski, Jennimaria Palomaki, Olivia Redfield, Michael Collins, Ankur P. Parikh, Chris Alberti, Danielle Epstein, Illia Polosukhin, Jacob Devlin, Kenton Lee, Kristina Toutanova, Llion Jones, Matthew Kelcey, Ming-Wei Chang, Andrew M. Dai, Jakob Uszkoreit, Quoc V. Le, and Slav Petrov. 2019 · 2019
Earlier work this paper cites.
Document Expansion by Query Prediction
Rodrigo Nogueira, Wei Yang, Jimmy Lin, and Kyunghyun Cho. 2019 · 2019
Cited alongside, same era.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks. In EMNLP
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Language Models are Few-Shot Learners. In NeurIPS
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, T. J. Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeff Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Cited alongside, same era.
Unsupervised Learning of Visual Features by Contrasting Cluster Assignments
Mathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal, Piotr Bojanowski, and Armand Joulin. 2020 · 2020
Cited alongside, same era.
BEIR: A Heterogenous Benchmark for Zero-shot Evaluation of Information Retrieval Models. In NeurIPS Datasets and Benchmarks Track (Round 2)
Nandan Thakur, Nils Reimers, Andreas Ruckl’e, Abhishek Srivastava, and Iryna Gurevych. 2021 · 2021
Later among the works it cites.
Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval. In ICLR
Lee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang, Jialin Liu, Paul Bennett, Junaid Ahmed, and Arnold Overwijk. 2021 · 2021
Later among the works it cites.
Autoregressive Search Engines: Generating Substrings as Document Identifiers. In NeurIPS
Michele Bevilacqua, Giuseppe Ottaviano, Patrick Lewis, Wen tau Yih, Sebastian Riedel, and Fabio Petroni. 2022 · 2022
Later among the works it cites.
CorpusBrain: Pre-train a Generative Retrieval Model for Knowledge-Intensive Language Tasks. In CIKM
Jiangui Chen, Ruqing Zhang, Jiafeng Guo, Y. Liu, Yixing Fan, and Xueqi Cheng. 2022 · 2022
Later among the works it cites.
Unsupervised Dense Information Retrieval with Contrastive Learning. In TMLR
Gautier Izacard, Mathilde Caron, Lucas Hosseini, Sebastian Riedel, Piotr Bojanowski, Armand Joulin, and Edouard Grave. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhuyun Dai and Jamie Callan. 2020 · 2020
Cited alongside, same era.
Discrete Latent Variable Representations for Low-Resource Text Classification. In ACL
Shuning Jin, Sam Wiseman, Karl Stratos, and Karen Livescu. 2020 · 2020
Cited alongside, same era.
Dense Passage Retrieval for Open-Domain Question Answering. In EMNLP
Vladimir Karpukhin, Barlas Oğuz, Sewon Min, Patrick Lewis, Ledell Yu Wu, Sergey Edunov, Danqi Chen, and Wen tau Yih. 2020 · 2020
Cited alongside, same era.
ColBERT: Efficient and Effective Passage Search via Contextualized Late Interaction over BERT. In SIGIR
Omar Khattab and Matei Zaharia. 2020 · 2020
Cited alongside, same era.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam M. Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter Liu. 2020 · 2020
Cited alongside, same era.
Zero-Shot Text-to-Image Generation. In ICML
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2020 · 2020
Cited alongside, same era.
Autoregressive Entity Retrieval. In ICLR
Nicola De Cao, Gautier Izacard, Sebastian Riedel, and Fabio Petroni. 2021 · 2021
Cited alongside, same era.
Efficiently Teaching an Effective Dense Retriever with Balanced Topic Aware Sampling. In SIGIR
Sebastian Hofstätter, Sheng-Chieh Lin, Jheng-Hong Yang, Jimmy J. Lin, and Allan Hanbury. 2021 · 2021
Cited alongside, same era.
Later among the works it cites.
Contextualized Generative Retrieval
Hyunji Lee, Jaeyoung Kim, Hoyeon Chang, Hanseok Oh, Sohee Yang, Vladimir Karpukhin, Yi Lu, and Minjoon Seo. 2022 · 2022
Later among the works it cites.
Pretrained Transformers for Text Ranking: BERT and Beyond
Jimmy Lin, Rodrigo Nogueira, and Andrew Yates. 2022 · 2022
Later among the works it cites.
Yuxiang Lu, Yiding Liu, Jiaxiang Liu, Yunsheng Shi, Zhengjie Huang, Shi Feng, Yu Sun, Hao Tian, Hua Wu, Shuaiqiang Wang, Dawei Yin, and Haifeng Wang. 2022 · 2022
Later among the works it cites.
DSI++: Updating Transformer Memory with New Documents
Sanket Vaibhav Mehta, Jai Gupta, Yi Tay, Mostafa Dehghani, Vinh Quang Tran, Jinfeng Rao, Marc-Alexander Najork, Emma Strubell, and Donald Metzler. 2022 · 2022
Later among the works it cites.
Text and Code Embeddings by Contrastive Pre-Training
Arvind Neelakantan, Tao Xu, Raul Puri, Alec Radford, Jesse Michael Han, Jerry Tworek, Qiming Yuan, Nikolas A. Tezak, Jong Wook Kim, Chris Hallacy, Johannes Heidecke, Pranav Shyam, Boris Power, Tyna Eloundou Nekoul, Girish Sastry, Gretchen Krueger, David P. Schnurr, Felipe Petroski Such, Kenny Sai-Kin Hsu, Madeleine Thompson, Tabarak Khan, Toki Sherbakov, Joanne Jang, Peter Welinder, and Lilian Weng. 2022 · 2022
Later among the works it cites.
Sentence-T5: Scalable Sentence Encoders from Pre-trained Text-to-Text Models. In Findings of ACL
Jianmo Ni, Gustavo Hern’andez ’Abrego, Noah Constant, Ji Ma, Keith B. Hall, Daniel Matthew Cer, and Yinfei Yang. 2022 · 2022
Later among the works it cites.
ColBERTv2: Effective and Efficient Retrieval via Lightweight Late Interaction. In NAACL
Keshav Santhanam, Omar Khattab, Jon Saad-Falcon, Christopher Potts, and Matei Zaharia. 2022 · 2022
Later among the works it cites.
Transformer Memory as a Differentiable Search Index. In NeurIPS
Yi Tay, Vinh Quang Tran, Mostafa Dehghani, Jianmo Ni, Dara Bahri, Harsh Mehta, Zhen Qin, Kai Hui, Zhe Zhao, Jai Gupta, Tal Schuster, William W. Cohen, and Donald Metzler. 2022 · 2022
Later among the works it cites.
Learning Semantic Textual Similarity via Topic-informed Discrete Latent Variables
Erxin Yu, Lan Du, Yuan Jin, Zhepei Wei, and Yi Chang. 2022 · 2022
Later among the works it cites.
Learning Discrete Representations via Constrained Clustering for Effective and Efficient Dense Retrieval. In WSDM
Jingtao Zhan, Jiaxin Mao, Yiqun Liu, Jiafeng Guo, M. Zhang, and Shaoping Ma. 2022 · 2022
Later among the works it cites.
Ultron: An Ultimate Retriever on Corpus with a Model-based Indexer
Yujia Zhou, Jing Yao, Zhicheng Dou, Ledell Yu Wu, Peitian Zhang, and Ji rong Wen. 2022 · 2022
Later among the works it cites.
Shengyao Zhuang, Houxing Ren, Linjun Shou, Jian Pei, Ming Gong, G. Zuccon, and Daxin Jiang. 2022 · 2022
Later among the works it cites.