HuggingFace’s Transformers: State-of-the-art Natural Language Processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, et al · 1910
Earlier work this paper cites.
Introduction to information retrieval
Christopher D Manning, Hinrich Schütze, and Prabhakar Raghavan. 2008 · 2008
Earlier work this paper cites.
Simplified TinyBERT: Knowledge Distillation for Document Retrieval
Xuanang Chen, Ben He, Kai Hui, Le Sun, and Yingfei Sun. 2020 · 2009
Earlier work this paper cites.
From ranknet to lambdarank to lambdamart: An overview
Christopher JC Burges. 2010 · 2010
Earlier work this paper cites.
Adam: A method for stochastic optimization
Original
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Distilling the knowledge in a neural network
Original
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean. 2015 · 2015
Earlier work this paper cites.
MS MARCO : A Human Generated MAchine Reading COmprehension Dataset. In Proc. of NIPS
Payal Bajaj, Daniel Campos, Nick Craswell, Li Deng, Jianfeng Gao, Xiaodong Liu, Rangan Majumder, Andrew Mcnamara, Bhaskar Mitra, and Tri Nguyen. 2016 · 2016
Earlier work this paper cites.
Learning to learn from weak supervision by full supervision
Mostafa Dehghani, Aliaksei Severyn, Sascha Rothe, and Jaap Kamps. 2017 · 2017
Earlier work this paper cites.
Automatic differentiation in PyTorch. In Proc. of NIPS-W
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
Earlier work this paper cites.
End-to-End Neural Ad-hoc Ranking with Kernel Pooling. In Proc. of SIGIR
Chenyan Xiong, Zhuyun Dai, Jamie Callan, Zhiyuan Liu, and Russell Power. 2017 · 2017
Earlier work this paper cites.
Anserini: Enabling the use of Lucene for information retrieval research. In Proc. of SIGIR
Peilin Yang, Hui Fang, and Jimmy Lin. 2017 · 2017
Earlier work this paper cites.
Fidelity-weighted learning
Mostafa Dehghani, Arash Mehrjou, Stephan Gouws, Jaap Kamps, and Bernhard Schölkopf. 2018 · 2018
Earlier work this paper cites.
Ranking distillation: Learning compact ranking models with high performance for recommender system. In Proc. of SIGKDD
Jiaxi Tang and Ke Wang. 2018 · 2018
Earlier work this paper cites.
Learning a Better Negative Sampling Policy with Deep Neural Networks for Search. In Proc. of ICTIR
Daniel Cohen, Scott M. Jordan, and W. Bruce Croft. 2019 · 2019
Earlier work this paper cites.
Overview of the TREC 2019 deep learning track. In TREC
Nick Craswell, Bhaskar Mitra, Emine Yilmaz, and Daniel Campos. 2019 · 2019
Earlier work this paper cites.