Fetching the paper…
Reading the bibliography…
InfoNCE loss is a widely used loss function for contrastive model training.
BPR: Bayesian personalized ranking from implicit feedback. In UAI . 452–461
Steffen Rendle, Christoph Freudenthaler, Zeno Gantner, and Lars Schmidt-Thieme. 2009 · 2009
Earlier work this paper cites.
Learning deep structured semantic models for web search using clickthrough data. In CIKM . 2333–2338
Po-Sen Huang, Xiaodong He, Jianfeng Gao, Li Deng, Alex Acero, and Larry Heck. 2013 · 2013
Earlier work this paper cites.
Learning with noisy labels
Nagarajan Natarajan, Inderjit S Dhillon, Pradeep K Ravikumar, and Ambuj Tewari. 2013 · 2013
Earlier work this paper cites.
The movielens datasets: History and context
F Maxwell Harper and Joseph A Konstan. 2015 · 2015
Earlier work this paper cites.
Accurate, large minibatch sgd: Training imagenet in 1 hour
Priya Goyal, Piotr Dollár, Ross Girshick, Pieter Noordhuis, Lukasz Wesolowski, Aapo Kyrola, Andrew Tulloch, Yangqing Jia, and Kaiming He. 2017 · 2017
Earlier work this paper cites.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Earlier work this paper cites.
Neural News Recommendation with Long-and Short-term User Representations. In ACL . 336–345
Mingxiao An, Fangzhao Wu, Chuhan Wu, Kun Zhang, Zheng Liu, and Xing Xie. 2019 · 2019
Earlier work this paper cites.
Learning deep representations by mutual information estimation and maximization. In ICLR
R. Devon Hjelm, Alex Fedorov, Samuel Lavoie-Marchildon, Karan Grewal, Philip Bachman, Adam Trischler, and Yoshua Bengio. 2019 · 2019
Earlier work this paper cites.
A Mutual Information Maximization Perspective of Language Representation Learning. In ICLR
Lingpeng Kong, Cyprien de Masson d’Autume, Lei Yu, Wang Ling, Zihang Dai, and Dani Yogatama. 2019 · 2019
Cited alongside, same era.
Can gradient clipping mitigate label noise?. In ICLR
Aditya Krishna Menon, Ankit Singh Rawat, Sashank J Reddi, and Sanjiv Kumar. 2019 · 2019
Cited alongside, same era.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks. In EMNLP-IJCNLP . 3973–3983
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
BERT4Rec: Sequential recommendation with bidirectional encoder representations from transformer. In CIKM . 1441–1450
Fei Sun, Jun Liu, Jian Wu, Changhua Pei, Xiao Lin, Wenwu Ou, and Peng Jiang. 2019 · 2019
Cited alongside, same era.
Unsupervised Video Representation Learning by Bidirectional Feature Prediction. In CVPR . 1670–1679
Nadine Behrmann, Jurgen Gall, and Mehdi Noroozi. 2020 · 2020
Cited alongside, same era.
Momentum contrast for unsupervised visual representation learning. In CVPR . 9729–9738
Kaiming He, Haoqi Fan, Yuxin Wu, Saining Xie, and Ross Girshick. 2020 · 2020
Later among the works it cites.
Hard negative mixing for contrastive learning
Yannis Kalantidis, Mert Bulent Sariyildiz, Noe Pion, Philippe Weinzaepfel, and Diane Larlus. 2020 · 2020
Later among the works it cites.
Contrastive learning with hard negative samples
Joshua Robinson, Ching-Yao Chuang, Suvrit Sra, and Stefanie Jegelka. 2020 · 2020
Later among the works it cites.
Context Encoding for Video Retrieval with Contrastive Learning
Jie Shao, Xin Wen, Bingchen Zhao, Changhu Wang, and Xiangyang Xue. 2020 · 2020
Later among the works it cites.
Contrastive Distillation on Intermediate Representations for Language Model Compression. In EMNLP . 498–508
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Wei-Cheng Chang, Felix X Yu, Yin-Wen Chang, Yiming Yang, and Sanjiv Kumar. 2020 · 2020
Cited alongside, same era.
Ching-Yao Chuang, Joshua Robinson, Lin Yen-Chen, Antonio Torralba, and Stefanie Jegelka. 2020 · 2020
Cited alongside, same era.
Neural News Recommendation with Attentive Multi-View Learning. In IJCAI . 3863–3869
Chuhan Wu, Fangzhao Wu, Mingxiao An, Jianqiang Huang, Yongfeng Huang, and Xing Xie. 2019a
Cited in the paper.
Npa: Neural news recommendation with personalized attention. In KDD . 2576–2584
Chuhan Wu, Fangzhao Wu, Mingxiao An, Jianqiang Huang, Yongfeng Huang, and Xing Xie. 2019b
Cited in the paper.
Neural News Recommendation with Multi-Head Self-Attention. In EMNLP-IJCNLP . 6390–6395
Chuhan Wu, Fangzhao Wu, Suyu Ge, Tao Qi, Yongfeng Huang, and Xing Xie. 2019c
Cited in the paper.
Siqi Sun, Zhe Gan, Yuwei Fang, Yu Cheng, Shuohang Wang, and Jingjing Liu. 2020 · 2020
Later among the works it cites.
Mind: A large-scale dataset for news recommendation. In ACL . 3597–3606
Fangzhao Wu, Ying Qiao, Jiun-Hung Chen, Chuhan Wu, Tao Qi, Jianxun Lian, Danyang Liu, Xing Xie, Jianfeng Gao, Winnie Wu, et al · 2020
Later among the works it cites.
Infoxlm: An information-theoretic framework for cross-lingual language model pre-training
Zewen Chi, Li Dong, Furu Wei, Nan Yang, Saksham Singhal, Wenhui Wang, Xia Song, Xian-Ling Mao, Heyan Huang, and Ming Zhou. 2021 · 2021
Closest in time.