Fetching the paper…
Reading the bibliography…
In recent years, the rapid growth of online multimedia services, such as e-commerce platforms, has necessitated the development of personalised recommendation approaches that can encode diverse content about each item.
BPR: Bayesian personalized ranking from implicit feedback. In Proc. of UAI
Steffen Rendle, Christoph Freudenthaler, Zeno Gantner, and Lars Schmidt-Thieme. 2009 · 2009
Earlier work this paper cites.
Decaf: A deep convolutional activation feature for generic visual recognition. In Proc. of ICML
Jeff Donahue, Yangqing Jia, Oriol Vinyals, Judy Hoffman, Ning Zhang, Eric Tzeng, and Trevor Darrell. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization. In Proc. of ICLR
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Distributed representations of sentences and documents. In Proc. of ICML
Quoc Le and Tomas Mikolov. 2014 · 2014
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition. In Proc. of ICLR
Karen Simonyan and Andrew Zisserman. 2015 · 2015
Earlier work this paper cites.
Deep residual learning for image recognition. In Proc. of CVPR
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
VBPR: Visual bayesian personalized ranking from implicit feedback. In Proc. of AAAI
Ruining He and Julian McAuley. 2016 · 2016
Earlier work this paper cites.
VSE++: Improving visual-semantic embeddings with hard negatives
Fartash Faghri, David J Fleet, Jamie Ryan Kiros, and Sanja Fidler. 2017 · 2017
Earlier work this paper cites.
Audio set: An ontology and human-labeled dataset for audio events. In Proc. of ICASSP
Jort F Gemmeke, Daniel PW Ellis, Dylan Freedman, Aren Jansen, Wade Lawrence, R Channing Moore, Manoj Plakal, and Marvin Ritter. 2017 · 2017
Earlier work this paper cites.
Personalised fashion recommendation with visual explanations based on multimodal attention network: Towards visually explainable recommendation. In Proc. of SIGIR
Xu Chen, Hanxiong Chen, Hongteng Xu, Yongfeng Zhang, Yixin Cao, Zheng Qin, and Hongyuan Zha. 2019 · 2019
Earlier work this paper cites.
VilBERT: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks. In Proc. of NeurIPS
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019 · 2019
Earlier work this paper cites.
Justifying recommendations using distantly-labeled reviews and fine-grained aspects. In Proc. of EMNLP-IJCNLP
Jianmo Ni, Jiacheng Li, and Julian McAuley. 2019 · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al · 2019
Cited alongside, same era.
Sentence-BERT: Sentence embeddings using siamese bert-networks. In Proc. of EMNLP-IJCNLP
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
MMGCN: Multi-modal graph convolution network for personalised recommendation of micro-video. In Proc. of MM
Yinwei Wei, Xiang Wang, Liqiang Nie, Xiangnan He, Richang Hong, and Tat-Seng Chua. 2019 · 2019
Cited alongside, same era.
Learning transferable visual models from natural language supervision. In Proc. of ICML
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Cited alongside, same era.
Multi-modal contrastive pre-training for recommendation. In Proc. of ICMR
Zhuang Liu, Yunpu Ma, Matthias Schubert, Yuanxin Ouyang, and Zhang Xiong. 2022 · 2022
Later among the works it cites.
Multimodal meta-Learning for cold-Start sequential recommendation. In Proc. of CIKM
Xingyu Pan, Yushuo Chen, Changxin Tian, Zihan Lin, Jinpeng Wang, He Hu, and Wayne Xin Zhao. 2022 · 2022
Later among the works it cites.
Where does the performance improvement come from? -A reproducibility concern about image-text retrieval. In Proc. of SIGIR
Jun Rao, Fei Wang, Liang Ding, Shuhan Qi, Yibing Zhan, Weifeng Liu, and Dacheng Tao. 2022 · 2022
Later among the works it cites.
Self-supervised learning for multimedia recommendation
Zhulin Tao, Xiaohao Liu, Yewei Xia, Xiang Wang, Lifang Yang, Xianglin Huang, and Tat-Seng Chua. 2022 · 2022
Later among the works it cites.
Multi-modal graph contrastive learning for micro-video recommendation. In Proc. of SIGIR
Zixuan Yi, Xi Wang, Iadh Ounis, and Craig Macdonald. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Self-supervised graph learning for recommendation. In Proc. of SIGIR
Jiancan Wu, Xiang Wang, Fuli Feng, Xiangnan He, Liang Chen, Jianxun Lian, and Xing Xie. 2021 · 2021
Cited alongside, same era.
Mining latent structures for multimedia recommendation. In Proc. of MM
Liu Qiang Zhang Jinghao, Zhu Yanqiao, Wu Shu, Wang Shuhui, and Wang Liang. 2021 · 2021
Cited alongside, same era.
Vlmo: Unified vision-language pre-training with mixture-of-modality-experts. In Proc. of NeurIPS
Hangbo Bao, Wenhui Wang, Li Dong, Qiang Liu, Owais Khan Mohammed, Kriti Aggarwal, Subhojit Som, Songhao Piao, and Furu Wei. 2022 · 2022
Cited alongside, same era.
MARIO: Modality-aware attention and modality-preserving decoders for multimedia recommendation. In Proc. of CIKM
Taeri Kim, Yeon-Chang Lee, Kijung Shin, and Sang-Wook Kim. 2022 · 2022
Cited alongside, same era.
An image is worth 16x16 words: Transformers for image recognition at scale. In Proc. of ICLR
Alexander Kolesnikov, Alexey Dosovitskiy, Dirk Weissenborn, Georg Heigold, Jakob Uszkoreit, Lucas Beyer, Matthias Minderer, Mostafa Dehghani, Neil Houlsby, Sylvain Gelly, Thomas Unterthiner, and Xiaohua Zhai. 2022 · 2022
Cited alongside, same era.
Graph contrastive learning with positional representation for recommendation. In Proc. of ECIR
Zixuan Yi, Iadh Ounis, and Craig Macdonald. 2023b
Cited in the paper.
Multi-modal recommender systems: A Survey
Qidong Liu, Jiaxi Hu, Yutian Xiao, Jingtong Gao, and Xiangyu Zhao. 2023 · 2023
Closest in time.
Large-scale multi-modal pre-trained models: A comprehensive survey
Xiao Wang, Guangyao Chen, Guangwu Qian, Pengcheng Gao, Xiao-Yong Wei, Yaowei Wang, Yonghong Tian, and Wen Gao. 2023 · 2023
Closest in time.
LightGT: A light graph transformer for multimedia recommendation. In Proc. of SIGIR
Yinwei Wei, Wenqi Liu, Fan Liu, Xiang Wang, Liqiang Nie, and Tat-Seng Chua. 2023 · 2023
Closest in time.
Contrastive graph prompt-tuning for cross-domain recommendation
Zixuan Yi, Iadh Ounis, and Craig Macdonald. 2023a · 2023
Closest in time.
Hongyu Zhou, Xin Zhou, Zhiwei Zeng, Lingzi Zhang, and Zhiqi Shen. 2023 · 2023
Closest in time.