Fetching the paper…
Reading the bibliography…
Multimodal entity linking (MEL) task, which aims at resolving ambiguous mentions to a multimodal knowledge graph, has attracted wide attention in recent years.
Entity search engine: Towards agile best-effort information integration over the web
Tao Cheng and Kevin Chen-Chuan Chang · 2007
Earlier work this paper cites.
Wikidata: a free collaborative knowledgebase
Denny Vrandecic and Markus Krötzsch · 2014
Earlier work this paper cites.
Capturing semantic similarity for entity linking with convolutional neural networks
Matthew Francis-Landau, Greg Durrett, and Dan Klein · 2016
Earlier work this paper cites.
Rethinking the inception architecture for computer vision
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jonathon Shlens, and Zbigniew Wojna · 2016
Earlier work this paper cites.
Cross-lingual wikification using multilingual embeddings
Chen-Tse Tsai and Dan Roth · 2016
Earlier work this paper cites.
Joint learning of the embedding of words and entities for named entity disambiguation
Ikuya Yamada, Hiroyuki Shindo, Hideaki Takeda, and Yoshiyasu Takefuji · 2016
Earlier work this paper cites.
Bridge text and knowledge by learning multi-prototype entity mention embedding
Yixin Cao, Lifu Huang, Heng Ji, Xu Chen, and Juanzi Li · 2017
Earlier work this paper cites.
Named entity disambiguation for noisy text
Yotam Eshel, Noam Cohen, Kira Radinsky, Shaul Markovitch, Ikuya Yamada, and Omer Levy · 2017
Earlier work this paper cites.
Entity linking via joint encoding of types, descriptions, and context
Nitish Gupta, Sameer Singh, and Dan Roth · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Neural collective entity linking
Yixin Cao, Lei Hou, Juanzi Li, and Zhiyuan Liu · 2018
Earlier work this paper cites.
Improving entity linking by modeling latent relations between mentions
Phong Le and Ivan Titov · 2018
Earlier work this paper cites.
Knowledge-aware multimodal dialogue systems
Lizi Liao, Yunshan Ma, Xiangnan He, Richang Hong, and Tat-Seng Chua · 2018
Earlier work this paper cites.
Multimodal named entity disambiguation for noisy social media posts
Seungwhan Moon, Leonardo Neves, and Vitor Carvalho · 2018
Earlier work this paper cites.
Concet: Entity-aware topic classification for open-domain conversational agents
Ali Ahmadvand, Harshita Sahijwani, Jason Ingyu Choi, and Eugene Agichtein · 2019
Earlier work this paper cites.
Cross-modal image-text retrieval with semantic consistency
Hui Chen, Guiguang Ding, Zijia Lin, Sicheng Zhao, and Jungong Han · 2019
Cited alongside, same era.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Cited alongside, same era.
Joint entity linking with deep reinforcement learning
Zheng Fang, Yanan Cao, Qian Li, Dongjie Zhang, Zhenyu Zhang, and Yanbing Liu · 2019
Cited alongside, same era.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2019
Cited alongside, same era.
Autoregressive entity retrieval
Nicola De Cao, Gautier Izacard, Sebastian Riedel, and Fabio Petroni · 2021
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, Jakob Uszkoreit, and Neil Houlsby · 2021
Later among the works it cites.
Multimodal entity linking: A new dataset and A baseline
Jingru Gan, Jinchang Luo, Haiwei Wang, Shuhui Wang, Wei He, and Qingming Huang · 2021
Later among the works it cites.
Vilt: Vision-and-language transformer without convolution or region supervision
Wonjae Kim, Bokyung Son, and Ildoo Kim · 2021
Later among the works it cites.
Align before fuse: Vision and language representation learning with momentum distillation
Junnan Li, Ramprasaath R. Selvaraju, Akhilesh Gotmare, Shafiq R. Joty, Caiming Xiong, and Steven Chu-Hong Hoi · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Köpf, Edward Z. Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala · 2019
Cited alongside, same era.
Knowledge enhanced contextual word representations
Matthew E. Peters, Mark Neumann, Robert L. Logan IV, Roy Schwartz, Vidur Joshi, Sameer Singh, and Noah A. Smith · 2019
Cited alongside, same era.
Improving question answering over incomplete kbs with knowledge-aware reader
Wenhan Xiong, Mo Yu, Shiyu Chang, Xiaoxiao Guo, and William Yang Wang · 2019
Cited alongside, same era.
Learning dynamic context augmentation for global entity linking
Xiyuan Yang, Xiaotao Gu, Sheng Lin, Siliang Tang, Yueting Zhuang, Fei Wu, Zhigang Chen, Guoping Hu, and Xiang Ren · 2019
Cited alongside, same era.
Multimodal entity linking for tweets
Omar Adjali, Romaric Besançon, Olivier Ferret, Hervé Le Borgne, and Brigitte Grau · 2020
Cited alongside, same era.
BART: denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer · 2020
Cited alongside, same era.
Richpedia: A large-scale, comprehensive multi-modal knowledge graph
Meng Wang, Haofen Wang, Guilin Qi, and Qiushuo Zheng · 2020
Cited alongside, same era.
Entity-based knowledge conflicts in question answering
Shayne Longpre, Kartik Perisetla, Anthony Chen, Nikhil Ramesh, Chris DuBois, and Sameer Singh · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Later among the works it cites.
Attention-based multimodal entity linking with high-quality images
Li Zhang, Zhixu Li, and Qiang Yang · 2021
Later among the works it cites.
Multi-modal siamese network for entity alignment
Liyi Chen, Zhi Li, Tong Xu, Han Wu, Zhefeng Wang, Nicholas Jing Yuan, and Enhong Chen · 2022
Later among the works it cites.
An empirical study of training end-to-end vision-and-language transformers
Zi-Yi Dou, Yichong Xu, Zhe Gan, Jianfeng Wang, Shuohang Wang, Lijuan Wang, Chenguang Zhu, Pengchuan Zhang, Lu Yuan, Nanyun Peng, Zicheng Liu, and Michael Zeng · 2022
Later among the works it cites.
Entity-aware transformers for entity search
Emma J. Gerritse, Faegheh Hasibi, and Arjen P. de Vries · 2022
Later among the works it cites.
Multimodal entity linking with gated hierarchical fusion and contrastive training
Peng Wang, Jiangheng Wu, and Xiaohang Chen · 2022
Later among the works it cites.
Wikidiverse: A multimodal entity linking dataset with diversified contextual topics and entity types
Xuwu Wang, Junfeng Tian, Min Gui, Zhixu Li, Rui Wang, Ming Yan, Lihan Chen, and Yanghua Xiao · 2022
Later among the works it cites.
Visual entity linking via multi-modal learning
Qiushuo Zheng, Hao Wen, Meng Wang, and Guilin Qi · 2022
Later among the works it cites.