Fetching the paper…
Reading the bibliography…
Knowledge Graphs (KGs) play a pivotal role in advancing various AI applications, with the semantic web community's exploration into multi-modal dimensions unlocking new avenues for innovation.
W. Cai, W. Ma, J. Zhan, and Y. Jiang, “Entity alignment with reliable path reasoning and relation-aware heterogeneous graph transformer,” in IJCAI . ijcai.org, 2022, pp. 1930–1937
1937
Earlier work this paper cites.
J. Cho, J. Lei, H. Tan, and M. Bansal, “Unifying vision-and-language tasks via text generation,” in ICML , ser. Proceedings of Machine Learning Research, vol. 139. PMLR, 2021, pp. 1931–1942
1942
Earlier work this paper cites.
J. Liu, Y. Chen, and J. Xu, “Multimedia event extraction from news with a unified contrastive learning framework,” in ACM Multimedia . ACM, 2022, pp. 1945–1953
1953
Earlier work this paper cites.
J. Gu, H. Zhao, Z. Lin, S. Li, J. Cai, and M. Ling, “Scene graph generation with external knowledge and image reconstruction,” in CVPR . Computer Vision Foundation / IEEE, 2019, pp. 1969–1978
1978
Earlier work this paper cites.
B. A. Kipfer, “Roget’s 21st century thesaurus in dictionary form: the essential reference for home, school, or office,” (No Title) , 1992
1992
Earlier work this paper cites.
Z. Wu and M. Palmer, “Verb semantics and lexical selection,” arXiv preprint cmp-lg/9406033 , 1994
1994
Earlier work this paper cites.
G. A. Miller, “WordNet: A lexical database for english,” Communications of the ACM , vol. 38, no. 11, pp. 39–41, 1995
1995
Earlier work this paper cites.
C. F. Baker, C. J. Fillmore, and J. B. Lowe, “The berkeley framenet project,” in COLING-ACL . Morgan Kaufmann Publishers / ACL, 1998, pp. 86–90
1998
Earlier work this paper cites.
D. Lu, L. Neves, V. Carvalho, N. Zhang, and H. Ji, “Visual attention model for name tagging in multimodal social media,” in ACL (1) . Association for Computational Linguistics, 2018, pp. 1990–1999
1999
Earlier work this paper cites.
K. Papineni, S. Roukos, T. Ward, and W. Zhu, “Bleu: a method for automatic evaluation of machine translation,” in ACL . ACL, 2002, pp. 311–318
2002
Earlier work this paper cites.
O. Bodenreider, “The unified medical language system (UMLS): integrating biomedical terminology,” Nucleic Acids Res. , vol. 32, no. Database-Issue, pp. 267–270, 2004
2004
Earlier work this paper cites.
L. B. Smith and M. Gasser, “The development of embodied cognition: Six lessons from babies,” Artif. Life , vol. 11, no. 1-2, pp. 13–29, 2005
2005
Earlier work this paper cites.
J. R. Finkel, T. Grenager, and C. D. Manning, “Incorporating non-local information into information extraction systems by gibbs sampling,” in ACL . The Association for Computer Linguistics, 2005, pp. 363–370
2005
Earlier work this paper cites.
L. Denoyer and P. Gallinari, “The wikipedia xml corpus,” in ACM SIGIR Forum , vol. 40, no. 1. ACM New York, NY, USA, 2006, pp. 64–69
2006
Earlier work this paper cites.
F. M. Suchanek, G. Kasneci, and G. Weikum, “Yago: a core of semantic knowledge,” in WWW . ACM, 2007, pp. 697–706
2007
Earlier work this paper cites.
S. Auer, C. Bizer, G. Kobilarov, J. Lehmann, R. Cyganiak, and Z. G. Ives, “Dbpedia: A nucleus for a web of open data,” in ISWC/ASWC , ser. Lecture Notes in Computer Science, vol. 4825. Springer, 2007, pp. 722–735
2007
Earlier work this paper cites.
K. D. Bollacker, C. Evans, P. K. Paritosh, T. Sturge, and J. Taylor, “Freebase: a collaboratively created graph database for structuring human knowledge,” in SIGMOD Conference . ACM, 2008, pp. 1247–1250
2008
Earlier work this paper cites.
I. Horrocks, “Ontologies and the semantic web,” Communications of the ACM , vol. 51, no. 12, pp. 58–67, 2008
2008
Earlier work this paper cites.
S. Moon, L. Neves, and V. Carvalho, “Multimodal named entity disambiguation for noisy social media posts,” in ACL (1) . Association for Computational Linguistics, 2018, pp. 2000–2008
2008
Earlier work this paper cites.
J. Deng, W. Dong, R. Socher, L. Li, K. Li, and L. Fei-Fei, “Imagenet: A large-scale hierarchical image database,” in CVPR . IEEE Computer Society, 2009, pp. 248–255
2009
Earlier work this paper cites.
S. Bird, E. Klein, and E. Loper, Natural Language Processing with Python . O’Reilly, 2009
2009
Earlier work this paper cites.
T. Chua, J. Tang, R. Hong, H. Li, Z. Luo, and Y. Zheng, “NUS-WIDE: a real-world web image database from national university of singapore,” in CIVR . ACM, 2009
2009
Earlier work this paper cites.
R. Navigli and S. P. Ponzetto, “Babelnet: The automatic construction, evaluation and application of a wide-coverage multilingual semantic network,” Artif. Intell. , vol. 193, pp. 217–250, 2012
2012
Earlier work this paper cites.
W. Wu, H. Li, H. Wang, and K. Q. Zhu, “Probase: a probabilistic taxonomy for text understanding,” in SIGMOD Conference . ACM, 2012, pp. 481–492
2012
Earlier work this paper cites.
A. Gaulton, L. J. Bellis, A. P. Bento, J. Chambers, M. Davies, A. Hersey, Y. Light, S. McGlinchey, D. Michalovich, B. Al-Lazikani et al. , “Chembl: a large-scale bioactivity database for drug discovery,” Nucleic acids research , vol. 40, no. D1, pp. D1100–D1107, 2012
2012
Earlier work this paper cites.
H. Kilicoglu, D. Shin, M. Fiszman, G. Rosemblat, and T. C. Rindflesch, “Semmeddb: a pubmed-scale repository of biomedical semantic predications,” Bioinformatics , vol. 28, no. 23, pp. 3158–3160, 2012
2012
Earlier work this paper cites.
C. Xu, D. Tao, and C. Xu, “A survey on multi-view learning,” CoRR , vol. abs/1304.5634, 2013
2013
Earlier work this paper cites.
X. Chen, A. Shrivastava, and A. Gupta, “NEIL: extracting visual knowledge from web data,” in ICCV . IEEE Computer Society, 2013, pp. 1409–1416
2013
Earlier work this paper cites.
A. Bordes, N. Usunier, A. García-Durán, J. Weston, and O. Yakhnenko, “Translating embeddings for modeling multi-relational data,” in NIPS , 2013, pp. 2787–2795
2013
Earlier work this paper cites.
A. Frome, G. S. Corrado, J. Shlens, S. Bengio, J. Dean, M. Ranzato, and T. Mikolov, “Devise: A deep visual-semantic embedding model,” in NIPS , 2013, pp. 2121–2129
2013
Earlier work this paper cites.
Z. Akata, F. Perronnin, Z. Harchaoui, and C. Schmid, “Label-embedding for attribute-based classification,” in CVPR . IEEE Computer Society, 2013, pp. 819–826
2013
Earlier work this paper cites.
N. Tandon, G. de Melo, F. M. Suchanek, and G. Weikum, “Webchild: harvesting and organizing commonsense knowledge from the web,” in WSDM . ACM, 2014, pp. 523–532
2014
Earlier work this paper cites.
D. Vrandecic and M. Krötzsch, “Wikidata: a free collaborative knowledgebase,” Commun. ACM , vol. 57, no. 10, pp. 78–85, 2014
2014
Earlier work this paper cites.
D. Chen and C. D. Manning, “A fast and accurate dependency parser using neural networks,” in EMNLP . ACL, 2014, pp. 740–750
2014
Earlier work this paper cites.
J. Pennington, R. Socher, and C. D. Manning, “Glove: Global vectors for word representation,” in EMNLP . ACL, 2014, pp. 1532–1543
2014
Earlier work this paper cites.
C. H. Lampert, H. Nickisch, and S. Harmeling, “Attribute-based classification for zero-shot visual object categorization,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 36, no. 3, pp. 453–465, 2014
2014
Earlier work this paper cites.
P. Young, A. Lai, M. Hodosh, and J. Hockenmaier, “From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,” Trans. Assoc. Comput. Linguistics , vol. 2, pp. 67–78, 2014
2014
Earlier work this paper cites.
T. Lin, M. Maire, S. J. Belongie, J. Hays, P. Perona, D. Ramanan, P. Dollár, and C. L. Zitnick, “Microsoft COCO: common objects in context,” in ECCV (5) , ser. Lecture Notes in Computer Science, vol. 8693. Springer, 2014, pp. 740–755
2014
Earlier work this paper cites.
W. Shen, J. Wang, and J. Han, “Entity linking with a knowledge base: Issues, techniques, and solutions,” TKDE , vol. 27, no. 2, pp. 443–460, 2014
2014
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. S. Bernstein, A. C. Berg, and L. Fei-Fei, “Imagenet large scale visual recognition challenge,” Int. J. Comput. Vis. , vol. 115, no. 3, pp. 211–252, 2015
2015
Earlier work this paper cites.
F. Sadeghi, S. K. Divvala, and A. Farhadi, “Viske: Visual knowledge extraction and question answering by visual verification of relation phrases,” in CVPR . IEEE Computer Society, 2015, pp. 1456–1464
2015
Earlier work this paper cites.
J. Johnson, R. Krishna, M. Stark, L. Li, D. A. Shamma, M. S. Bernstein, and L. Fei-Fei, “Image retrieval using scene graphs,” in CVPR . IEEE Computer Society, 2015, pp. 3668–3678
2015
Earlier work this paper cites.
S. Antol, A. Agrawal, J. Lu, M. Mitchell, D. Batra, C. L. Zitnick, and D. Parikh, “VQA: visual question answering,” in ICCV . IEEE Computer Society, 2015, pp. 2425–2433
2015
Earlier work this paper cites.
R. Vedantam, C. L. Zitnick, and D. Parikh, “Cider: Consensus-based image description evaluation,” in CVPR . IEEE Computer Society, 2015, pp. 4566–4575
2015
Earlier work this paper cites.
X. Li, S. Liao, W. Lan, X. Du, and G. Yang, “Zero-shot image tagging by hierarchical semantic embedding,” in SIGIR . ACM, 2015, pp. 879–882
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
K. Toutanova and D. Chen, “Observed versus latent features for knowledge base and text inference,” in CVSC . Association for Computational Linguistics, 2015, pp. 57–66
2015
Earlier work this paper cites.
T. H. Nguyen and R. Grishman, “Event detection and domain adaptation with convolutional neural networks,” in ACL (2) . The Association for Computer Linguistics, 2015, pp. 365–371
2015
Earlier work this paper cites.
B. Thomee, D. A. Shamma, G. Friedland, B. Elizalde, K. Ni, D. Poland, D. Borth, and L. Li, “YFCC100M: the new data in multimedia research,” Commun. ACM , vol. 59, no. 2, pp. 64–73, 2016
2016
Earlier work this paper cites.
C. Lu, R. Krishna, M. S. Bernstein, and L. Fei-Fei, “Visual relationship detection with language priors,” in ECCV (1) , ser. Lecture Notes in Computer Science, vol. 9905. Springer, 2016, pp. 852–869
2016
Earlier work this paper cites.
M. Yatskar, V. Ordonez, and A. Farhadi, “Stating the obvious: Extracting visual common sense knowledge,” in HLT-NAACL . The Association for Computational Linguistics, 2016, pp. 193–198
2016
Earlier work this paper cites.
Q. Wu, P. Wang, C. Shen, A. R. Dick, and A. van den Hengel, “Ask me anything: Free-form visual question answering based on knowledge from external sources,” in CVPR . IEEE Computer Society, 2016, pp. 4622–4630
2016
Earlier work this paper cites.
Z. Yang, X. He, J. Gao, L. Deng, and A. J. Smola, “Stacked attention networks for image question answering,” in CVPR . IEEE Computer Society, 2016, pp. 21–29
2016
Earlier work this paper cites.
A. Kumar, O. Irsoy, P. Ondruska, M. Iyyer, J. Bradbury, I. Gulrajani, V. Zhong, R. Paulus, and R. Socher, “Ask me anything: Dynamic memory networks for natural language processing,” in ICML , ser. JMLR Workshop and Conference Proceedings, vol. 48. JMLR.org, 2016, pp. 1378–1387
2016
Earlier work this paper cites.
T. Chen and C. Guestrin, “Xgboost: A scalable tree boosting system,” in KDD . ACM, 2016, pp. 785–794
2016
Earlier work this paper cites.
Y. Zhu, O. Groth, M. S. Bernstein, and L. Fei-Fei, “Visual7w: Grounded question answering in images,” in CVPR . IEEE Computer Society, 2016, pp. 4995–5004
2016
Earlier work this paper cites.
N. Mostafazadeh, I. Misra, J. Devlin, M. Mitchell, X. He, and L. Vanderwende, “Generating natural questions about an image,” in ACL (1) . The Association for Computer Linguistics, 2016
2016
Earlier work this paper cites.
Zeynep Akata and Florent Perronnin and Zaïd Harchaoui and Cordelia Schmid, “Label-embedding for image classification,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 38, no. 7, pp. 1425–1438, 2016
2016
Earlier work this paper cites.
T. Salimans, I. J. Goodfellow, W. Zaremba, V. Cheung, A. Radford, and X. Chen, “Improved techniques for training gans,” in NIPS , 2016, pp. 2226–2234
2016
Earlier work this paper cites.
P. Anderson, B. Fernando, M. Johnson, and S. Gould, “SPICE: semantic propositional image caption evaluation,” in ECCV (5) , ser. Lecture Notes in Computer Science, vol. 9909. Springer, 2016, pp. 382–398
2016
Earlier work this paper cites.
G. Lample, M. Ballesteros, S. Subramanian, K. Kawakami, and C. Dyer, “Neural architectures for named entity recognition,” in Proceedings of the 2016 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . San Diego, California: Association for Computational Linguistics, 2016, pp. 260–270. [Online]. Available: https://aclanthology.org/N16-1030
2016
Earlier work this paper cites.
T. H. Nguyen, K. Cho, and R. Grishman, “Joint event extraction via recurrent neural networks,” in HLT-NAACL . The Association for Computational Linguistics, 2016, pp. 300–309
2016
Earlier work this paper cites.
M. Kuhn, I. Letunic, L. J. Jensen, and P. Bork, “The sider database of drugs and side effects,” Nucleic acids research , vol. 44, no. D1, pp. D1075–D1079, 2016
2016
Earlier work this paper cites.
R. Speer, J. Chin, and C. Havasi, “Conceptnet 5.5: An open multilingual graph of general knowledge,” in Thirty-first AAAI conference on artificial intelligence , 2017
2017
Earlier work this paper cites.
N. Tandon, G. de Melo, and G. Weikum, “Webchild 2.0 : Fine-grained commonsense knowledge distillation,” in ACL (System Demonstrations) . Association for Computational Linguistics, 2017, pp. 115–120
2017
Earlier work this paper cites.
R. Krishna, Y. Zhu, O. Groth, J. Johnson, K. Hata, J. Kravitz, S. Chen, Y. Kalantidis, L. Li, D. A. Shamma, M. S. Bernstein, and L. Fei-Fei, “Visual genome: Connecting language and vision using crowdsourced dense image annotations,” Int. J. Comput. Vis. , vol. 123, no. 1, pp. 32–73, 2017
2017
Earlier work this paper cites.
R. Xie, Z. Liu, H. Luan, and M. Sun, “Image-embodied knowledge representation learning,” in IJCAI . ijcai.org, 2017, pp. 3140–3146
2017
Earlier work this paper cites.
Z. Liu, S. Wang, L. Zheng, and Q. Tian, “Robust imagegraph: Rank-level feature fusion for image search,” IEEE Trans. Image Process. , vol. 26, no. 7, pp. 3128–3141, 2017
2017
Earlier work this paper cites.
S. Ferrada, B. Bustos, and A. Hogan, “Imgpedia: A linked dataset with content-based analysis of wikimedia images,” in ISWC (2) , ser. Lecture Notes in Computer Science, vol. 10588. Springer, 2017, pp. 84–93
2017
Earlier work this paper cites.
Z. Sun, W. Hu, and C. Li, “Cross-lingual entity alignment via joint attribute-preserving embedding,” in ISWC (1) , ser. Lecture Notes in Computer Science, vol. 10587. Springer, 2017, pp. 628–644
2017
Earlier work this paper cites.
J. Peyre, I. Laptev, C. Schmid, and J. Sivic, “Weakly-supervised learning of visual relations,” in ICCV . IEEE Computer Society, 2017, pp. 5189–5198
2017
Earlier work this paper cites.
D. Gong and D. Z. Wang, “Extracting visual knowledge from the web with multimodal learning,” in IJCAI . ijcai.org, 2017, pp. 1718–1724
2017
Earlier work this paper cites.
P. Wang, Q. Wu, C. Shen, A. R. Dick, and A. van den Hengel, “Explicit knowledge-based reasoning for visual question answering,” in IJCAI . ijcai.org, 2017, pp. 1290–1296
2017
Earlier work this paper cites.
H. Ben-Younes, R. Cadène, M. Cord, and N. Thome, “MUTAN: multimodal tucker fusion for visual question answering,” in ICCV . IEEE Computer Society, 2017, pp. 2631–2639
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in NIPS , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
F. Monti, D. Boscaini, J. Masci, E. Rodolà, J. Svoboda, and M. M. Bronstein, “Geometric deep learning on graphs and manifolds using mixture model cnns,” in CVPR . IEEE Computer Society, 2017, pp. 5425–5434
2017
Earlier work this paper cites.
R. Li, M. Tapaswi, R. Liao, J. Jia, R. Urtasun, and S. Fidler, “Situation recognition with graph neural networks,” in ICCV . IEEE Computer Society, 2017, pp. 4183–4192
2017
Earlier work this paper cites.
P. Wang, Q. Wu, C. Shen, A. R. Dick, and A. van den Hengel, “FVQA: fact-based visual question answering,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 40, no. 10, pp. 2413–2427, 2018
2018
Earlier work this paper cites.
Y. Chao, Y. Liu, X. Liu, H. Zeng, and J. Deng, “Learning to detect human-object interactions,” in WACV . IEEE Computer Society, 2018, pp. 381–389
2018
Earlier work this paper cites.
M. Narasimhan and A. G. Schwing, “Straight to the facts: Learning knowledge base retrieval for factual visual question answering,” in ECCV (8) , ser. Lecture Notes in Computer Science, vol. 11212. Springer, 2018, pp. 460–477
2018
Earlier work this paper cites.
M. Narasimhan, S. Lazebnik, and A. G. Schwing, “Out of the box: Reasoning with graph convolution nets for factual visual question answering,” in NeurIPS , 2018, pp. 2659–2670
2018
Earlier work this paper cites.
Z. Su, C. Zhu, Y. Dong, D. Cai, Y. Chen, and J. Li, “Learning visual knowledge memory networks for visual question answering,” in CVPR . Computer Vision Foundation / IEEE Computer Society, 2018, pp. 7736–7745
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
M. S. Schlichtkrull, T. N. Kipf, P. Bloem, R. van den Berg, I. Titov, and M. Welling, “Modeling relational data with graph convolutional networks,” in ESWC , ser. Lecture Notes in Computer Science, vol. 10843. Springer, 2018, pp. 593–607
2018
Earlier work this paper cites.
P. Lu, L. Ji, W. Zhang, N. Duan, M. Zhou, and J. Wang, “R-VQA: learning visual relation facts with semantic attention for visual question answering,” in KDD . ACM, 2018, pp. 1880–1889
2018
Earlier work this paper cites.
Y. Zhu, M. Elhoseiny, B. Liu, X. Peng, and A. Elgammal, “A generative adversarial approach for zero-shot learning from noisy texts,” in CVPR . Computer Vision Foundation / IEEE Computer Society, 2018, pp. 1004–1013
2018
Earlier work this paper cites.
Y. Xian, T. Lorenz, B. Schiele, and Z. Akata, “Feature generating networks for zero-shot learning,” in CVPR . Computer Vision Foundation / IEEE Computer Society, 2018, pp. 5542–5551
2018
Earlier work this paper cites.
X. Wang, Y. Ye, and A. Gupta, “Zero-shot recognition via semantic embeddings and knowledge graphs,” in CVPR . Computer Vision Foundation / IEEE Computer Society, 2018, pp. 6857–6866
2018
Earlier work this paper cites.
C. Lee, W. Fang, C. Yeh, and Y. F. Wang, “Multi-label zero-shot learning with structured knowledge graphs,” in CVPR . Computer Vision Foundation / IEEE Computer Society, 2018, pp. 1576–1585
2018
Earlier work this paper cites.
J. Redmon and A. Farhadi, “Yolov3: An incremental improvement,” CoRR , vol. abs/1804.02767, 2018
2018
Earlier work this paper cites.
D. Lu, S. Whitehead, L. Huang, H. Ji, and S. Chang, “Entity-aware image caption generation,” in EMNLP . Association for Computational Linguistics, 2018, pp. 4013–4023
2018
Earlier work this paper cites.
P. Velickovic, G. Cucurull, A. Casanova, A. Romero, P. Liò, and Y. Bengio, “Graph attention networks,” in ICLR (Poster) . OpenReview.net, 2018
2018
Earlier work this paper cites.
T. Xu, P. Zhang, Q. Huang, H. Zhang, Z. Gan, X. Huang, and X. He, “Attngan: Fine-grained text to image generation with attentional generative adversarial networks,” in CVPR . Computer Vision Foundation / IEEE Computer Society, 2018, pp. 1316–1324
2018
Earlier work this paper cites.
R. Zellers, M. Yatskar, S. Thomson, and Y. Choi, “Neural motifs: Scene graph parsing with global context,” in CVPR . Computer Vision Foundation / IEEE Computer Society, 2018, pp. 5831–5840
2018
Earlier work this paper cites.
H. Wan, Y. Luo, B. Peng, and W. Zheng, “Representation learning for scene graph completion via jointly structural and visual embedding,” in IJCAI . ijcai.org, 2018, pp. 949–956
2018
Earlier work this paper cites.
K. Chen, J. Gao, and R. Nevatia, “Knowledge aided consistency for weakly supervised phrase grounding,” in CVPR . Computer Vision Foundation / IEEE Computer Society, 2018, pp. 4042–4050
2018
Earlier work this paper cites.
H. M. Sergieh, T. Botschen, I. Gurevych, and S. Roth, “A multimodal translation-based approach for knowledge graph representation learning,” in *SEM@NAACL-HLT . Association for Computational Linguistics, 2018, pp. 225–234
2018
Earlier work this paper cites.
P. Pezeshkpour, L. Chen, and S. Singh, “Embedding multimodal relational data for knowledge base completion,” in EMNLP . Association for Computational Linguistics, 2018, pp. 3208–3218
2018
Earlier work this paper cites.
S. Moon, L. Neves, and V. Carvalho, “Multimodal named entity recognition for short social media posts,” in NAACL-HLT . Association for Computational Linguistics, 2018, pp. 852–860
2018
Earlier work this paper cites.
Q. Zhang, J. Fu, X. Liu, and X. Huang, “Adaptive co-attention network for named entity recognition in tweets,” in AAAI . AAAI Press, 2018, pp. 5674–5681
2018
Earlier work this paper cites.
P. Qi, T. Dozat, Y. Zhang, and C. D. Manning, “Universal dependency parsing from scratch,” in CoNLL Shared Task (2) . Association for Computational Linguistics, 2018, pp. 160–170
2018
Earlier work this paper cites.
R. Das, S. Dhuliawala, M. Zaheer, L. Vilnis, I. Durugkar, A. Krishnamurthy, A. Smola, and A. McCallum, “Go for a walk and arrive at the answer: Reasoning over paths in knowledge bases using reinforcement learning,” in ICLR (Poster) . OpenReview.net, 2018
2018
Earlier work this paper cites.
B. Percha and R. B. Altman, “A global network of biomedical relationships derived from text,” Bioinformatics , vol. 34, no. 15, pp. 2614–2624, 2018
2018
Earlier work this paper cites.
K. Marino, M. Rastegari, A. Farhadi, and R. Mottaghi, “OK-VQA: A visual question answering benchmark requiring external knowledge,” in CVPR . Computer Vision Foundation / IEEE, 2019, pp. 3195–3204
2019
Earlier work this paper cites.
X. Li and S. Jiang, “Know more say less: Image captioning based on scene graphs,” IEEE Trans. Multim. , vol. 21, no. 8, pp. 2117–2130, 2019
2019
Earlier work this paper cites.
D. Oñoro-Rubio, M. Niepert, A. García-Durán, R. Gonzalez-Sanchez, and R. J. López-Sastre, “Answering visual-relational queries in web-extracted knowledge graphs,” in AKBC , 2019
2019
Earlier work this paper cites.
Y. Liu, H. Li, A. García-Durán, M. Niepert, D. Oñoro-Rubio, and D. S. Rosenblum, “MMKG: multi-modal knowledge graphs,” in ESWC , ser. Lecture Notes in Computer Science, vol. 11503. Springer, 2019, pp. 459–474
2019
Earlier work this paper cites.
L. Guo, Z. Sun, and W. Hu, “Learning to exploit long-term relational dependencies in knowledge graphs,” in ICML , ser. Proceedings of Machine Learning Research, vol. 97. PMLR, 2019, pp. 2505–2514
2019
Earlier work this paper cites.
J. Lu, D. Batra, D. Parikh, and S. Lee, “Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks,” in NeurIPS , 2019, pp. 13–23
2019
Earlier work this paper cites.
S. Shah, A. Mishra, N. Yadati, and P. P. Talukdar, “KVQA: knowledge-aware visual question answering,” in AAAI . AAAI Press, 2019, pp. 8876–8884
2019
Earlier work this paper cites.
H. Wang, W. Lu, and Z. Tang, “Incorporating external knowledge to boost machine comprehension based question answering,” in ECIR (1) , ser. Lecture Notes in Computer Science, vol. 11437. Springer, 2019, pp. 819–827
2019
Earlier work this paper cites.
H. Tan and M. Bansal, “LXMERT: learning cross-modality encoder representations from transformers,” in EMNLP/IJCNLP (1) . Association for Computational Linguistics, 2019, pp. 5099–5110
2019
Earlier work this paper cites.
F. Petroni, T. Rocktäschel, S. Riedel, P. S. H. Lewis, A. Bakhtin, Y. Wu, and A. H. Miller, “Language models as knowledge bases?” in EMNLP/IJCNLP (1) . Association for Computational Linguistics, 2019, pp. 2463–2473
2019
Earlier work this paper cites.
A. Lerer, L. Wu, J. Shen, T. Lacroix, L. Wehrstedt, A. Bose, and A. Peysakhovich, “Pytorch-biggraph: A large scale graph embedding system,” in MLSys . mlsys.org, 2019
2019
Earlier work this paper cites.
N. Reimers and I. Gurevych, “Sentence-bert: Sentence embeddings using siamese bert-networks,” in EMNLP/IJCNLP (1) . Association for Computational Linguistics, 2019, pp. 3980–3990
2019
Earlier work this paper cites.
J. Devlin, M. Chang, K. Lee, and K. Toutanova, “BERT: pre-training of deep bidirectional transformers for language understanding,” in NAACL-HLT (1) . Association for Computational Linguistics, 2019, pp. 4171–4186
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
M. Sap, R. L. Bras, E. Allaway, C. Bhagavatula, N. Lourie, H. Rashkin, B. Roof, N. A. Smith, and Y. Choi, “ATOMIC: an atlas of machine commonsense for if-then reasoning,” in AAAI . AAAI Press, 2019, pp. 3027–3035
2019
Earlier work this paper cites.
R. Zellers, Y. Bisk, A. Farhadi, and Y. Choi, “From recognition to cognition: Visual commonsense reasoning,” in CVPR . Computer Vision Foundation / IEEE, 2019, pp. 6720–6731
2019
Earlier work this paper cites.
M. Kampffmeyer, Y. Chen, X. Liang, H. Wang, Y. Zhang, and E. P. Xing, “Rethinking knowledge graph propagation for zero-shot learning,” in CVPR . Computer Vision Foundation / IEEE, 2019, pp. 11 487–11 496
2019
Earlier work this paper cites.
J. Gao, T. Zhang, and C. Xu, “I know the relationships: Zero-shot action recognition via two-stream graph convolutional networks and knowledge graphs,” in AAAI . AAAI Press, 2019, pp. 8303–8311
2019
Earlier work this paper cites.
C. Zhang, X. Lyu, and Z. Tang, “TGG: transferable graph generation for zero-shot and few-shot learning,” in ACM Multimedia . ACM, 2019, pp. 1641–1649
2019
Earlier work this paper cites.
Y. Xian, C. H. Lampert, B. Schiele, and Z. Akata, “Zero-shot learning - A comprehensive evaluation of the good, the bad and the ugly,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 41, no. 9, pp. 2251–2265, 2019
2019
Earlier work this paper cites.
M. Z. Hossain, F. Sohel, M. F. Shiratuddin, and H. Laga, “A comprehensive survey of deep learning for image captioning,” ACM Comput. Surv. , vol. 51, no. 6, pp. 118:1–118:36, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y. Zhou, Y. Sun, and V. G. Honavar, “Improving image captioning by leveraging knowledge graphs,” in WACV . IEEE, 2019, pp. 283–293
2019
Earlier work this paper cites.
P. Yang, F. Luo, P. Chen, L. Li, Z. Yin, X. He, and X. Sun, “Knowledgeable storyteller: A commonsense-driven generative model for visual storytelling,” in IJCAI . ijcai.org, 2019, pp. 5356–5362
2019
Earlier work this paper cites.
T. Qiao, J. Zhang, D. Xu, and D. Tao, “Learn, imagine and create: Text-to-image generation from prior knowledge,” in NeurIPS , 2019, pp. 885–895
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
T. Chen, W. Yu, R. Chen, and L. Lin, “Knowledge-embedded routing network for scene graph generation,” in CVPR . Computer Vision Foundation / IEEE, 2019, pp. 6163–6171
2019
Earlier work this paper cites.
J. M. Gómez-Pérez and R. Ortega, “Look, read and enrich - learning from scientific figures and their captions,” in K-CAP . ACM, 2019, pp. 101–108
2019
Earlier work this paper cites.
B. Shi, L. Ji, P. Lu, Z. Niu, and N. Duan, “Knowledge aware semantic concept expansion for image-text matching,” in IJCAI . ijcai.org, 2019, pp. 5182–5189
2019
Earlier work this paper cites.
S. Yang, G. Li, and Y. Yu, “Cross-modal relationship inference for grounding referring expressions,” in CVPR . Computer Vision Foundation / IEEE, 2019, pp. 4145–4154
2019
Earlier work this paper cites.
A. Bosselut, H. Rashkin, M. Sap, C. Malaviya, A. Celikyilmaz, and Y. Choi, “Comet: Commonsense transformers for automatic knowledge graph construction,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , 2019, pp. 4762–4779
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Z. Sun, Z. Deng, J. Nie, and J. Tang, “Rotate: Knowledge graph embedding by relational rotation in complex space,” in ICLR (Poster) . OpenReview.net, 2019
2019
Earlier work this paper cites.
Z. Wang, L. Li, Q. Li, and D. Zeng, “Multimodal data enhanced representation learning for knowledge graphs,” in IJCNN . IEEE, 2019, pp. 1–8
2019
Earlier work this paper cites.
J. Liu, Y. Chen, K. Liu, and J. Zhao, “Neural cross-lingual event detection with minimal parallel resources,” in EMNLP/IJCNLP (1) . Association for Computational Linguistics, 2019, pp. 738–748
2019
Earlier work this paper cites.
D. Wadden, U. Wennberg, Y. Luan, and H. Hajishirzi, “Entity, relation, and event extraction with contextualized span representations,” in EMNLP/IJCNLP (1) . Association for Computational Linguistics, 2019, pp. 5783–5788
2019
Earlier work this paper cites.
Z. Wang, L. Li, Q. Li, and D. Zeng, “Multimodal data enhanced representation learning for knowledge graphs,” in IJCNN . IEEE, 2019, pp. 1–8
2019
Earlier work this paper cites.
U. Consortium, “Uniprot: a worldwide hub of protein knowledge,” Nucleic acids research , vol. 47, no. D1, pp. D506–D515, 2019
2019
Earlier work this paper cites.
M. Federici, A. Dutta, P. Forré, N. Kushman, and Z. Akata, “Learning robust representations via multi-view information bottleneck,” in ICLR . OpenReview.net, 2020
2020
Earlier work this paper cites.
Y. Wang, W. Huang, F. Sun, T. Xu, Y. Rong, and J. Huang, “Deep multimodal fusion by channel exchanging,” in NeurIPS , 2020
2020
Earlier work this paper cites.
J. Zhao, X. Lin, J. Zhou, J. Yang, L. He, and Z. Yang, “Knowledge-based fine-grained classification for few-shot learning,” in 2020 IEEE International Conference on Multimedia and Expo (ICME) . IEEE, 2020, pp. 1–6
2020
Earlier work this paper cites.
Z. Zhu, J. Yu, Y. Wang, Y. Sun, Y. Hu, and Q. Wu, “Mucko: Multi-layer cross-modal knowledge reasoning for fact-based visual question answering,” in IJCAI . ijcai.org, 2020, pp. 1097–1103
2020
Earlier work this paper cites.
M. Li, A. Zareian, Y. Lin, X. Pan, S. Whitehead, B. Chen, B. Wu, H. Ji, S. Chang, C. R. Voss, D. Napierski, and M. Freedman, “GAIA: A fine-grained multimedia knowledge extraction system,” in ACL (demo) . Association for Computational Linguistics, 2020, pp. 77–86
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
M. Wang, H. Wang, G. Qi, and Q. Zheng, “Richpedia: A large-scale, comprehensive multi-modal knowledge graph,” Big Data Res. , vol. 22, p. 100159, 2020
2020
Earlier work this paper cites.
Z. Sun, Q. Zhang, W. Hu, C. Wang, M. Chen, F. Akrami, and C. Li, “A benchmarking study of embedding-based entity alignment for knowledge graphs,” Proc. VLDB Endow. , vol. 13, no. 11, pp. 2326–2340, 2020
2020
Earlier work this paper cites.
M. Baumgartner, L. Rossetto, and A. Bernstein, “Towards using semantic-web technologies for multi-modal knowledge graph construction,” in ACM Multimedia . ACM, 2020, pp. 4645–4649
2020
Earlier work this paper cites.
J. Yu, Z. Zhu, Y. Wang, W. Zhang, Y. Hu, and J. Tan, “Cross-modal knowledge reasoning for knowledge-based visual question answering,” Pattern Recognit. , vol. 108, p. 107563, 2020
2020
Earlier work this paper cites.
F. Gardères, M. Ziaeefard, B. Abeloos, and F. Lécué, “Conceptbert: Concept-aware representation for visual question answering,” in EMNLP (Findings) , ser. Findings of ACL, vol. EMNLP 2020. Association for Computational Linguistics, 2020, pp. 489–498
2020
Earlier work this paper cites.
M. Ziaeefard and F. Lécué, “Towards knowledge-augmented visual question answering,” in COLING . International Committee on Computational Linguistics, 2020, pp. 1863–1873
2020
Earlier work this paper cites.
G. Li, X. Wang, and W. Zhu, “Boosting visual question answering with context-aware knowledge aggregation,” in ACM Multimedia . ACM, 2020, pp. 1227–1235
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
P. Qi, Y. Zhang, Y. Zhang, J. Bolton, and C. D. Manning, “Stanza: A python natural language processing toolkit for many human languages,” in ACL (demo) . Association for Computational Linguistics, 2020, pp. 101–108
2020
Earlier work this paper cites.
R. Saqur and K. Narasimhan, “Multimodal graph networks for compositional generalization in visual question answering,” in NeurIPS , 2020
2020
Earlier work this paper cites.
J. Lu, V. Goswami, M. Rohrbach, D. Parikh, and S. Lee, “12-in-1: Multi-task vision and language representation learning,” in CVPR . Computer Vision Foundation / IEEE, 2020, pp. 10 434–10 443
2020
Earlier work this paper cites.
V. Karpukhin, B. Oguz, S. Min, P. S. H. Lewis, L. Wu, S. Edunov, D. Chen, and W. Yih, “Dense passage retrieval for open-domain question answering,” in EMNLP (1) . Association for Computational Linguistics, 2020, pp. 6769–6781
2020
Earlier work this paper cites.
C. Malaviya, C. Bhagavatula, A. Bosselut, and Y. Choi, “Commonsense knowledge base completion with structural and semantic context,” in AAAI . AAAI Press, 2020, pp. 2925–2933
2020
Earlier work this paper cites.
Z. Lan, M. Chen, S. Goodman, K. Gimpel, P. Sharma, and R. Soricut, “ALBERT: A lite BERT for self-supervised learning of language representations,” in ICLR . OpenReview.net, 2020
2020
Earlier work this paper cites.
Y. Chen, L. Li, L. Yu, A. E. Kholy, F. Ahmed, Z. Gan, Y. Cheng, and J. Liu, “UNITER: universal image-text representation learning,” in ECCV (30) , ser. Lecture Notes in Computer Science, vol. 12375. Springer, 2020, pp. 104–120
2020
Earlier work this paper cites.
W. Su, X. Zhu, Y. Cao, B. Li, L. Lu, F. Wei, and J. Dai, “VL-BERT: pre-training of generic visual-linguistic representations,” in ICLR . OpenReview.net, 2020
2020
Earlier work this paper cites.
T. B. Brown, B. Mann, N. Ryder, M. Subbiah, J. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell, S. Agarwal, A. Herbert-Voss, G. Krueger, T. Henighan, R. Child, A. Ramesh, D. M. Ziegler, J. Wu, C. Winter, C. Hesse, M. Chen, E. Sigler, M. Litwin, S. Gray, B. Chess, J. Clark, C. Berner, S. McCandlish, A. Radford, I. Sutskever, and D. Amodei, “Language models are few-shot learners,” in NeurIPS , 2020
2020
Earlier work this paper cites.
C. Raffel, N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu, “Exploring the limits of transfer learning with a unified text-to-text transformer,” J. Mach. Learn. Res. , vol. 21, pp. 140:1–140:67, 2020
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
N. Garcia, M. Otani, C. Chu, and Y. Nakashima, “Knowit VQA: answering knowledge-based questions about videos,” in AAAI . AAAI Press, 2020, pp. 10 826–10 834
2020
Earlier work this paper cites.
D. Guo, H. Wang, H. Zhang, Z. Zha, and M. Wang, “Iterative context-aware graph inference for visual dialog,” in CVPR . Computer Vision Foundation / IEEE, 2020, pp. 10 052–10 061
2020
Earlier work this paper cites.
X. Jiang, S. Du, Z. Qin, Y. Sun, and J. Yu, “KBGN: knowledge-bridge graph network for adaptive vision-text reasoning in visual dialogue,” in ACM Multimedia . ACM, 2020, pp. 1265–1273
2020
Earlier work this paper cites.
J. Chen, F. Lécué, Y. Geng, J. Z. Pan, and H. Chen, “Ontology-guided semantic composition for zero-shot learning,” in KR , 2020, pp. 850–854
2020
Earlier work this paper cites.
J. Chen, L. Pan, Z. Wei, X. Wang, C. Ngo, and T. Chua, “Zero-shot ingredient recognition by multi-relational graph convolutional network,” in AAAI . AAAI Press, 2020, pp. 10 542–10 550
2020
Earlier work this paper cites.
Y. Wang, S. Qian, J. Hu, Q. Fang, and C. Xu, “Fake news detection via knowledge-driven multimodal graph convolutional networks,” in ICMR . ACM, 2020, pp. 540–547
2020
Earlier work this paper cites.
M. Bain, A. Nagrani, A. Brown, and A. Zisserman, “Condensed movies: Story based retrieval with contextual embeddings,” in ACCV (5) , ser. Lecture Notes in Computer Science, vol. 12626. Springer, 2020, pp. 460–479
2020
Earlier work this paper cites.
Q. Huang, Y. Xiong, A. Rao, J. Wang, and D. Lin, “Movienet: A holistic dataset for movie understanding,” in ECCV (4) , ser. Lecture Notes in Computer Science, vol. 12349. Springer, 2020, pp. 709–727
2020
Earlier work this paper cites.
F. Huang, Z. Li, S. Chen, C. Zhang, and H. Ma, “Image captioning with internal and external knowledge,” in CIKM . ACM, 2020, pp. 535–544
2020
Earlier work this paper cites.
J. Hou, X. Wu, X. Zhang, Y. Qi, Y. Jia, and J. Luo, “Joint commonsense and relation reasoning for image and video captioning,” in AAAI . AAAI Press, 2020, pp. 10 973–10 980
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Y. Zhang, X. Wang, Z. Xu, Q. Yu, A. L. Yuille, and D. Xu, “When radiology report generation meets knowledge graph,” in AAAI . AAAI Press, 2020, pp. 12 910–12 917
2020
Earlier work this paper cites.
C. Hsu, Z. Chen, C. Hsu, C. Li, T. Lin, T. K. Huang, and L. Ku, “Knowledge-enriched visual storytelling,” in AAAI . AAAI Press, 2020, pp. 7952–7960
2020
Earlier work this paper cites.
M. Lewis, Y. Liu, N. Goyal, M. Ghazvininejad, A. Mohamed, O. Levy, V. Stoyanov, and L. Zettlemoyer, “BART: denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension,” in ACL . Association for Computational Linguistics, 2020, pp. 7871–7880
2020
Earlier work this paper cites.
J. Cheng, F. Wu, Y. Tian, L. Wang, and D. Tao, “Rifegan: Rich feature generation for text-to-image synthesis from prior knowledge,” in CVPR . Computer Vision Foundation / IEEE, 2020, pp. 10 908–10 917
2020
Earlier work this paper cites.
T. He, L. Gao, J. Song, J. Cai, and Y. Li, “Learning from the scene and borrowing from the rich: Tackling the long tail in scene graph generation,” in IJCAI . ijcai.org, 2020, pp. 587–593
2020
Earlier work this paper cites.
A. Zareian, Z. Wang, H. You, and S. Chang, “Learning visual commonsense for robust scene graph generation,” in ECCV (23) , ser. Lecture Notes in Computer Science, vol. 12368. Springer, 2020, pp. 642–657
2020
Earlier work this paper cites.
A. Zareian, S. Karaman, and S. Chang, “Bridging knowledge graphs to generate scene graphs,” in ECCV (23) , ser. Lecture Notes in Computer Science, vol. 12368. Springer, 2020, pp. 606–623
2020
Earlier work this paper cites.
H. Wang, Y. Zhang, Z. Ji, Y. Pang, and L. Ma, “Consensus-aware visual-semantic embedding for image-text matching,” in ECCV (24) , ser. Lecture Notes in Computer Science, vol. 12369. Springer, 2020, pp. 18–34
2020
Earlier work this paper cites.
P. Wang, D. Liu, H. Li, and Q. Wu, “Give me something to eat: Referring expression comprehension with commonsense knowledge,” in ACM Multimedia . ACM, 2020, pp. 28–36
2020
Earlier work this paper cites.
D. Xu, C. Ruan, E. Körpeoglu, S. Kumar, and K. Achan, “Product knowledge graph embedding for e-commerce,” in WSDM . ACM, 2020, pp. 672–680
2020
Earlier work this paper cites.
T. Safavi and D. Koutra, “Codex: A comprehensive knowledge graph completion benchmark,” in EMNLP (1) . Association for Computational Linguistics, 2020, pp. 8328–8350
2020
Earlier work this paper cites.
Z. Wu, C. Zheng, Y. Cai, J. Chen, H. Leung, and Q. Li, “Multimodal representation with embedded visual guiding objects for named entity recognition in social media posts,” in ACM Multimedia . ACM, 2020, pp. 1038–1046
2020
Earlier work this paper cites.
L. Sun, J. Wang, Y. Su, F. Weng, Y. Sun, Z. Zheng, and Y. Chen, “RIVA: A pre-trained tweet multimodal model based on text-image relation for multimodal NER,” in COLING . International Committee on Computational Linguistics, 2020, pp. 1852–1862
2020
Earlier work this paper cites.
J. Yu, J. Jiang, L. Yang, and R. Xia, “Improving multimodal named entity recognition via entity span detection with unified multimodal transformer,” in ACL . Association for Computational Linguistics, 2020, pp. 3342–3352
2020
Earlier work this paper cites.
M. Li, A. Zareian, Q. Zeng, S. Whitehead, D. Lu, H. Ji, and S. Chang, “Cross-media structured common space for multimedia event extraction,” in ACL . Association for Computational Linguistics, 2020, pp. 2557–2568
2020
Earlier work this paper cites.
L. Chen, Z. Li, Y. Wang, T. Xu, Z. Wang, and E. Chen, “MMEA: entity alignment for multi-modal knowledge graph,” in KSEM (1) , ser. Lecture Notes in Computer Science, vol. 12274. Springer, 2020, pp. 134–147
2020
Earlier work this paper cites.
L. Wu, F. Petroni, M. Josifoski, S. Riedel, and L. Zettlemoyer, “Scalable zero-shot entity linking with dense entity retrieval,” in EMNLP (1) . Association for Computational Linguistics, 2020, pp. 6397–6407
2020
Earlier work this paper cites.
O. Adjali, R. Besançon, O. Ferret, H. L. Borgne, and B. Grau, “Multimodal entity linking for tweets,” in ECIR (1) , ser. Lecture Notes in Computer Science, vol. 12035. Springer, 2020, pp. 463–478
2020
Earlier work this paper cites.
X. Lin, Z. Quan, Z. Wang, T. Ma, and X. Zeng, “KGNN: knowledge graph neural network for drug-drug interaction prediction,” in IJCAI . ijcai.org, 2020, pp. 2739–2745
2020
Earlier work this paper cites.
R. Sun, X. Cao, Y. Zhao, J. Wan, K. Zhou, F. Zhang, Z. Wang, and K. Zheng, “Multi-modal knowledge graphs for recommender systems,” in CIKM . ACM, 2020, pp. 1405–1414
2020
Earlier work this paper cites.
J. Liu, Y. Chen, K. Liu, W. Bi, and X. Liu, “Event extraction as machine reading comprehension,” in EMNLP (1) . Association for Computational Linguistics, 2020, pp. 1641–1651
2020
Earlier work this paper cites.
S. M. Pratt, M. Yatskar, L. Weihs, A. Farhadi, and A. Kembhavi, “Grounded situation recognition,” in ECCV (4) , ser. Lecture Notes in Computer Science, vol. 12349. Springer, 2020, pp. 314–332
2020
Earlier work this paper cites.
A. Tran, A. P. Mathews, and L. Xie, “Transform and tell: Entity-aware news image captioning,” in CVPR . Computer Vision Foundation / IEEE, 2020, pp. 13 032–13 042
2020
Earlier work this paper cites.
B. Kim, T. Hong, Y. Ko, and J. Seo, “Multi-task learning for knowledge graph completion with pre-trained language models,” in COLING . International Committee on Computational Linguistics, 2020, pp. 1737–1743
2020
Earlier work this paper cites.
V. N. Ioannidis, X. Song, S. Manchanda, M. Li, X. Pan, D. Zheng, X. Ning, X. Zeng, and G. Karypis, “Drkg - drug repurposing knowledge graph for covid-19,” https://github.com/gnn4dr/DRKG/ , 2020
2020
Earlier work this paper cites.
Y. Huang, C. Du, Z. Xue, X. Chen, H. Zhao, and L. Huang, “What makes multi-modal learning better than single (provably),” in NeurIPS , 2021, pp. 10 944–10 956
2021
Earlier work this paper cites.
Y. Hu, G. Wen, A. Chapman, P. Yang, M. Luo, Y. Xu, D. Dai, and W. Hall, “Graph-based visual-semantic entanglement network for zero-shot image recognition,” IEEE Transactions on Multimedia , 2021
2021
Earlier work this paper cites.
Y. Geng, J. Chen, Z. Chen, J. Z. Pan, Z. Ye, Z. Yuan, Y. Jia, and H. Chen, “Ontozsl: Ontology-enhanced zero-shot learning,” in WWW . ACM / IW3C2, 2021, pp. 3325–3336
2021
Cited alongside, same era.
K. Marino, X. Chen, D. Parikh, A. Gupta, and M. Rohrbach, “KRISP: integrating implicit and symbolic knowledge for open-domain knowledge-based VQA,” in CVPR . Computer Vision Foundation / IEEE, 2021, pp. 14 111–14 121
2021
Cited alongside, same era.
Z. Chen, J. Chen, Y. Geng, J. Z. Pan, Z. Yuan, and H. Chen, “Zero-shot visual question answering using knowledge graph,” in ISWC , ser. Lecture Notes in Computer Science, vol. 12922. Springer, 2021, pp. 146–162
2021
Cited alongside, same era.
X. Wang, T. Gao, Z. Zhu, Z. Zhang, Z. Liu, J. Li, and J. Tang, “KEPLER: A unified model for knowledge embedding and pre-trained language representation,” Trans. Assoc. Comput. Linguistics , vol. 9, pp. 176–194, 2021
2021
Cited alongside, same era.
2023
Later among the works it cites.
X. He and X. Wang, “Multimodal graph transformer for multimodal question answering,” in EACL . Association for Computational Linguistics, 2023, pp. 189–200
2023
Later among the works it cites.
A. Salaberria, G. Azkune, O. L. de Lacalle, A. Soroa, and E. Agirre, “Image captioning for effective use of language models in knowledge-based visual question answering,” Expert Syst. Appl. , vol. 212, p. 118669, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
F. Liu, M. Chen, D. Roth, and N. Collier, “Visual pivoting for (unsupervised) entity alignment,” in AAAI . AAAI Press, 2021, pp. 4257–4266
2021
Cited alongside, same era.
H. Wen, Y. Lin, T. M. Lai, X. Pan, S. Li, X. Lin, B. Zhou, M. Li, H. Wang, H. Zhang, X. Yu, A. Dong, Z. Wang, Y. R. Fung, P. Mishra, Q. Lyu, D. Surís, B. Chen, S. W. Brown, M. Palmer, C. Callison-Burch, C. Vondrick, J. Han, D. Roth, S. Chang, and H. Ji, “RESIN: A dockerized schema-guided cross-document cross-lingual cross-media information extraction and event tracking system,” in NAACL-HLT (Demonstrations) . Association for Computational Linguistics, 2021, pp. 133–143
2021
Cited alongside, same era.
L. Zhang, S. Liu, D. Liu, P. Zeng, X. Li, J. Song, and L. Gao, “Rich visual knowledge-based augmentation network for visual question answering,” IEEE Trans. Neural Networks Learn. Syst. , vol. 32, no. 10, pp. 4362–4373, 2021
2021
Cited alongside, same era.
W. Zheng, L. Yin, X. Chen, Z. Ma, S. Liu, and B. Yang, “Knowledge base graph embedding module design for visual question answering model,” Pattern Recognit. , vol. 120, p. 108153, 2021
2021
Cited alongside, same era.
P. Vickers, N. Aletras, E. Monti, and L. Barrault, “In factuality: Efficient integration of relevant facts for visual question answering,” in ACL/IJCNLP (2) . Association for Computational Linguistics, 2021, pp. 468–475
2021
Cited alongside, same era.
A. Jain, M. Kothyari, V. Kumar, P. Jyothi, G. Ramakrishnan, and S. Chakrabarti, “Select, substitute, search: A new benchmark for knowledge-augmented visual question answering,” in SIGIR . ACM, 2021, pp. 2491–2498
2021
Cited alongside, same era.
2021
Cited alongside, same era.
M. Luo, Y. Zeng, P. Banerjee, and C. Baral, “Weakly-supervised visual-retriever-reader for knowledge-based question answering,” in EMNLP (1) . Association for Computational Linguistics, 2021, pp. 6417–6431
2021
Cited alongside, same era.
Q. Si, Y. Mo, Z. Lin, H. Ji, and W. Wang, “Combo of thinking and observing for outside-knowledge VQA,” in ACL (1) . Association for Computational Linguistics, 2023, pp. 10 959–10 975
2023
Later among the works it cites.
O. Adjali, P. Grimal, O. Ferret, S. Ghannay, and H. L. Borgne, “Explicit knowledge integration for knowledge-aware visual question answering about named entities,” in ICMR . ACM, 2023, pp. 29–38
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Wu, G. Zhao, and X. Qian, “Resolving zero-shot and fact-based visual question answering via enhanced fact retrieval,” IEEE Transactions on Multimedia , 2023
2023
Later among the works it cites.
A. Salemi, J. A. Pizzorno, and H. Zamani, “A symmetric dual encoding dense retrieval framework for knowledge-intensive visual question answering,” in SIGIR . ACM, 2023, pp. 110–120
2023
Later among the works it cites.
Z. Hu, A. Iscen, C. Sun, Z. Wang, K. Chang, Y. Sun, C. Schmid, D. A. Ross, and A. Fathi, “Reveal: Retrieval-augmented visual-language pre-training with multi-source multimodal knowledge memory,” in CVPR . IEEE, 2023, pp. 23 369–23 379
2023
Later among the works it cites.
2023
Later among the works it cites.
Z. Shao, Z. Yu, M. Wang, and J. Yu, “Prompting large language models with answer heuristics for knowledge-based visual question answering,” in CVPR . IEEE, 2023, pp. 14 974–14 983
2023
Later among the works it cites.
S. Subramanian, M. Narasimhan, K. Khangaonkar, K. Yang, A. Nagrani, C. Schmid, A. Zeng, T. Darrell, and D. Klein, “Modular visual question answering via code generation,” in ACL (2) . Association for Computational Linguistics, 2023, pp. 747–761
2023
Later among the works it cites.
A. Chevalier, A. Wettig, A. Ajith, and D. Chen, “Adapting language models to compress contexts,” in EMNLP . Association for Computational Linguistics, 2023, pp. 3829–3846
2023
Later among the works it cites.
Z. Yang, X. Du, E. Cambria, and C. Cardie, “End-to-end case-based reasoning for commonsense knowledge base completion,” in EACL . Association for Computational Linguistics, 2023, pp. 3491–3504
2023
Later among the works it cites.
2023
Later among the works it cites.
R. Yu, C. Pan, X. Fei, M. Chen, and D. Shen, “Multi-graph attention networks with bilinear convolution for diagnosis of schizophrenia,” IEEE J. Biomed. Health Informatics , vol. 27, no. 3, pp. 1443–1454, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
E. Kamalloo, N. Dziri, C. L. A. Clarke, and D. Rafiei, “Evaluating open-domain question answering in the era of large language models,” in ACL (1) . Association for Computational Linguistics, 2023, pp. 5591–5606
2023
Later among the works it cites.
2023
Later among the works it cites.
W. Lin, Z. Wang, and B. Byrne, “FVQA 2.0: Introducing adversarial samples into fact-based visual question answering,” in EACL (Findings) . Association for Computational Linguistics, 2023, pp. 149–157
2023
Later among the works it cites.
B. Z. Reichman, A. Sundar, C. Richardson, T. Zubatiy, P. Chowdhury, A. Shah, J. Truxal, M. Grimes, D. Shah, W. J. Chee, S. Punjwani, A. Jain, and L. Heck, “Outside knowledge visual question answering version 2.0,” in ICASSP . IEEE, 2023, pp. 1–5
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Ye, Y. Xie, D. Chen, Y. Xu, L. Yuan, C. Zhu, and J. Liao, “Improving commonsense in vision-language models via knowledge graph riddles,” in CVPR . IEEE, 2023, pp. 2634–2645
2023
Later among the works it cites.
J. Gao, Q. Wu, A. Blair, and M. Pagnucco, “Lora: A logical reasoning augmented dataset for visual question answering,” in Thirty-seventh Conference on Neural Information Processing Systems Datasets and Benchmarks Track , 2023
2023
Later among the works it cites.
S. Tan, M. Ge, D. Guo, H. Liu, and F. Sun, “Knowledge-based embodied question answering,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 45, no. 10, pp. 11 948–11 960, 2023
2023
Later among the works it cites.
J. Chen, Z. Guo, J. Xie, Y. Cai, and Q. Li, “Deconfounded visual question generation with causal inference,” in ACM Multimedia . ACM, 2023, pp. 5132–5142
2023
Later among the works it cites.
A. Salemi, M. Rafiee, and H. Zamani, “Pre-training multi-modal dense retrievers for outside-knowledge visual question answering,” in ICTIR . ACM, 2023, pp. 169–176
2023
Later among the works it cites.
K. Uehara and T. Harada, “K-VQG: knowledge-aware visual question generation for common-sense acquisition,” in WACV . IEEE, 2023, pp. 4390–4398
2023
Later among the works it cites.
L. Zhao, J. Li, L. Gao, Y. Rao, J. Song, and H. T. Shen, “Heterogeneous knowledge network for visual dialog,” IEEE Trans. Circuits Syst. Video Technol. , vol. 33, no. 2, pp. 861–871, 2023
2023
Later among the works it cites.
Z. Zhang, Y. Ji, and C. Liu, “Knowledge-aware causal inference network for visual dialog,” in ICMR . ACM, 2023, pp. 253–261
2023
Later among the works it cites.
A.-A. Liu, C. Huang, N. Xu, H. Tian, J. Liu, and Y. Zhang, “Counterfactual visual dialog: Robust commonsense knowledge learning from unbiased training,” IEEE Transactions on Multimedia , 2023
2023
Later among the works it cites.
Y. Geng, J. Chen, X. Zhuang, Z. Chen, J. Z. Pan, J. Li, Z. Yuan, and H. Chen, “Benchmarking knowledge-driven zero-shot learning,” J. Web Semant. , vol. 75, p. 100757, 2023
2023
Later among the works it cites.
Z. Chen, Y. Huang, J. Chen, Y. Geng, W. Zhang, Y. Fang, J. Z. Pan, and H. Chen, “DUET: cross-modal semantic grounding for contrastive zero-shot learning,” in AAAI . AAAI Press, 2023, pp. 405–413
2023
Later among the works it cites.
L. Wu, Z. Li, H. Zhao, Z. Wang, Q. Liu, B. Huai, N. J. Yuan, and E. Chen, “Recognizing unseen objects via multimodal intensive knowledge graph propagation,” in Proceedings of the 29th ACM SIGKDD Conference on Knowledge Discovery and Data Mining, KDD 2023, Long Beach, CA, USA, August 6-10, 2023 , A. K. Singh, Y. Sun, L. Akoglu, D. Gunopulos, X. Yan, R. Kumar, F. Ozcan, and J. Ye, Eds. ACM, 2023, pp. 2618–2628
2023
Later among the works it cites.
M. Sun, X. Zhang, J. Ma, S. Xie, Y. Liu, and P. S. Yu, “Inconsistent matters: A knowledge-guided dual-consistency network for multi-modal rumor detection,” IEEE Trans. Knowl. Data Eng. , vol. 35, no. 12, pp. 12 736–12 749, 2023
2023
Later among the works it cites.
Y. Zhang, X. Su, J. Wu, J. Yang, H. Fan, and X. Zheng, “Emoknow: Emotion- and knowledge-oriented model for COVID-19 fake news detection,” in ADMA (1) , ser. Lecture Notes in Computer Science, vol. 14176. Springer, 2023, pp. 352–367
2023
Later among the works it cites.
J. Li, G. Qi, C. Zhang, Y. Chen, Y. Tan, C. Xia, and Y. Tian, “Incorporating domain knowledge graph into multimodal movie genre classification with self-supervised attention and contrastive learning,” in ACM Multimedia . ACM, 2023, pp. 3337–3345
2023
Later among the works it cites.
J. Zhong and D. Wang, “Image caption generation based on object detection and knowledge enhancement,” in International Conference on Image, Signal Processing, and Pattern Recognition (ISPP 2023) , vol. 12707. SPIE, 2023, pp. 226–232
2023
Later among the works it cites.
T. Li, H. Wang, B. He, and C. W. Chen, “Knowledge-enriched attention network with group-wise semantic for visual storytelling,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 45, no. 7, pp. 8634–8645, 2023
2023
Later among the works it cites.
A.-A. Liu, Z. Sun, N. Xu, R. Kang, J. Cao, F. Yang, W. Qin, S. Zhang, J. Zhang, and X. Li, “Prior knowledge guided text to image generation,” Pattern Recognition Letters , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
X. Chang, P. Ren, P. Xu, Z. Li, X. Chen, and A. Hauptmann, “A comprehensive survey of scene graphs: Generation and application,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 45, no. 1, pp. 1–26, 2023
2023
Later among the works it cites.
Z. Chen, S. Rezayi, and S. Li, “More knowledge, less bias: Unbiasing scene graph generation with explicit ontological adjustment,” in WACV . IEEE, 2023, pp. 4012–4021
2023
Later among the works it cites.
J. Lu, L. Chen, Y. Song, S. Lin, C. Wang, and G. He, “Prior knowledge-driven dynamic scene graph generation with causal inference,” in ACM Multimedia . ACM, 2023, pp. 4877–4885
2023
Later among the works it cites.
H. Tian, N. Xu, Y. Wang, C. Yan, B. Zheng, X. Li, and A. Liu, “Towards confidence-aware commonsense knowledge integration for scene graph generation,” in ICME . IEEE, 2023, pp. 2255–2260
2023
Later among the works it cites.
W. Li, S. Yang, Q. Li, X. Li, and A.-A. Liu, “Commonsense-guided semantic and relational consistencies for image-text retrieval,” IEEE Transactions on Multimedia , 2023
2023
Later among the works it cites.
S. Yang, Q. Li, W. Li, M. Liu, X. Li, and A. Liu, “External knowledge dynamic modeling for image-text retrieval,” in ACM Multimedia . ACM, 2023, pp. 5330–5338
2023
Later among the works it cites.
D. Feng, X. He, and Y. Peng, “MKVSE: multimodal knowledge enhanced visual-semantic embedding for image-text retrieval,” ACM Trans. Multim. Comput. Commun. Appl. , vol. 19, no. 5, pp. 162:1–162:21, 2023
2023
Later among the works it cites.
X. Dong, X. Zhan, Y. Wei, X. Wei, Y. Wang, M. Lu, X. Cao, and X. Liang, “Entity-graph enhanced cross-modal pretraining for instance-level product retrieval,” IEEE Trans. Pattern Anal. Mach. Intell. , vol. 45, no. 11, pp. 13 117–13 133, 2023
2023
Later among the works it cites.
X. Wang, L. Li, Z. Li, X. Wang, X. Zhu, C. Wang, J. Huang, and Y. Xiao, “AGREE: aligning cross-modal entities for image-text retrieval upon vision-language pre-trained models,” in WSDM . ACM, 2023, pp. 456–464
2023
Later among the works it cites.
W. Tang, L. Li, X. Liu, L. Jin, J. Tang, and Z. Li, “Context disentangling and prototype inheriting for robust visual grounding,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2023
2023
Later among the works it cites.
Z. Zhang, H. Yannakoudakis, X. Zhen, and E. Shutova, “Ck-transformer: Commonsense knowledge enhanced transformers for referring expression comprehension,” in EACL (Findings) . Association for Computational Linguistics, 2023, pp. 2541–2551
2023
Later among the works it cites.
Y. Bu, X. Wu, L. Li, Y. Cai, Q. Liu, and Q. Huang, “Segment-level and category-oriented network for knowledge-based referring expression comprehension,” in ACL (Findings) . Association for Computational Linguistics, 2023, pp. 8745–8757
2023
Later among the works it cites.
Z. Chen, R. Zhang, Y. Song, X. Wan, and G. Li, “Advancing visual grounding with scene knowledge: Benchmark and method,” in CVPR . IEEE, 2023, pp. 15 039–15 049
2023
Later among the works it cites.
W. Zhang, Y. Zhu, M. Chen, Y. Geng, Y. Huang, Y. Xu, W. Song, and H. Chen, “Structure pretraining and prompt tuning for knowledge graph transfer,” in WWW . ACM, 2023, pp. 2581–2590
2023
Later among the works it cites.
H. Liu, C. Li, Q. Wu, and Y. J. Lee, “Visual instruction tuning,” CoRR , vol. abs/2304.08485, 2023
2023
Later among the works it cites.
P. Wang, X. Chen, Z. Shang, and W. Ke, “Multimodal named entity recognition with bottleneck fusion and contrastive learning,” IEICE Trans. Inf. Syst. , vol. 106, no. 4, pp. 545–555, 2023
2023
Later among the works it cites.
Z. Zhang, J. Chen, X. Liu, W. Mai, and Q. Cai, “‘what’and ‘where’both matter: dual cross-modal graph convolutional networks for multimodal named entity recognition,” International Journal of Machine Learning and Cybernetics , pp. 1–11, 2023
2023
Later among the works it cites.
X. Bao, M. Tian, Z. Zha, and B. Qin, “MPMRC-MNER: A unified MRC framework for multimodal named entity recognition based multimodal prompt,” in CIKM . ACM, 2023, pp. 47–56
2023
Later among the works it cites.
M. Jia, L. Shen, X. Shen, L. Liao, M. Chen, X. He, Z. Chen, and J. Li, “MNER-QG: an end-to-end MRC framework for multimodal named entity recognition with query grounding,” in AAAI . AAAI Press, 2023, pp. 8032–8040
2023
Later among the works it cites.
X. Zhang, J. Yuan, L. Li, and J. Liu, “Reducing the bias of visual objects in multimodal named entity recognition,” in WSDM . ACM, 2023, pp. 958–966
2023
Later among the works it cites.
W. Mai, Z. Zhang, K. Li, Y. Xue, and F. Li, “Dynamic graph construction framework for multimodal named entity recognition in social media,” IEEE Transactions on Computational Social Systems , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
A. Guo, X. Zhao, Z. Tan, and W. Xiao, “MGICL: multi-grained interaction contrastive learning for multimodal named entity recognition,” in CIKM . ACM, 2023, pp. 639–648
2023
Later among the works it cites.
J. Li, H. Li, Z. Pan, D. Sun, J. Wang, W. Zhang, and G. Pan, “Prompting chatgpt in MNER: enhanced multimodal named entity recognition with auxiliary refined knowledge,” in EMNLP (Findings) . Association for Computational Linguistics, 2023, pp. 2787–2802
2023
Later among the works it cites.
X. Hu, J. Chen, A. Liu, S. Meng, L. Wen, and P. S. Yu, “Prompt me up: Unleashing the power of alignments for multimodal entity and relation extraction,” in ACM Multimedia . ACM, 2023, pp. 5185–5194
2023
Later among the works it cites.
2023
Later among the works it cites.
L. Li, X. Chen, S. Qiao, F. Xiong, H. Chen, and N. Zhang, “On analyzing the role of image for visual-enhanced relation extraction (student abstract),” in AAAI . AAAI Press, 2023, pp. 16 254–16 255
2023
Later among the works it cites.
S. Wu, H. Fei, Y. Cao, L. Bing, and T. Chua, “Information screening whilst exploiting! multimodal relation extraction with feature denoising and multimodal topic modeling,” in ACL (1) . Association for Computational Linguistics, 2023, pp. 14 734–14 751
2023
Later among the works it cites.
Q. Li, S. Guo, C. Ji, X. Peng, S. Cui, J. Li, and L. Wang, “Dual-gated fusion with prefix-tuning for multi-modal relation extraction,” in ACL (Findings) . Association for Computational Linguistics, 2023, pp. 8982–8994
2023
Later among the works it cites.
X. Hu, Z. Guo, Z. Teng, I. King, and P. S. Yu, “Multimodal relation extraction with cross-modal retrieval and synthesis,” in ACL (2) . Association for Computational Linguistics, 2023, pp. 303–311
2023
Later among the works it cites.
Z. Wang and J. Shang, “Towards zero-shot relation extraction in web mining: A multimodal approach with relative XML path,” in EMNLP (Findings) . Association for Computational Linguistics, 2023, pp. 4254–4265
2023
Later among the works it cites.
Z. Du, Y. Li, X. Guo, Y. Sun, and B. Li, “Training multimedia event extraction with generated images and captions,” in ACM Multimedia . ACM, 2023, pp. 5504–5513
2023
Later among the works it cites.
S. Wang, M. Ju, Y. Zhang, Y. Zheng, M. Wang, and G. Qi, “Cross-modal contrastive learning for event extraction,” in DASFAA (3) , ser. Lecture Notes in Computer Science, vol. 13945. Springer, 2023, pp. 699–715
2023
Later among the works it cites.
J. Li, C. Zhang, M. Du, D. Min, Y. Chen, and G. Qi, “Three stream based multi-level event contrastive learning for text-video event extraction,” in EMNLP . Association for Computational Linguistics, 2023
2023
Later among the works it cites.
F. Moghimifar, F. Shiri, R. Haffari, Y. Li, and V. Nguyen, “Few-shot domain-adaptative visually-fused event detection from text,” in FUSION . IEEE, 2023, pp. 1–8
2023
Later among the works it cites.
Q. Li, S. Guo, Y. Luo, C. Ji, L. Wang, J. Sheng, and J. Li, “Attribute-consistent knowledge graph representation learning for multi-modal entity alignment,” in WWW . ACM, 2023, pp. 2499–2508
2023
Later among the works it cites.
F. Su, C. Xu, H. Yang, Z. Chen, and N. Jing, “Neural entity alignment with cross-modal supervision,” Inf. Process. Manag. , vol. 60, no. 2, p. 103174, 2023
2023
Later among the works it cites.
J. Li, Q. Zhou, W. Chen, and L. Zhao, “Enhanced entity interaction modeling for multi-modal entity alignment,” in KSEM (2) , ser. Lecture Notes in Computer Science, vol. 14118. Springer, 2023, pp. 214–227
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
B. Xu, C. Xu, and B. Su, “Cross-modal graph attention network for entity alignment,” in ACM Multimedia . ACM, 2023, pp. 3715–3723
2023
Later among the works it cites.
J. Zhu, C. Huang, and P. D. Meo, “DFMKE: A dual fusion multi-modal knowledge graph embedding framework for entity alignment,” Inf. Fusion , vol. 90, pp. 111–119, 2023
2023
Later among the works it cites.
M. Wang, Y. Shi, H. Yang, Z. Zhang, Z. Lin, and Y. Zheng, “Probing the impacts of visual context in multimodal entity alignment,” Data Sci. Eng. , vol. 8, no. 2, pp. 124–134, 2023
2023
Later among the works it cites.
Z. Chen, J. Chen, W. Zhang, L. Guo, Y. Fang, Y. Huang, Y. Zhang, Y. Geng, J. Z. Pan, W. Song, and H. Chen, “Meaformer: Multi-modal entity alignment transformer for meta modality hybrid,” in ACM Multimedia . ACM, 2023, pp. 3317–3327
2023
Later among the works it cites.
2023
Later among the works it cites.
W. Ni, Q. Xu, Y. Jiang, Z. Cao, X. Cao, and Q. Huang, “PSNEA: pseudo-siamese network for entity alignment between multi-modal knowledge graphs,” in ACM Multimedia . ACM, 2023, pp. 3489–3497
2023
Later among the works it cites.
C. Yang, B. He, Y. Wu, C. Xing, L. He, and C. Ma, “MMEL: A joint learning framework for multi-mention entity linking,” in UAI , ser. Proceedings of Machine Learning Research, vol. 216. PMLR, 2023, pp. 2411–2421
2023
Later among the works it cites.
P. Luo, T. Xu, S. Wu, C. Zhu, L. Xu, and E. Chen, “Multi-grained multimodal interaction network for entity linking,” in KDD . ACM, 2023, pp. 1583–1594
2023
Later among the works it cites.
S. Wang, A. H. Li, H. Zhu, S. Zhang, P. Perera, C. Hang, J. Ma, W. Y. Wang, Z. Wang, V. Castelli, B. Xiang, and P. Ng, “Benchmarking diverse-modal entity linking with generative models,” in Findings of ACL . Association for Computational Linguistics, 2023, pp. 7841–7857
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
X. Li, X. Zhao, J. Xu, Y. Zhang, and C. Xing, “IMF: interactive multimodal fusion model for link prediction,” in WWW . ACM, 2023, pp. 2572–2580
2023
Later among the works it cites.
D. Xu, J. Zhou, T. Xu, Y. Xia, J. Liu, E. Chen, and D. Dou, “Multimodal biological knowledge graph completion via triple co-attention mechanism,” in ICDE . IEEE, 2023, pp. 3928–3941
2023
Later among the works it cites.
Q. Fang, X. Zhang, J. Hu, X. Wu, and C. Xu, “Contrastive multi-modal knowledge graph representation learning,” IEEE Trans. Knowl. Data Eng. , vol. 35, no. 9, pp. 8983–8996, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Liang, A. Zhu, J. Zhang, and J. Shao, “Hyper-node relational graph attention network for multi-modal knowledge graph completion,” ACM Trans. Multim. Comput. Commun. Appl. , vol. 19, no. 2, pp. 62:1–62:21, 2023
2023
Later among the works it cites.
S. Zheng, W. Wang, J. Qu, H. Yin, W. Chen, and L. Zhao, “MMKGR: multi-hop multi-modal knowledge graph reasoning,” in ICDE . IEEE, 2023, pp. 96–109
2023
Later among the works it cites.
2023
Later among the works it cites.
N. Zhang, L. Li, X. Chen, X. Liang, S. Deng, and H. Chen, “Multimodal analogical reasoning over knowledge graphs,” in ICLR . OpenReview.net, 2023
2023
Later among the works it cites.
Y. Zeng, Q. Jin, T. Bao, and W. Li, “Multi-modal knowledge hypergraph for diverse image retrieval,” in AAAI . AAAI Press, 2023, pp. 3376–3383
2023
Later among the works it cites.
L. Jin and J. Chen, “Self-supervised opinion summarization with multi-modal knowledge graph,” Journal of Intelligent Information Systems , pp. 1–18, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Fang, Q. Zhang, N. Zhang, Z. Chen, X. Zhuang, X. Shao, X. Fan, and H. Chen, “Knowledge graph-enhanced molecular contrastive learning with functional prompt,” Nature Machine Intelligence , pp. 1–12, 2023
2023
Later among the works it cites.
D. Xu, J. Zhou, T. Xu, Y. Xia, J. Liu, E. Chen, and D. Dou, “Multimodal biological knowledge graph completion via triple co-attention mechanism,” in ICDE . IEEE, 2023, pp. 3928–3941
2023
Later among the works it cites.
S. Cheng, X. Liang, Z. Bi, H. Chen, and N. Zhang, “Multi-modal protein knowledge graph construction and applications (student abstract),” in AAAI . AAAI Press, 2023, pp. 16 190–16 191
2023
Later among the works it cites.
X. Wang, C. Wang, L. Li, Z. Li, B. Chen, L. Jin, J. Huang, Y. Xiao, and M. Gao, “Fashionklip: Enhancing e-commerce image-text retrieval with fashion multi-modal conceptual knowledge graph,” in ACL (industry) . Association for Computational Linguistics, 2023, pp. 149–158
2023
Later among the works it cites.
C. Sun, W. Chen, L. Lin, and L. Shan, “Enhancing recommender system with multi-modal knowledge graph,” in PRCV (1) , ser. Lecture Notes in Computer Science, vol. 14425. Springer, 2023, pp. 395–407
2023
Later among the works it cites.
Q. Fang, X. Zhang, J. Hu, X. Wu, and C. Xu, “Contrastive multi-modal knowledge graph representation learning,” IEEE Trans. Knowl. Data Eng. , vol. 35, no. 9, pp. 8983–8996, 2023
2023
Later among the works it cites.
Y. Wei, W. Chen, S. Wen, A. Liu, and L. Zhao, “Knowledge graph incremental embedding for unseen modalities,” Knowl. Inf. Syst. , vol. 65, no. 9, pp. 3611–3631, 2023
2023
Later among the works it cites.
Y. Chen, X. Ge, S. Yang, L. Hu, J. Li, and J. Zhang, “A survey on multimodal knowledge graphs: Construction, completion and applications,” Mathematics , vol. 11, no. 8, p. 1815, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
J. Wang, Z. Li, J. Yu, L. Yang, and R. Xia, “Fine-grained multimodal named entity recognition and grounding with a generative framework,” in ACM Multimedia . ACM, 2023, pp. 3934–3943
2023
Later among the works it cites.
L. Yuan, Y. Cai, J. Wang, and Q. Li, “Joint multimodal entity-relation extraction based on edge-enhanced graph alignment network and word-pair relation tagging,” in AAAI . AAAI Press, 2023, pp. 11 051–11 059
2023
Later among the works it cites.
2023
Later among the works it cites.
J. Li, H. Li, Z. Pan, D. Sun, J. Wang, W. Zhang, and G. Pan, “Prompting chatgpt in MNER: enhanced multimodal named entity recognition with auxiliary refined knowledge,” pp. 2787–2802, 2023
2023
Later among the works it cites.
Z. Sun, W. Hu, C. Wang, Y. Wang, and Y. Qu, “Revisiting embedding-based entity alignment: A robust and adaptive method,” IEEE Trans. Knowl. Data Eng. , vol. 35, no. 8, pp. 8461–8475, 2023
2023
Later among the works it cites.
X. Liu, J. Wu, T. Li, L. Chen, and Y. Gao, “Unsupervised entity alignment for temporal knowledge graphs,” in WWW . ACM, 2023, pp. 2528–2538
2023
Later among the works it cites.
Z. Sun, J. Huang, X. Xu, Q. Chen, W. Ren, and W. Hu, “What makes entities similar? A similarity flooding perspective for multi-sourced knowledge graph embeddings,” in ICML , ser. Proceedings of Machine Learning Research, vol. 202. PMLR, 2023, pp. 32 875–32 885
2023
Later among the works it cites.
B. Zhu, M. Wu, Y. Hong, Y. Chen, B. Xie, F. Liu, C. Bu, and W. Ding, “MMIEA: multi-modal interaction entity alignment model for knowledge graphs,” Inf. Fusion , vol. 100, p. 101935, 2023
2023
Later among the works it cites.
B. Liu, T. Lan, W. Hua, and G. Zuccon, “Dependency-aware self-training for entity alignment,” in WSDM . ACM, 2023, pp. 796–804
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Xu, S. Liu, T. Culhane, E. Pertseva, M. Wu, S. J. Semnani, and M. S. Lam, “Fine-tuned llms know more, hallucinate less with few-shot sequence-to-sequence semantic parsing over wikidata,” in EMNLP . Association for Computational Linguistics, 2023, pp. 5778–5791
2023
Later among the works it cites.
2023
Later among the works it cites.
C. Chen, Y. Wang, A. Sun, B. Li, and K. Lam, “Dipping plms sauce: Bridging structure and text for effective knowledge graph completion via conditional soft prompting,” in ACL (Findings) . Association for Computational Linguistics, 2023, pp. 11 489–11 503
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Y. Zhang, Z. Chen, and W. Zhang, “MACO: A modality adversarial and contrastive framework for modality-missing multi-modal knowledge graph completion,” in NLPCC (1) , ser. Lecture Notes in Computer Science, vol. 14302. Springer, 2023, pp. 123–134
2023
Later among the works it cites.
2023
Later among the works it cites.
P. Lo Surdo, M. Iannuccelli, S. Contino, L. Castagnoli, L. Licata, G. Cesareni, and L. Perfetto, “Signor 3.0, the signaling network open resource 3.0: 2022 update,” Nucleic Acids Research , vol. 51, no. D1, pp. D631–D637, 2023
2023
Later among the works it cites.
X. Huang, Y.-J. Huang, Y. Zhang, W. Tian, R. Feng, Y. Zhang, Y. Xie, Y. Li, and L. Zhang, “Open-set image tagging with multi-grained text supervision,” arXiv e-prints , pp. arXiv–2310, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
X. Liu, J. Wu, T. Li, L. Chen, and Y. Gao, “Unsupervised entity alignment for temporal knowledge graphs,” in WWW . ACM, 2023, pp. 2528–2538
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
F. Wan, X. Huang, T. Yang, X. Quan, W. Bi, and S. Shi, “Explore-instruct: Enhancing domain-specific instruction coverage through active exploration,” in EMNLP . Association for Computational Linguistics, 2023, pp. 9435–9454
2023
Later among the works it cites.
Y. Wang, Y. Kordi, S. Mishra, A. Liu, N. A. Smith, D. Khashabi, and H. Hajishirzi, “Self-instruct: Aligning language models with self-generated instructions,” in ACL (1) . Association for Computational Linguistics, 2023, pp. 13 484–13 508
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
W. Yu, D. Iter, S. Wang, Y. Xu, M. Ju, S. Sanyal, C. Zhu, M. Zeng, and M. Jiang, “Generate rather than retrieve: Large language models are strong context generators,” in ICLR . OpenReview.net, 2023
2023
Later among the works it cites.
S. Cheng, B. Tian, Q. Liu, X. Chen, Y. Wang, H. Chen, and N. Zhang, “Can we edit multimodal large language models?” in EMNLP . Association for Computational Linguistics, 2023, pp. 13 877–13 888
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
R. Zhang, Y. Li, Y. Ma, M. Zhou, and L. Zou, “Llmaaa: Making large language models as active annotators,” in EMNLP (Findings) . Association for Computational Linguistics, 2023, pp. 13 088–13 103
2023
Later among the works it cites.
A. A. Ismail, S. O. Arik, J. Yoon, A. Taly, S. Feizi, and T. Pfister, “Interpretable mixture of experts,” Transactions on Machine Learning Research , 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
L.-M. Team, “Llama-moe: Building mixture-of-experts from llama with continual pre-training,” Dec 2023. [Online]. Available: https://github.com/pjlab-sys4nlp/llama-moe
2023
Later among the works it cites.
V. Pahuja, W. Luo, Y. Gu, C. Tu, H. Chen, T. Y. Berger-Wolf, C. V. Stewart, S. Gao, W. Chao, and Y. Su, “Bringing back the context: Camera trap species identification as link prediction on multimodal knowledge graphs,” 2024
2024
Closest in time.
J. Dong, Q. Zhang, H. Zhou, D. Zha, P. Zheng, and X. Huang, “Modality-aware integration with large language models for knowledge-based visual question answering,” 2024
2024
Closest in time.
P. Liu, G. Wang, H. Li, J. Liu, Y. Ren, H. Zhu, and L. Sun, “Multi-granularity cross-modal representation learning for named entity recognition on social media,” Inf. Process. Manag. , vol. 61, no. 1, p. 103546, 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
D. Zhang, Y. Yu, C. Li, J. Dong, D. Su, C. Chu, and D. Yu, “Mm-llms: Recent advances in multimodal large language models,” 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
A. Mishra, A. Asai, V. Balachandran, Y. Wang, G. Neubig, Y. Tsvetkov, and H. Hajishirzi, “Fine-grained hallucination detection and editing for language models,” 2024
2024
Closest in time.
K. Lu, B. Yu, C. Zhou, and J. Zhou, “Large language models are superpositions of all characters: Attaining arbitrary role-play via self-alignment,” 2024
2024
Closest in time.
S. Qiao, N. Zhang, R. Fang, Y. Luo, W. Zhou, Y. E. Jiang, C. Lv, and H. Chen, “AUTOACT: automatic agent learning from scratch via self-planning,” 2024
2024
Closest in time.
D. Mondal, S. Modi, S. Panda, R. Singh, and G. S. Rao, “Kam-cot: Knowledge augmented multimodal chain-of-thoughts reasoning,” 2024
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Z. Qi, Z. Zhang, J. Chen, X. Chen, Y. Xiang, N. Zhang, and Y. Zheng, “Unsupervised knowledge graph alignment by probabilistic reasoning and semantic embedding,” in IJCAI , 2021, pp. 2019–2025
2025
Closest in time.
X. Chen, A. Shrivastava, and A. Gupta, “Enriching visual knowledge bases via object discovery and segmentation,” in CVPR . IEEE Computer Society, 2014, pp. 2035–2042
2042
Closest in time.
J. Lu, D. Zhang, J. Zhang, and P. Zhang, “Flat multi-modal interaction transformer for named entity recognition,” in COLING . International Committee on Computational Linguistics, 2022, pp. 2055–2064
2064
Closest in time.
Y. Guo, L. Nie, Y. Wong, Y. Liu, Z. Cheng, and M. S. Kankanhalli, “A unified end-to-end retriever-reader framework for knowledge-based VQA,” in ACM Multimedia . ACM, 2022, pp. 2061–2069
2069
Closest in time.