Fetching the paper…
Reading the bibliography…
With massive explosion of social media such as Twitter and Instagram, people daily share billions of multimedia posts, containing images and text.
D. Lu, L. Neves, V. Carvalho, N. Zhang, and H. Ji, “Visual attention model for name tagging in multimodal social media,” in Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , vol. 1, 2018, pp. 1990–1999
1999
Earlier work this paper cites.
J. D. Lafferty, A. McCallum, and F. C. N. Pereira, “Conditional random fields: Probabilistic models for segmenting and labeling sequence data,” pp. 282–289, 2001
2001
Earlier work this paper cites.
C. W. Leong and R. Mihalcea, “Going beyond text: A hybrid image-text approach for measuring word relatedness,” in International Joint Conference on Natural Language Processing , 2011, pp. 1403–1407
2011
Earlier work this paper cites.
D. Kiela and L. Bottou, “Learning image embeddings using convolutional neural networks for improved multi-modal semantics,” in Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) , 2014, pp. 36–45
2014
Earlier work this paper cites.
T. Baldwin, M.-C. de Marneffe, B. Han, Y.-B. Kim, A. Ritter, and W. Xu, “Shared tasks of the 2015 workshop on noisy user-generated text: Twitter lexical normalization and named entity recognition,” in Proceedings of the Workshop on Noisy User-generated Text , 2015, pp. 126–135
2015
Earlier work this paper cites.
L. Wang, Y. Li, and S. Lazebnik, “Learning deep structure-preserving image-text embeddings,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, pp. 5005–5013
2016
Earlier work this paper cites.
A. Fukui, D. H. Park, D. Yang, A. Rohrbach, T. Darrell, and M. Rohrbach, “Multimodal compact bilinear pooling for visual question answering and visual grounding,” Proceedings of Empirical Methods in Natural Language Processing, EMNLP 2016 , pp. 457–468, 2016
2016
Earlier work this paper cites.
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2017
Cited alongside, same era.
I. Gallo, A. Calefati, and S. Nawaz, “Multimodal classification fusion in real-world scenarios,” in Document Analysis and Recognition (ICDAR) , vol. 5. IEEE, 2017, pp. 36–41
2017
Cited alongside, same era.
2018
Later among the works it cites.
P. Anderson, X. He, C. Buehler, D. Teney, M. Johnson, S. Gould, and L. Zhang, “Bottom-up and top-down attention for image captioning and visual question answering,” in CVPR , vol. 3, no. 5, 2018, p. 6
2018
Later among the works it cites.
Q. Zhang, J. Fu, X. Liu, and X. Huang, “Adaptive co-attention network for named entity recognition in tweets.” in AAAI , 2018
2018
Later among the works it cites.
S. Moon, L. Neves, and V. Carvalho, “Multimodal named entity recognition for short social media posts,” in Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers . Association for Computational Linguistics, 2018, pp. 852–860
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Nam, J.-W. Ha, and J. Kim, “Dual attention networks for multimodal reasoning and matching,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 299–307
2017
Cited alongside, same era.
G. Aguilar, S. Maharjan, A. P. L. Monroy, and T. Solorio, “A multi-task approach for named entity recognition in social media data,” in Proceedings of Noisy User-generated Text , 2017, pp. 148–153
2017
Cited alongside, same era.
D. Kiela, E. Grave, A. Joulin, and T. Mikolov, “Efficient large-scale multi-modal classification,” Proceedings of the Thirty-Second AAAI Conference on Artificial Intelligence , pp. 5198–5204, 2018
2018
Cited alongside, same era.
T. Shen, T. Zhou, G. Long, J. Jiang, S. Pan, and C. Zhang, “Disan: Directional self-attention network for rnn/cnn-free language understanding,” AAAI Conference on Artificial Intelligence , pp. 5446–5455, 2018
2018
Later among the works it cites.
E. Grave, P. Bojanowski, P. Gupta, A. Joulin, and T. Mikolov, “Learning word vectors for 157 languages,” in Proceedings of the International Conference on Language Resources and Evaluation (LREC 2018) , 2018
2018
Later among the works it cites.
K. Xu, J. Ba, R. Kiros, K. Cho, A. Courville, R. Salakhudinov, R. Zemel, and Y. Bengio, “Show, attend and tell: Neural image caption generation with visual attention,” in ICML , 2015, pp. 2048–2057
2057
Closest in time.