Fetching the paper…
Reading the bibliography…
The new era of technology has brought us to the point where it is convenient for people to share their opinions over an abundance of platforms.
LIII. On lines and planes of closest fit to systems of points in space
Karl Pearson. 1901 · 1901
Earlier work this paper cites.
UNITER: UNiversal Image-TExt Representation Learning
Yen-Chun Chen, Linjie Li, Licheng Yu, Ahmed El Kholy, Faisal Ahmed, Zhe Gan, Yu Cheng, and Jingjing Liu. 2020 · 1909
Earlier work this paper cites.
Better Summarization Evaluation with Word Embeddings for ROUGE. In Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Lisbon, Portugal, 1925–1930
Jun-Ping Ng and Viktoria Abrecht. 2015 · 1930
Earlier work this paper cites.
The automatic creation of literature abstracts
Hans Peter Luhn. 1958 · 1958
Earlier work this paper cites.
An analysis of approximations for maximizing submodular set functions—I
George L Nemhauser, Laurence A Wolsey, and Marshall L Fisher. 1978 · 1978
Earlier work this paper cites.
Automatic text processing: The transformation, analysis, and retrieval of
Gerard Salton. 1989 · 1989
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Latent dirichlet allocation
David M Blei, Andrew Y Ng, and Michael I Jordan. 2003 · 2003
Earlier work this paper cites.
Multimodal summarization of meeting recordings. In 2003 International Conference on Multimedia and Expo. ICME’03. Proceedings (Cat. No. 03TH8698) , Vol. 3. IEEE, III–25
Berna Erol, D-S Lee, and Jonathan Hull. 2003 · 2003
Earlier work this paper cites.
Multimodal biometric authentication methods: a COTS approach. In Proc. of Workshop on Multimodal User Authentication . Citeseer, 99–106
M Indovina, U Uludag, R Snelick, A Mink, and A Jain. 2003 · 2003
Earlier work this paper cites.
Lexrank: Graph-based lexical centrality as salience in text summarization
Günes Erkan and Dragomir R Radev. 2004 · 2004
Earlier work this paper cites.
Performance of a Deep-Learning Algorithm vs Manual Grading for Detecting Diabetic Retinopathy in India
Varun Gulshan, Renu P. Rajan, Kasumi Widner, Derek Wu, Peter Wubbels, Tyler Rhodes, Kira Whitehouse, Marc Coram, Greg Corrado, Kim Ramasamy, Rajiv Raman, Lily Peng, and Dale R. Webster. 2019 · 2004
Earlier work this paper cites.
ROUGE: A Package for Automatic Evaluation of Summaries. In Text Summarization Branches Out . Association for Computational Linguistics, Barcelona, Spain, 74–81
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Graph-based ranking algorithms for sentence extraction, applied to text summarization. In Proceedings of the ACL Interactive Poster and Demonstration Sessions . 170–173
Rada Mihalcea. 2004 · 2004
Earlier work this paper cites.
Textrank: Bringing order into text. In Proceedings of the 2004 conference on empirical methods in natural language processing . 404–411
Rada Mihalcea and Paul Tarau. 2004 · 2004
Earlier work this paper cites.
Validity index for crisp and fuzzy clusters
Malay K Pakhira, Sanghamitra Bandyopadhyay, and Ujjwal Maulik. 2004 · 2004
Earlier work this paper cites.
Multimodal approach to seismic pavement testing
Nils Ryden, Choon B Park, Peter Ulriksen, and Richard D Miller. 2004 · 2004
Earlier work this paper cites.
Discriminative multimodal biometric authentication based on quality measures
Julian Fierrez-Aguilar, Javier Ortega-Garcia, Joaquin Gonzalez-Rodriguez, and Josef Bigun. 2005 · 2005
Earlier work this paper cites.
Multimodal approaches for emotion recognition: a survey. In Internet Imaging VI , Vol. 5670. International Society for Optics and Photonics, 56–67
Nicu Sebe, Ira Cohen, Theo Gevers, and Thomas S Huang. 2005 · 2005
Earlier work this paper cites.
Large-scale evaluation of multimodal biometric authentication using state-of-the-art systems
Robert Snelick, Umut Uludag, Alan Mink, Mike Indovina, and Anil Jain. 2005 · 2005
Earlier work this paper cites.
Language-agnostic BERT Sentence Embedding
Fangxiaoyu Feng, Yinfei Yang, Daniel Cer, Naveen Arivazhagan, and Wei Wang. 2020 · 2007
Earlier work this paper cites.
Multimodal human–computer interaction: A survey
Alejandro Jaimes and Nicu Sebe. 2007 · 2007
Earlier work this paper cites.
A survey automatic text summarization
Oguzhan Tas and Farzad Kiyani. 2007 · 2007
Earlier work this paper cites.
Movie summarization based on audiovisual saliency detection. In 2008 15th IEEE International Conference on Image Processing . IEEE, 2528–2531
Georgios Evangelopoulos, Konstantinos Rapantzikos, Alexandros Potamianos, Petros Maragos, A Zlatintsi, and Yannis Avrithis. 2008 · 2008
Earlier work this paper cites.
Video summarisation: A conceptual framework and survey of the state of the art
Arthur G Money and Harry Agius. 2008 · 2008
Earlier work this paper cites.
Multimodal named entity disambiguation for noisy social media posts. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 2000–2008
Seungwhan Moon, Leonardo Neves, and Vitor Carvalho. 2018a · 2008
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database. In 2009 IEEE conference on computer vision and pattern recognition . Ieee, 248–255
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei. 2009 · 2009
Earlier work this paper cites.
Video event detection and summarization using audio, visual and text saliency. In 2009 IEEE International Conference on Acoustics, Speech and Signal Processing . IEEE, 3553–3556
Georgios Evangelopoulos, Athanasia Zlatintsi, Georgios Skoumas, Konstantinos Rapantzikos, Alexandros Potamianos, Petros Maragos, and Yannis Avrithis. 2009 · 2009
Earlier work this paper cites.
Gather customer concerns from online product reviews–A text summarization approach
Jiaming Zhan, Han Tong Loh, and Ying Liu. 2009 · 2009
Earlier work this paper cites.
Multi-document summarization model based on integer linear programming
Rasim Alguliev, Ramiz Aliguliyev, and Makrufa Hajirahimova. 2010 · 2010
Earlier work this paper cites.
Multimodal fusion for multimedia analysis: a survey
Pradeep K Atrey, M Anwar Hossain, Abdulmotaleb El Saddik, and Mohan S Kankanhalli. 2010 · 2010
Earlier work this paper cites.
Automatic summarization of audio-visual soccer feeds. In 2010 IEEE International Conference on Multimedia and Expo . IEEE, 837–842
Fan Chen, Christophe De Vleeschouwer, H Duxans Barrobés, J Gregorio Escalada, and David Conejero. 2010 · 2010
Earlier work this paper cites.
A survey of text summarization extractive techniques
Vishal Gupta and Gurpreet Singh Lehal. 2010 · 2010
Earlier work this paper cites.
Multi-document summarization via budgeted maximization of submodular functions. In Human Language Technologies: The 2010 Annual Conference of the North American Chapter of the Association for Computational Linguistics . 912–920
Hui Lin and Jeff Bilmes. 2010 · 2010
Earlier work this paper cites.
Collecting image annotations using amazon’s mechanical turk. In Proceedings of the NAACL HLT 2010 Workshop on Creating Speech and Language Data with Amazon’s Mechanical Turk . 139–147
Cyrus Rashtchian, Peter Young, Micah Hodosh, and Julia Hockenmaier. 2010 · 2010
Earlier work this paper cites.
Multi-document Summarization via Deep Learning Techniques: A Survey
Congbo Ma, Wei Emma Zhang, Mingyu Guo, Hu Wang, and Quan Z. Sheng. 2020 · 2011
Earlier work this paper cites.
Towards multimodal sentiment analysis: Harvesting opinions from the web. In Proceedings of the 13th international conference on multimodal interfaces . 169–176
Louis-Philippe Morency, Rada Mihalcea, and Payal Doshi. 2011 · 2011
Earlier work this paper cites.
Summarizing a document stream. In European conference on information retrieval . Springer, 177–188
Hiroya Takamura, Hikaru Yokono, and Manabu Okumura. 2011 · 2011
Earlier work this paper cites.
Multi-modal summarization of key events and top players in sports tournament videos. In Applications of Computer Vision (WACV), 2011 IEEE Workshop on . IEEE, 471–478
Dian Tjondronegoro, Xiaohui Tao, Johannes Sasongko, and Cher Han Lau. 2011 · 2011
Earlier work this paper cites.
Multimodal summarization of complex sentences. In Proceedings of the 16th international conference on Intelligent user interfaces . ACM, 43–52
Naushad UzZaman, Jeffrey P Bigham, and James F Allen. 2011 · 2011
Earlier work this paper cites.
Water cycle algorithm–A novel metaheuristic optimization method for solving constrained engineering optimization problems
Hadi Eskandar, Ali Sadollah, Ardeshir Bahreininejad, and Mohd Hamdi. 2012 · 2012
Earlier work this paper cites.
Extractive multi-document summarization with integer linear programming and support vector regression. In Proceedings of COLING 2012 . 911–926
Dimitrios Galanis, Gerasimos Lampouras, and Ion Androutsopoulos. 2012 · 2012
Earlier work this paper cites.
A survey of text summarization techniques
Ani Nenkova and Kathleen McKeown. 2012 · 2012
Earlier work this paper cites.
Large-margin learning of submodular summarization models. In Proceedings of the 13th Conference of the European Chapter of the Association for Computational Linguistics . 224–233
Ruben Sipos, Pannaga Shivaswamy, and Thorsten Joachims. 2012 · 2012
Earlier work this paper cites.
Visualizing timelines: Evolutionary summarization via iterative reinforcement between text and image streams. In Proceedings of the 21st ACM international conference on Information and knowledge management . 275–284
Rui Yan, Xiaojun Wan, Mirella Lapata, Wayne Xin Zhao, Pu-Jen Cheng, and Xiaoming Li. 2012 · 2012
Earlier work this paper cites.
Multimedia summarization for trending topics in microblogs. In Proceedings of the 22nd ACM international conference on Information & Knowledge Management . 1807–1812
Jingwen Bian, Yang Yang, and Tat-Seng Chua. 2013 · 2013
Earlier work this paper cites.
Multimodal saliency and fusion for movie summarization based on aural, visual, and textual attention
Georgios Evangelopoulos, Athanasia Zlatintsi, Alexandros Potamianos, Petros Maragos, Konstantinos Rapantzikos, Georgios Skoumas, and Yannis Avrithis. 2013 · 2013
Earlier work this paper cites.
Framing image description as a ranking task: Data, models and evaluation metrics
Micah Hodosh, Peter Young, and Julia Hockenmaier. 2013 · 2013
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality. In Advances in neural information processing systems . 3111–3119
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg S Corrado, and Jeff Dean. 2013 · 2013
Earlier work this paper cites.
Multimodal sentiment analysis of spanish online videos
Verónica Pérez Rosas, Rada Mihalcea, and Louis-Philippe Morency. 2013 · 2013
Earlier work this paper cites.
Automatic text summarization: Past, present and future
Horacio Saggion and Thierry Poibeau. 2013 · 2013
Earlier work this paper cites.
Socially motivated multimedia topic timeline summarization. In Proceedings of the 2nd international workshop on Socially-aware multimedia . 19–24
Mathilde Sahuguet and Benoit Huet. 2013 · 2013
Earlier work this paper cites.
Sumblr: continuous summarization of evolving tweet streams. In Proceedings of the 36th international ACM SIGIR conference on Research and development in information retrieval . 533–542
Lidan Shou, Zhenhua Wang, Ke Chen, and Gang Chen. 2013 · 2013
Earlier work this paper cites.
A cross-media evolutionary timeline generation framework based on iterative recommendation. In Proceedings of the 3rd ACM conference on International conference on multimedia retrieval . 73–80
Shize Xu, Liang Kong, and Yan Zhang. 2013 · 2013
Earlier work this paper cites.
Query-chain focused summarization. In Proceedings of the 52nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 913–922
Tal Baumel, Raphael Cohen, and Michael Elhadad. 2014 · 2014
Earlier work this paper cites.
Multimedia summarization for social events in microblog stream
Jingwen Bian, Yang Yang, Hanwang Zhang, and Tat-Seng Chua. 2014 · 2014
Earlier work this paper cites.
Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Translation. In Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) . 1724–1734
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Deep fragment embeddings for bidirectional image sentence mapping. In Advances in neural information processing systems . 1889–1897
Andrej Karpathy, Armand Joulin, and Li F Fei-Fei. 2014 · 2014
Earlier work this paper cites.
Multimodal movement prediction-towards an individual assistance of patients
Elsa Andrea Kirchner, Marc Tabie, and Anett Seeland. 2014 · 2014
Earlier work this paper cites.
Fisher vectors derived from hybrid gaussian-laplacian mixture models for image annotation
Benjamin Klein, Guy Lev, Gil Sadeh, and Lior Wolf. 2014 · 2014
Earlier work this paper cites.
Grey wolf optimizer
Seyedali Mirjalili, Seyed Mohammad Mirjalili, and Andrew Lewis. 2014 · 2014
Earlier work this paper cites.
Multimodal feature extraction and fusion for semantic mining of soccer video: a survey
Payam Oskouie, Sara Alipour, and Amir-Masoud Eftekhari-Moghadam. 2014 · 2014
Earlier work this paper cites.
Glove: Global vectors for word representation. In Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP) . 1532–1543
Jeffrey Pennington, Richard Socher, and Christopher D Manning. 2014 · 2014
Earlier work this paper cites.
Landmark summarization with diverse viewpoints
Xueming Qian, Yao Xue, Xiyu Yang, Yuan Yan Tang, Xingsong Hou, and Tao Mei. 2014 · 2014
Earlier work this paper cites.
From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions
Peter Young, Alice Lai, Micah Hodosh, and Julia Hockenmaier. 2014 · 2014
Cited alongside, same era.
Perception-guided multimodal feature fusion for photo aesthetics assessment. In Proceedings of the 22nd ACM international conference on Multimedia . 237–246
Luming Zhang, Yue Gao, Chao Zhang, Hanwang Zhang, Qi Tian, and Roger Zimmermann. 2014 · 2014
Cited alongside, same era.
Incrests: Towards real-time incremental short text summarization on comment streams from social network services
Cheng-Ying Liu, Ming-Syan Chen, and Chi-Yao Tseng. 2015 · 2015
Cited alongside, same era.
Creating diverse product review summaries: a graph approach. In International Conference on Web Information Systems Engineering . Springer, 169–184
Natwar Modani, Elham Khabiri, Harini Srinivasan, and James Caverlee. 2015 · 2015
Cited alongside, same era.
A survey on video summarization techniques
Tinumol Sebastian and Jiby J Puthiyidam. 2015 · 2015
Survey of Compressed Domain Video Summarization Techniques
Madhushree Basavarajaiah and Priyanka Sharma. 2019 · 2019
Later among the works it cites.
Probing the need for visual context in multimodal machine translation
Ozan Caglayan, Pranava Madhyastha, Lucia Specia, and Loïc Barrault. 2019 · 2019
Later among the works it cites.
News Image Captioning Based on Text Summarization Using Image as Query. In 2019 15th International Conference on Semantics, Knowledge and Grids (SKG) . IEEE, 123–126
Jingqiang Chen and Hai Zhuge. 2019 · 2019
Later among the works it cites.
Towards knowledge-based personalized product description generation in e-commerce. In Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining . 3040–3050
Qibin Chen, Junyang Lin, Yichang Zhang, Hongxia Yang, Jingren Zhou, and Jie Tang. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Very Deep Convolutional Networks for Large-Scale Image Recognition. In International Conference on Learning Representations
Karen Simonyan and Andrew Zisserman. 2015 · 2015
Cited alongside, same era.
Going deeper with convolutions. In Proceedings of the IEEE conference on computer vision and pattern recognition . 1–9
Christian Szegedy, Wei Liu, Yangqing Jia, Pierre Sermanet, Scott Reed, Dragomir Anguelov, Dumitru Erhan, Vincent Vanhoucke, and Andrew Rabinovich. 2015 · 2015
Cited alongside, same era.
Neural Machine Translation by Jointly Learning to Align and Translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2016 · 2016
Cited alongside, same era.
Sentiment analysis and text summarization of online reviews: A survey. In 2016 International Conference on Communication and Signal Processing (ICCSP) . IEEE, 0241–0245
Pankaj Gupta, Ritu Tiwari, and Nirmal Robert. 2016 · 2016
Cited alongside, same era.
Deep residual learning for image recognition. In Proceedings of the IEEE conference on computer vision and pattern recognition . 770–778
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Attention-based multimodal neural machine translation. In Proceedings of the First Conference on Machine Translation: Volume 2, Shared Task Papers . 639–645
Po-Yao Huang, Frederick Liu, Sz-Rung Shiang, Jean Oh, and Chris Dyer. 2016 · 2016
Cited alongside, same era.
Multimodal residual learning for visual qa
Jin-Hwa Kim, Sang-Woo Lee, Donghyun Kwak, Min-Oh Heo, Jeonghee Kim, Jung-Woo Ha, and Byoung-Tak Zhang. 2016a · 2016
Cited alongside, same era.
Sentence Mover’s Similarity: Automatic Evaluation for Multi-Sentence Texts. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Florence, Italy
Elizabeth Clark, Asli Celikyilmaz, and Noah A. Smith. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Association for Computational Linguistics, Minneapolis, Minnesota, 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
A survey on evaluation of summarization methods
Liana Ermakova, Jean Valère Cossu, and Josiane Mothe. 2019 · 2019
Later among the works it cites.
A Survey on Video Summarization Techniques. In 2019 Innovations in Power and Advanced Computing Technologies (i-PACT) , Vol. 1. IEEE, 1–5
Mahesh Kini and Karthik Pai. 2019 · 2019
Later among the works it cites.
Visualbert: A simple and performant baseline for vision and language
Liunian Harold Li, Mark Yatskar, Da Yin, Cho-Jui Hsieh, and Kai-Wei Chang. 2019 · 2019
Later among the works it cites.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks. In Advances in Neural Information Processing Systems . 13–23
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019 · 2019
Later among the works it cites.
Multimodal abstractive summarization for how2 videos
Shruti Palaskar, Jindrich Libovickỳ, Spandana Gella, and Florian Metze. 2019 · 2019
Later among the works it cites.
Social media based event summarization by user–text–image co-clustering
Xueming Qian, Mingdi Li, Yayun Ren, and Shuhui Jiang. 2019 · 2019
Later among the works it cites.
Improvement of query-based text summarization using word sense disambiguation
Nazreena Rahman and Bhogeswar Borah. 2019 · 2019
Later among the works it cites.
Extractive single document summarization using binary differential evolution: Optimization of different sentence quality measures
Naveen Saini, Sriparna Saha, Dhiraj Chakraborty, and Pushpak Bhattacharyya. 2019a · 2019
Later among the works it cites.
Extractive single document summarization using multi-objective optimization: Exploring self-organized differential evolution, grey wolf optimizer and water cycle algorithm
Naveen Saini, Sriparna Saha, Anubhav Jangra, and Pushpak Bhattacharyya. 2019b · 2019
Later among the works it cites.
A Deep Architecture for Multimodal Summarization of Soccer Games. In Proceedings Proceedings of the 2nd International Workshop on Multimedia Content Analysis in Sports . 16–24
Melissa Sanabria, Frédéric Precioso, and Thomas Menguy. 2019 · 2019
Later among the works it cites.
Videobert: A joint model for video and language representation learning. In Proceedings of the IEEE International Conference on Computer Vision . 7464–7473
Chen Sun, Austin Myers, Carl Vondrick, Kevin Murphy, and Cordelia Schmid. 2019 · 2019
Later among the works it cites.
Language and perception: introduction to the special issue speakers and listeners in the visual world
Mila Vulchanova, Valentin Vulchanov, Isabella Fritz, and Evelyn A Milburn. 2019 · 2019
Later among the works it cites.
MoverScore: Text Generation Evaluating with Contextualized Embeddings and Earth Mover Distance. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . Association for Computational Linguistics, Hong Kong, China, 563–578
Wei Zhao, Maxime Peyrard, Fei Liu, Yang Gao, Christian M. Meyer, and Steffen Eger. 2019 · 2019
Later among the works it cites.
Topic and sentiment aware microblog summarization for twitter
Syed Muhammad Ali, Zeinab Noorian, Ebrahim Bagheri, Chen Ding, and Feras Al-Obeidat. 2020 · 2020
Later among the works it cites.
A news image captioning approach based on multimodal pointer-generator network
Jingqiang Chen and Hai Zhuge. 2020 · 2020
Later among the works it cites.
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale. In International Conference on Learning Representations
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Later among the works it cites.
Multi-modal Summarization for Video-containing Documents
Xiyan Fu, Jun Wang, and Zhenglu Yang. 2020 · 2020
Later among the works it cites.
SUPERT: Towards New Frontiers in Unsupervised Evaluation Metrics for Multi-Document Summarization. In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Online
Yang Gao, Wei Zhao, and Steffen Eger. 2020 · 2020
Later among the works it cites.
Exploring Explainable Selection to Control Abstractive Summarization
Wang Haonan, Gao Yang, Bai Yu, Mirella Lapata, and Huang Heyan. 2020 · 2020
Later among the works it cites.
Pixel-bert: Aligning image pixels with text by deep multi-modal transformers
Zhicheng Huang, Zhaoyang Zeng, Bei Liu, Dongmei Fu, and Jianlong Fu. 2020 · 2020
Later among the works it cites.
A comprehensive survey of multi-view video summarization
Tanveer Hussain, Khan Muhammad, Weiping Ding, Jaime Lloret, Sung Wook Baik, and Victor Hugo C de Albuquerque. 2020 · 2020
Later among the works it cites.
MAST: Multimodal abstractive summarization with trimodal hierarchical attention
Aman Khullar and Udit Arora. 2020 · 2020
Later among the works it cites.
VMSMO: Learning to Generate Multimodal Summary for Video-based News Articles
Mingzhe Li, Xiuying Chen, Shen Gao, Zhangming Chan, Dongyan Zhao, and Rui Yan. 2020a · 2020
Later among the works it cites.
Multistage Fusion with Forget Gate for Multimodal Summarization in Open-Domain Videos. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) . Association for Computational Linguistics, Online, 1834–1845
Nayu Liu, Xian Sun, Hongfeng Yu, Wenkai Zhang, and Guangluan Xu. 2020 · 2020
Later among the works it cites.
Aesthetic assessment of website design based on multimodal fusion
Xin Liu and Yujia Jiang. 2020 · 2020
Later among the works it cites.
On faithfulness and factuality in abstractive summarization
Joshua Maynez, Shashi Narayan, Bernd Bohnet, and Ryan McDonald. 2020 · 2020
Later among the works it cites.
Multi-modal product title compression
Lianhai Miao, Da Cao, Juntao Li, and Weili Guan. 2020 · 2020
Later among the works it cites.
Support-set bottlenecks for video-text representation learning
Mandela Patrick, Po-Yao Huang, Yuki Asano, Florian Metze, Alexander Hauptmann, Joao Henriques, and Andrea Vedaldi. 2020 · 2020
Later among the works it cites.
Multimodal multi-task financial risk forecasting. In Proceedings of the 28th ACM International Conference on Multimedia . 456–465
Ramit Sawhney, Puneet Mathur, Ayush Mangal, Piyush Khanna, Rajiv Ratn Shah, and Roger Zimmermann. 2020 · 2020
Later among the works it cites.
Why pay more? A simple and efficient named entity recognition system for tweets
Chanchal Suman, Saichethan Miriyala Reddy, Sriparna Saha, and Pushpak Bhattacharyya. 2020 · 2020
Later among the works it cites.
What Makes a Good Summary? Reconsidering the Focus of Automatic Summarization
Maartje ter Hoeve, Julia Kiseleva, and Maarten de Rijke. 2020 · 2020
Later among the works it cites.
Improving text summarization of online hotel reviews with review helpfulness and sentiment
Chih-Fong Tsai, Kuanchin Chen, Ya-Han Hu, and Wei-Kai Chen. 2020 · 2020
Later among the works it cites.
A Deep Multi-Level Attentive network for Multimodal Sentiment Analysis
Ashima Yadav and Dinesh Kumar Vishwakarma. 2020 · 2020
Later among the works it cites.
Improving multimodal named entity recognition via entity span detection with unified multimodal transformer. Association for Computational Linguistics
Jianfei Yu, Jing Jiang, Li Yang, and Rui Xia. 2020 · 2020
Later among the works it cites.
BERTScore: Evaluating Text Generation with BERT. In International Conference on Learning Representations
Tianyi Zhang*, Varsha Kishore*, Felix Wu*, Kilian Q. Weinberger, and Yoav Artzi. 2020 · 2020
Later among the works it cites.
Unified Vision-Language Pre-Training for Image Captioning and VQA.. In AAAI . 13041–13049
Luowei Zhou, Hamid Palangi, Lei Zhang, Houdong Hu, Jason J Corso, and Jianfeng Gao. 2020 · 2020
Later among the works it cites.
QIAI at MEDIQA 2021: Multimodal Radiology Report Summarization. In Proceedings of the 20th Workshop on Biomedical Language Processing . 285–290
Jean-Benoit Delbrouck, Cassie Zhang, and Daniel Rubin. 2021 · 2021
Closest in time.
SummEval: Re-evaluating Summarization Evaluation
A. R. Fabbri, Wojciech Kryscinski, Bryan McCann, R. Socher, and Dragomir Radev. 2021 · 2021
Closest in time.
Multi-Modal Supplementary-Complementary Summarization Using Multi-Objective Optimization
Anubhav Jangra, Sriparna Saha, Adam Jatowt, and Mohammed Hasanuzzaman. 2021 · 2021
Closest in time.
Multi-modal Summarization
Tsuneaki Kato. 2021 · 2021
Closest in time.
Transformers in vision: A survey
Salman Khan, Muzammal Naseer, Munawar Hayat, Syed Waqas Zamir, Fahad Shahbaz Khan, and Mubarak Shah. 2021 · 2021
Closest in time.
Exploiting BERT for Multimodal Target Sentiment Classification through Input Space Translation
Zaid Khan and Yun Raymond Fu. 2021 · 2021
Closest in time.
VX2TEXT: End-to-End Learning of Video-Based Text Generation From Multimodal Inputs. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 7005–7015
Xudong Lin, Gedas Bertasius, Jue Wang, Shih-Fu Chang, Devi Parikh, and Lorenzo Torresani. 2021 · 2021
Closest in time.
Zero-Shot Text-to-Image Generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever. 2021 · 2021
Closest in time.
On multimodal microblog summarization
Naveen Saini, Sriparna Saha, Pushpak Bhattacharyya, Shubhankar Mrinal, and Santosh Kumar Mishra. 2021 · 2021
Closest in time.
Mimoqa: Multimodal input multimodal output question answering. In Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . 5317–5332
Hrituraj Singh, Anshul Nasery, Denil Mehta, Aishwarya Agarwal, Jatin Lamba, and Balaji Vasan Srinivasan. 2021 · 2021
Closest in time.
Going deeper with image transformers
Hugo Touvron, Matthieu Cord, Alexandre Sablayrolles, Gabriel Synnaeve, and Hervé Jégou. 2021 · 2021
Closest in time.
Is Human Scoring the Best Criteria for Summary Evaluation?. In Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021 . Association for Computational Linguistics, Online
Oleg Vasilyev and John Bohannon. 2021 · 2021
Closest in time.
A Survey on Medical Document Summarization
Raghav Jain, Anubhav Jangra, Sriparna Saha, and Adam Jatowt. 2022a · 2022
Closest in time.
WIDAR - Weighted Input Document Augmented ROUGE. In Advances in Information Retrieval: 44th European Conference on IR Research, ECIR 2022, Stavanger, Norway, April 10–14, 2022, Proceedings, Part I (Stavanger, Norway). Springer-Verlag, Berlin, Heidelberg, 304–321
Raghav Jain, Vaibhav Mavi, Anubhav Jangra, and Sriparna Saha. 2022b · 2022
Closest in time.
Multimodal Summarization: A Concise Review. In Proceedings of the International Conference on Computational Intelligence and Sustainable Technologies . Springer, 613–623
Hira Javed, MM Sufyan Beg, and Nadeem Akhtar. 2022 · 2022
Closest in time.
Combining Vision and Language Representations for Patch-based Identification of Lexico-Semantic Relations. In Proceedings of the 30th ACM International Conference on Multimedia . 4406–4415
Prince Jha, Gaël Dias, Alexis Lechervy, Jose G Moreno, Anubhav Jangra, Sebastião Pais, and Sriparna Saha. 2022 · 2022
Closest in time.
BRIO: Bringing Order to Abstractive Summarization. In Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . 2890–2903
Yixin Liu, Pengfei Liu, Dragomir Radev, and Graham Neubig. 2022 · 2022
Closest in time.
Topic-aware Multimodal Summarization. In Findings of the Association for Computational Linguistics: AACL-IJCNLP 2022 . 387–398
Sourajit Mukherjee, Anubhav Jangra, Sriparna Saha, and Adam Jatowt. 2022 · 2022
Closest in time.
A Duo-generative Approach to Explainable Multimodal COVID-19 Misinformation Detection. In Proceedings of the ACM Web Conference 2022 . 3623–3631
Lanyu Shang, Ziyi Kou, Yang Zhang, and Dong Wang. 2022 · 2022
Closest in time.
MAKED: Multi-lingual Automatic Keyword Extraction Dataset. In Proceedings of the Thirteenth Language Resources and Evaluation Conference . 6170–6179
Yash Verma, Anubhav Jangra, Sriparna Saha, Adam Jatowt, and Dwaipayan Roy. 2022 · 2022
Closest in time.
Multimodal trajectory predictions for autonomous driving using deep convolutional networks. In 2019 International Conference on Robotics and Automation (ICRA) . IEEE, 2090–2096
Henggang Cui, Vladan Radosavljevic, Fang-Chieh Chou, Tsung-Han Lin, Thi Nguyen, Tzu-Kuo Huang, Jeff Schneider, and Nemanja Djuric. 2019 · 2096
Closest in time.