Fetching the paper…
Reading the bibliography…
There has been an increased interest in multimodal language processing including multimodal dialog, question answering, sentiment analysis, and speech recognition.
Analysis of individual differences in multidimensional scaling via an n-way generalization of “eckart-young” decomposition
J. Douglas Carroll and Jih-Jie Chang. 1970 · 1970
Earlier work this paper cites.
Bounds on the ranks of some 3-tensors
M.D. Atkinson and S. Lloyd. 1980 · 1980
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Interactive multimodal robot programming
Soshi Iba, Christiaan J. J. Paredis, and Pradeep K. Khosla. 2005 · 2005
Earlier work this paper cites.
Multimodal Dialogue Systems , pages 3–11. Springer Netherlands, Dordrecht
Alexander I. Rudnicky. 2005 · 2005
Earlier work this paper cites.
Tensor rank and the ill-posedness of the best low-rank approximation problem
Vin de Silva and Lek-Heng Lim. 2008 · 2008
Earlier work this paper cites.
Speaker identification on the scotus corpus
Jiahong Yuan and Mark Liberman. 2008 · 2008
Earlier work this paper cites.
Discriminative optical flow tensor for video semantic analysis
Xinbo Gao, Yimin Yang, Dacheng Tao, and Xuelong Li. 2009 · 2009
Earlier work this paper cites.
Tensor-based transductive learning for multimodality video semantic concept detection
F. Wu, Y. Liu, and Y. Zhuang. 2009 · 2009
Earlier work this paper cites.
Analyzing multimodal time series as dynamical systems
Shohei Hidaka and Chen Yu. 2010 · 2010
Earlier work this paper cites.
Tensor rank: Some lower and upper bounds
Boris Alexeev, Michael A. Forbes, and Jacob Tsimerman. 2011 · 2011
Earlier work this paper cites.
Towards multimodal sentiment analysis: Harvesting opinions from the web
Louis-Philippe Morency, Rada Mihalcea, and Payal Doshi. 2011 · 2011
Earlier work this paper cites.
Multimodal intent recognition for natural human-robotic interaction
James Rossiter. 2011 · 2011
Earlier work this paper cites.
Multimodal sentiment analysis
Rada Mihalcea. 2012 · 2012
Earlier work this paper cites.
Tensor decompositions for learning latent variable models
Animashree Anandkumar, Rong Ge, Daniel Hsu, Sham M. Kakade, and Matus Telgarsky. 2014 · 2014
Earlier work this paper cites.
An upper bound for the real tensor rank and the real symmetric tensor rank in terms of the complex ranks
E. Ballico. 2014 · 2014
Earlier work this paper cites.
Covarep - a collaborative voice analysis repository for speech technologies
Gilles Degottex, John Kane, Thomas Drugman, Tuomo Raitio, and Stefan Scherer. 2014 · 2014
Earlier work this paper cites.
Computational complexity of tensor nuclear norm
Shmuel Friedland and Lek-Heng Lim. 2014 · 2014
Earlier work this paper cites.
Relations of the Nuclear Norms of a Tensor and its Matrix Flattenings
Shenglong Hu. 2014 · 2014
Cited alongside, same era.
Low-rank tensors for scoring dependency structures
Tao Lei, Yu Xin, Yuan Zhang, Regina Barzilay, and Tommi Jaakkola. 2014 · 2014
Cited alongside, same era.
Max-margin tensor neural network for chinese word segmentation
Wenzhe Pei, Tao Ge, and Baobao Chang. 2014 · 2014
Cited alongside, same era.
Glove: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D Manning. 2014 · 2014
Cited alongside, same era.
Improved multimodal deep learning with variation of information
Kihyuk Sohn, Wenling Shang, and Honglak Lee. 2014 · 2014
Cited alongside, same era.
Learning distributed representations for structured output prediction
Hyperspectral image restoration using low-rank tensor recovery
Haiyan Fan, Yunjin Chen, Yulan Guo, Hongyan Zhang, and Gangyao Kuang. 2017 · 2017
Later among the works it cites.
Tensor product generation networks
Qiuyuan Huang, Paul Smolensky, Xiaodong He, Li Deng, and Dapeng Oliver Wu. 2017 · 2017
Later among the works it cites.
Facial expression analysis
iMotions. 2017 · 2017
Later among the works it cites.
Jean Kossaifi, Zachary C. Lipton, Aran Khanna, Tommaso Furlanello, and Anima Anandkumar. 2017 · 2017
Later among the works it cites.
Multimodal probabilistic model-based planning for human-robot interaction
Edward Schmerling, Karen Leung, Wolf Vollprecht, and Marco Pavone. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Vivek Srikumar and Christopher D Manning. 2014 · 2014
Cited alongside, same era.
Multimodal learning with deep boltzmann machines
Nitish Srivastava and Ruslan Salakhutdinov. 2014 · 2014
Cited alongside, same era.
VQA: Visual Question Answering
Stanislaw Antol, Aishwarya Agrawal, Jiasen Lu, Margaret Mitchell, Dhruv Batra, C. Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Extracting low-dimensional latent structure from time series in the presence of delays
Karthik Lakshmanan, Patrick T. Sadtler, Elizabeth C. Tyler-Kabara, Aaron P. Batista, and Byron M. Yu. 2015 · 2015
Cited alongside, same era.
Convolutional neural tensor network architecture for community-based question answering
Xipeng Qiu and Xuanjing Huang. 2015 · 2015
Cited alongside, same era.
Statistical machine translation features with multitask tensor networks
Hendra Setiawan, Zhongqiang Huang, Jacob Devlin, Thomas Lamar, Rabih Zbib, Richard Schwartz, and John Makhoul. 2015 · 2015
Cited alongside, same era.
Movieqa: Understanding stories in movies through question-answering
Makarand Tapaswi, Yukun Zhu, Rainer Stiefelhagen, Antonio Torralba, Raquel Urtasun, and Sanja Fidler. 2015 · 2015
Cited alongside, same era.
Missing modalities imputation via cascaded residual autoencoder
Luan Tran, Xiaoming Liu, Jiayu Zhou, and Rong Jin. 2017 · 2017
Later among the works it cites.
Learning to extract semantic structure from documents using multimodal fully convolutional neural networks
Xiao Yang, Ersin Yumer, Paul Asente, Mike Kraley, Daniel Kifer, and C. Lee Giles. 2017 · 2017
Later among the works it cites.
Tensor fusion network for multimodal sentiment analysis
Amir Zadeh, Minghai Chen, Soujanya Poria, Erik Cambria, and Louis-Philippe Morency. 2017 · 2017
Later among the works it cites.
A tucker deep computation model for mobile multimedia feature learning
Qingchen Zhang, Laurence T. Yang, Xingang Liu, Zhikui Chen, and Peng Li. 2017 · 2017
Later among the works it cites.
Deep adversarial learning for multi-modality missing data completion
Lei Cai, Zhengyang Wang, Hongyang Gao, Dinggang Shen, and Shuiwang Ji. 2018 · 2018
Later among the works it cites.
Embodied Question Answering
Abhishek Das, Samyak Datta, Georgia Gkioxari, Stefan Lee, Devi Parikh, and Dhruv Batra. 2018 · 2018
Later among the works it cites.
Multimodal affective analysis using hierarchical attention strategy with word-level alignment
Yue Gu, Kangning Yang, Shiyu Fu, Shuhong Chen, Xinyu Li, and Ivan Marsic. 2018 · 2018
Later among the works it cites.
Multimodal language analysis with recurrent multistage fusion
Paul Pu Liang, Ziyin Liu, Amir Zadeh, and Louis-Philippe Morency. 2018 · 2018
Later among the works it cites.
Efficient low-rank multimodal fusion with modality-specific factors
Zhun Liu, Ying Shen, Varun Bharadhwaj Lakshminarasimhan, Paul Pu Liang, AmirAli Bagher Zadeh, and Louis-Philippe Morency. 2018 · 2018
Later among the works it cites.
Low rank tensor completion for multiway visual data
Zhen Long, Yipeng Liu, Longxi Chen, and Ce Zhu. 2018 · 2018
Later among the works it cites.
A dual framework for low-rank tensor completion
Madhav Nimishakavi, Pratik Kumar Jawanpuria, and Bamdev Mishra. 2018 · 2018
Later among the works it cites.
End-to-end multimodal speech recognition
Shruti Palaskar, Ramon Sanabria, and Florian Metze. 2018 · 2018
Later among the works it cites.
Found in translation: Learning robust joint representations by cyclic translations between modalities
Hai Pham, Paul Pu Liang, Thomas Manzini, Louis-Philippe Morency, and Barnabas Poczos. 2019 · 2019
Closest in time.