Fetching the paper…
Reading the bibliography…
Deep multimodal learning has achieved great progress in recent years.
Indoor segmentation and support inference from rgbd images
Pushmeet Kohli Nathan Silberman, Derek Hoiem and Rob Fergus · 2012
Earlier work this paper cites.
Mixture of experts: a literature survey
Saeed Masoudnia and Reza Ebrahimpour · 2014
Earlier work this paper cites.
Conditional computation in neural networks for faster models
Emmanuel Bengio, Pierre-Luc Bacon, Joelle Pineau, and Doina Precup · 2015
Earlier work this paper cites.
Utd-mhad: A multimodal dataset for human action recognition utilizing a depth camera and a wearable inertial sensor
Chen Chen, Roozbeh Jafari, and Nasser Kehtarnavaz · 2015
Earlier work this paper cites.
Adaptive computation time for recurrent neural networks
Alex Graves · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2016
Earlier work this paper cites.
Categorical reparameterization with gumbel-softmax
Eric Jang, Shixiang Gu, and Ben Poole · 2016
Earlier work this paper cites.
Branchynet: Fast inference via early exiting from deep neural networks
Surat Teerapittayanon, Bradley McDanel, and Hsiang-Tsung Kung · 2016
Earlier work this paper cites.
Gated multimodal units for information fusion
John Arevalo, Thamar Solorio, Manuel Montes-y Gómez, and Fabio A González · 2017
Earlier work this paper cites.
Adaptive neural networks for efficient inference
Tolga Bolukbasi, Joseph Wang, Ofer Dekel, and Venkatesh Saligrama · 2017
Earlier work this paper cites.
Attention-based multimodal fusion for video description
Chiori Hori, Takaaki Hori, Teng-Yok Lee, Ziming Zhang, Bret Harsham, John R Hershey, Tim K Marks, and Kazuhiko Sumi · 2017
Earlier work this paper cites.
Dynamic routing between capsules
Sara Sabour, Nicholas Frosst, and Geoffrey E Hinton · 2017
Earlier work this paper cites.
Deep multimodal feature analysis for action recognition in rgb+ d videos
Amir Shahroudy, Tian-Tsong Ng, Yihong Gong, and Gang Wang · 2017
Earlier work this paper cites.
Outrageously large neural networks: The sparsely-gated mixture-of-experts layer
Noam Shazeer, Azalia Mirhoseini, Krzysztof Maziarz, Andy Davis, Quoc Le, Geoffrey Hinton, and Jeff Dean · 2017
Earlier work this paper cites.
A survey of multimodal sentiment analysis
Mohammad Soleymani, David Garcia, Brendan Jou, Björn Schuller, Shih-Fu Chang, and Maja Pantic · 2017
Earlier work this paper cites.
Temporal multimodal fusion for video emotion classification in the wild
Valentin Vielzeuf, Stéphane Pateux, and Frédéric Jurie · 2017
Earlier work this paper cites.
Tensor fusion network for multimodal sentiment analysis
Amir Zadeh, Minghai Chen, Soujanya Poria, Erik Cambria, and Louis-Philippe Morency · 2017
Earlier work this paper cites.
Adaptive feeding: Achieving fast and accurate detections by adaptively combining object detectors
Hong-Yu Zhou, Bin-Bin Gao, and Jianxin Wu · 2017
Earlier work this paper cites.
Multimodal machine learning: A survey and taxonomy
Tadas Baltrušaitis, Chaitanya Ahuja, and Louis-Philippe Morency · 2018
Cited alongside, same era.
Squeeze-and-excitation networks
Jie Hu, Li Shen, and Gang Sun · 2018
Cited alongside, same era.
Condensenet: An efficient densenet using learned group convolutions
Gao Huang, Shichen Liu, Laurens Van der Maaten, and Kilian Q Weinberger · 2018
Cited alongside, same era.
Learn to combine modalities in multimodal deep learning
Kuan Liu, Yanen Li, Ning Xu, and Prem Natarajan · 2018
Cited alongside, same era.
Efficient low-rank multimodal fusion with modality-specific factors
Zhun Liu, Ying Shen, Varun Bharadhwaj Lakshminarasimhan, Paul Pu Liang, Amir Zadeh, and Louis-Philippe Morency · 2018
Cited alongside, same era.
Listen to look: Action recognition by previewing audio
Ruohan Gao, Tae-Hyun Oh, Kristen Grauman, and Lorenzo Torresani · 2020
Later among the works it cites.
Multiplicative interactions and where to find them
Siddhant M Jayakumar, Wojciech M Czarnecki, Jacob Menick, Jonathan Schwarz, Jack Rae, Simon Osindero, Yee Whye Teh, Tim Harley, and Razvan Pascanu · 2020
Later among the works it cites.
Mmtm: Multimodal transfer module for cnn fusion
Hamid Reza Vaezi Joze, Amirreza Shaban, Michael L Iuzzolino, and Kazuhito Koishida · 2020
Later among the works it cites.
Learning dynamic routing for semantic segmentation
Yanwei Li, Lin Song, Yukang Chen, Zeming Li, Xiangyu Zhang, Xingang Wang, and Jian Sun · 2020
Later among the works it cites.
Deep multimodal fusion by channel exchanging
Yikai Wang, Wenbing Huang, Fuchun Sun, Tingyang Xu, Yu Rong, and Junzhou Huang · 2020
Later among the works it cites.
Glance and focus: a dynamic approach to reducing spatial redundancy in image classification
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hydranets: Specialized dynamic architectures for efficient inference
Ravi Teja Mullapudi, William R Mark, Noam Shazeer, and Kayvon Fatahalian · 2018
Cited alongside, same era.
Light-weight refinenet for real-time semantic segmentation
Vladimir Nekrasov, Chunhua Shen, and Ian Reid · 2018
Cited alongside, same era.
Centralnet: a multilayer approach for multimodal fusion
Valentin Vielzeuf, Alexis Lechervy, Stéphane Pateux, and Frédéric Jurie · 2018
Cited alongside, same era.
Skipnet: Learning dynamic routing in convolutional networks
Xin Wang, Fisher Yu, Zi-Yi Dou, Trevor Darrell, and Joseph E Gonzalez · 2018
Cited alongside, same era.
Multimodal language analysis in the wild: Cmu-mosei dataset and interpretable dynamic fusion graph
Amir Zadeh and Paul Pu · 2018
Cited alongside, same era.
Adaptive convolution for object detection
Chunlin Chen and Qiang Ling · 2019
Cited alongside, same era.
Acnet: Attention based network to exploit complementary features for rgbd semantic segmentation
Xinxin Hu, Kailun Yang, Lei Fei, and Kaiwei Wang · 2019
Cited alongside, same era.
Yulin Wang, Kangchen Lv, Rui Huang, Shiji Song, Le Yang, and Gao Huang · 2020
Later among the works it cites.
Deep multimodal neural architecture search
Zhou Yu, Yuhao Cui, Jun Yu, Meng Wang, Dacheng Tao, and Qi Tian · 2020
Later among the works it cites.
Dynamic routing networks
Shaofeng Cai, Yao Shu, and Wei Wang · 2021
Later among the works it cites.
Dynamic neural networks: A survey
Yizeng Han, Gao Huang, Shiji Song, Le Yang, Honghui Wang, and Yulin Wang · 2021
Later among the works it cites.
Detect, reject, correct: Crossmodal compensation of corrupted sensors
Michelle A Lee, Matthew Tan, Yuke Zhu, and Jeannette Bohg · 2021
Later among the works it cites.
Multibench: Multiscale benchmarks for multimodal representation learning
Paul Pu Liang, Yiwei Lyu, Xiang Fan, Zetian Wu, Yun Cheng, Jason Wu, Leslie Chen, Peter Wu, Michelle A Lee, Yuke Zhu, et al · 2021
Later among the works it cites.
Attention bottlenecks for multimodal fusion
Arsha Nagrani, Shan Yang, Anurag Arnab, Aren Jansen, Cordelia Schmid, and Chen Sun · 2021
Later among the works it cites.
Adamml: Adaptive multi-modal learning for efficient video recognition
Rameswar Panda, Chun-Fu Richard Chen, Quanfu Fan, Ximeng Sun, Kate Saenko, Aude Oliva, and Rogerio Feris · 2021
Later among the works it cites.
Efficient rgb-d semantic segmentation for indoor scene analysis
Daniel Seichter, Mona Köhler, Benjamin Lewandowski, Tim Wengefeld, and Horst-Michael Gross · 2021
Later among the works it cites.
Deep rgb-d saliency detection with depth-sensitive attention and automatic multi-modal fusion
Peng Sun, Wenhu Zhang, Huanyu Wang, Songyuan Li, and Xi Li · 2021
Later among the works it cites.
Multimodal dynamics: Dynamical fusion for trustworthy multimodal classification
Zongbo Han, Fan Yang, Junzhou Huang, Changqing Zhang, and Jianhua Yao · 2022
Closest in time.
Efficient deep visual and inertial odometry with adaptive visual modality selection
Mingyu Yang, Yu Chen, and Hun-Seok Kim · 2022
Closest in time.