Fetching the paper…
Reading the bibliography…
Text response generation for multimodal task-oriented dialog systems, which aims to generate the proper text response given the multimodal context, is an essential yet challenging task.
Symmetric Regularization based BERT for Pair-wise Semantic Reasoning. In Proceedings of the International ACM SIGIR conference on research and development in Information Retrieval . ACM, 1901–1904
Weidi Xu, Xingyi Cheng, Kunlong Chen, and Taifeng Wang. 2020 · 1904
Earlier work this paper cites.
Graph-Structured Context Understanding for Knowledge-grounded Response Generation. In Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval . ACM, 1930–1934
Yanran Li, Wenjie Li, and Zhitao Wang. 2021 · 1934
Earlier work this paper cites.
Minimum cross entropy thresholding
Chun Hung Li and C. K. Lee. 1993 · 1993
Earlier work this paper cites.
Automatic Evaluation of Machine Translation Quality Using N-Gram Co-Occurrence Statistics. In Proceedings of the Second International Conference on Human Language Technology Research . Morgan Kaufmann Publishers Inc., 138–145
George Doddington. 2002 · 2002
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation. In Proceedings of the Annual Meeting of the Association for Computational Linguistics . ACL, 311–318
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Distributed Representations of Words and Phrases and Their Compositionality. In Proceedings of the International Conference on Neural Information Processing Systems . Curran Associates Inc., 3111–3119
Tomas Mikolov, Ilya Sutskever, Kai Chen, Greg Corrado, and Jeffrey Dean. 2013 · 2013
Earlier work this paper cites.
Empirical Evaluation of Gated Recurrent Neural Networks on Sequence Modeling
Junyoung Chung, Çaglar Gülçehre, KyungHyun Cho, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
GloVe: Global Vectors for Word Representation. In Proceedings of the Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, 1532–1543
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Sequence to Sequence Learning with Neural Networks. In Proceedings of the International Conference on Neural Information Processing Systems . MIT Press, 3104–3112
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le. 2014 · 2014
Earlier work this paper cites.
End-to-End Task-Completion Neural Dialogue Systems. In Proceedings of the Eighth International Joint Conference on Natural Language Processing . Asian Federation of Natural Language Processing, 733–743
Xiujun Li, Yun-Nung Chen, Lihong Li, Jianfeng Gao, and Asli Celikyilmaz. 2017 · 2017
Earlier work this paper cites.
Attention Is All You Need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017b · 2017
Earlier work this paper cites.
Gated-Attention Architectures for Task-Oriented Language Grounding. In Proceedings of the AAAI Conference on Artificial Intelligence . AAAI Press, 2819–2826
Devendra Singh Chaplot, Kanthashree Mysore Sathyendra, Rama Kumar Pasumarthi, Dheeraj Rajagopal, and Ruslan Salakhutdinov. 2018 · 2018
Earlier work this paper cites.
Deep Neural Architecture for Multi-Modal Retrieval based on Joint Embedding Space for Text and Images. In Proceedings of the Eleventh ACM International Conference on Web Search and Data Mining . ACM, 28–36
Saeid Balaneshin Kordan and Alexander Kotov. 2018 · 2018
Earlier work this paper cites.
Sequicity: Simplifying Task-oriented Dialogue Systems with Single Sequence-to-Sequence Architectures. In Proceedings of the Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, 1437–1447
Wenqiang Lei, Xisen Jin, Min-Yen Kan, Zhaochun Ren, Xiangnan He, and Dawei Yin. 2018 · 2018
Earlier work this paper cites.
Knowledge-aware Multimodal Dialogue Systems. In Proceedings of the ACM International Conference on Multimedia . ACM, 801–809
Lizi Liao, Yunshan Ma, Xiangnan He, Richang Hong, and Tat-Seng Chua. 2018 · 2018
Earlier work this paper cites.
Mem2Seq: Effectively Incorporating Knowledge Bases into End-to-End Task-Oriented Dialog Systems. In Proceedings ofAnnual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, 1468–1478
Andrea Madotto, Chien-Sheng Wu, and Pascale Fung. 2018 · 2018
Earlier work this paper cites.
Improving Language Understanding by Generative Pre-Training
Alec Radford and Karthik Narasimhan. 2018 · 2018
Earlier work this paper cites.
Towards Building Large Scale Multimodal Domain-Aware Conversation Systems. In Proceedings of the AAAI Conference on Artificial Intelligence . AAAI Press, 696–704
Amrita Saha, Mitesh M. Khapra, and Karthik Sankaranarayanan. 2018 · 2018
Earlier work this paper cites.
Augmenting End-to-End Dialogue Systems With Commonsense Knowledge. In Proceedings of the AAAI Conference on Artificial Intelligence . AAAI Press, 4970–4977
Tom Young, Erik Cambria, Iti Chaturvedi, Hao Zhou, Subham Biswas, and Minlie Huang. 2018 · 2018
Earlier work this paper cites.
Ordinal and Attribute Aware Response Generation in a Multimodal Dialogue System. In Proceedings of the Conference of the Association for Computational Linguistics . Association for Computational Linguistics, 5437–5447
Hardik Chauhan, Mauajama Firdaus, Asif Ekbal, and Pushpak Bhattacharyya. 2019 · 2019
Earlier work this paper cites.
User Attention-guided Multimodal Dialog Systems. In Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval . ACM, 445–454
Chen Cui, Wenjie Wang, Xuemeng Song, Minlie Huang, Xin-Shun Xu, and Liqiang Nie. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the Conference of the North American Chapter of the Association for Computational Linguistics . Association for Computational Linguistics, 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Text-based editing of talking-head video
Ohad Fried, Ayush Tewari, Michael Zollhöfer, Adam Finkelstein, Eli Shechtman, Dan B. Goldman, Kyle Genova, Zeyu Jin, Christian Theobalt, and Maneesh Agrawala. 2019 · 2019
Cited alongside, same era.
Multimodal Dialog System: Generating Responses via Adaptive Decoders. In Proceedings of the ACM International Conference on Multimedia . ACM, 1098–1106
Liqiang Nie, Wenjie Wang, Richang Hong, Meng Wang, and Qi Tian. 2019 · 2019
Cited alongside, same era.
EMScore: Evaluating Video Captioning via Coarse-Grained and Fine-Grained Embedding Matching
Yaya Shi, Xu Yang, Haiyang Xu, Chunfeng Yuan, Bing Li, Weiming Hu, and Zheng-Jun Zha. 2021 · 2021
Later among the works it cites.
Comprehensive Linguistic-Visual Composition Network for Image Retrieval. In Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval . ACM, 1369–1378
Haokun Wen, Xuemeng Song, Xin Yang, Yibing Zhan, and Liqiang Nie. 2021 · 2021
Later among the works it cites.
Vision Guided Generative Pre-trained Language Models for Multimodal Abstractive Summarization. In Proceedings of the Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, 3995–4007
Tiezheng Yu, Wenliang Dai, Zihan Liu, and Pascale Fung. 2021 · 2021
Later among the works it cites.
Understanding WeChat User Preferences and “Wow” Diffusion
Fanjin Zhang, Jie Tang, Xueyi Liu, Zhenyu Hou, Yuxiao Dong, Jing Zhang, Xiao Liu, Ruobing Xie, Kai Zhuang, Xu Zhang, Leyu Lin, and Philip Yu. 2021b · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
PyTorch: An Imperative Style, High-Performance Deep Learning Library. In Proceedings of the Annual Conference on Neural Information Processing Systems . 8024–8035
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Köpf, Edward Z. Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019 · 2019
Cited alongside, same era.
In Proceedings of the International ACM SIGIR conference on research and development in Information Retrieval . ACM, 1665–1668
Jie Cai, Zhengzhou Zhu, Ping Nie, and Qian Liu. 2020 · 2020
Cited alongside, same era.
Fine-Grained Privacy Detection with Graph-Regularized Hierarchical Attentive Representation Learning
Xiaolin Chen, Xuemeng Song, Ruiyang Ren, Lei Zhu, Zhiyong Cheng, and Liqiang Nie. 2020 · 2020
Cited alongside, same era.
Multimodal Dialogue Systems via Capturing Context-aware Dependencies of Semantic Elements. In Proceedings of the ACM International Conference on Multimedia . ACM, 2755–2764
Weidong He, Zhi Li, Dongcai Lu, Enhong Chen, Tong Xu, Baoxing Huai, and Jing Yuan. 2020 · 2020
Cited alongside, same era.
BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension. In Proceedings of the Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, 7871–7880
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2020 · 2020
Cited alongside, same era.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Cited alongside, same era.
Dual Dynamic Memory Network for End-to-End Multi-turn Task-oriented Dialog Systems. In Proceedings of the International Conference on Computational Linguistics . International Committee on Computational Linguistics, 4100–4110
Jian Wang, Junhao Liu, Wei Bi, Xiaojiang Liu, Kejing He, Ruifeng Xu, and Min Yang. 2020 · 2020
Cited alongside, same era.
Adversarial-Enhanced Hybrid Graph Network for User Identity Linkage. In Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval . ACM, 1084–1093
Xiaolin Chen, Xuemeng Song, Guozhen Peng, Shanshan Feng, and Liqiang Nie. 2021 · 2021
Cited alongside, same era.
Leveraging Lead Bias for Zero-shot Abstractive News Summarization. In Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval . ACM, 1462–1471
Chenguang Zhu, Ziyi Yang, Robert Gmyr, Michael Zeng, and Xuedong Huang. 2021 · 2021
Later among the works it cites.
CATS: Customizable Abstractive Topic-based Summarization
Seyed Ali Bahrainian, George Zerveas, Fabio Crestani, and Carsten Eickhoff. 2022 · 2022
Closest in time.
BERT-ER: Query-specific BERT Entity Representations for Entity Ranking. In Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval . ACM, 1466–1477
Shubham Chatterjee and Laura Dietz. 2022 · 2022
Closest in time.
KGGen: A Generative Approach for Incipient Knowledge Graph Population
Hao Chen, Chenwei Zhang, Jun Li, Philip S. Yu, and Ning Jing. 2022 · 2022
Closest in time.
Toward Personalized Answer Generation in E-Commerce via Multi-perspective Preference Modeling
Yang Deng, Yaliang Li, Wenxuan Zhang, Bolin Ding, and Wai Lam. 2022 · 2022
Closest in time.
"What Can I Cook with these Ingredients?" - Understanding Cooking-Related Information Needs in Conversational Search
Alexander Frummet, David Elsweiler, and Bernd Ludwig. 2022 · 2022
Closest in time.
HeteroQA: Learning towards Question-and-Answering through Multiple Information Sources via Heterogeneous Graph Modeling. In Proceedings of the ACM International Conference on Web Search and Data Mining . ACM, 307–315
Shen Gao, Yuchi Zhang, Yongliang Wang, Yang Dong, Xiuying Chen, Dongyan Zhao, and Rui Yan. 2022 · 2022
Closest in time.
Hierarchical Prediction and Adversarial Learning For Conditional Response Generation
Yanran Li, Ruixiang Zhang, Wenjie Li, and Ziqiang Cao. 2022b · 2022
Closest in time.
Topic-Guided Conversational Recommender in Multiple Domains
Lizi Liao, Ryuichi Takanobu, Yunshan Ma, Xun Yang, Minlie Huang, and Tat-Seng Chua. 2022 · 2022
Closest in time.
Graph-Grounded Goal Planning for Conversational Recommendation
Zeming Liu, Ding Zhou, Hao Liu, Haifeng Wang, Zheng-Yu Niu, Hua Wu, Wanxiang Che, Ting Liu, and Hui Xiong. 2022 · 2022
Closest in time.
UniTranSeR: A Unified Transformer Semantic Representation Framework for Multimodal Task-Oriented Dialog System. In Proceedings of the Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, 103–114
Zhiyuan Ma, Jianjun Li, Guohui Li, and Yongjing Cheng. 2022 · 2022
Closest in time.
On the Study of Transformers for Query Suggestion
Agnès Mustar, Sylvain Lamprier, and Benjamin Piwowarski. 2022 · 2022
Closest in time.
V2P: Vision-to-Prompt based Multi-Modal Product Summary Generation. In Proceedings of the International ACM SIGIR Conference on Research and Development in Information Retrieval . ACM, 992–1001
Xuemeng Song, Liqiang Jing, Dengtian Lin, Zhongzhou Zhao, Haiqing Chen, and Liqiang Nie. 2022 · 2022
Closest in time.
AnyFace: Free-style Text-to-Face Synthesis and Manipulation
Jianxin Sun, Qiyao Deng, Qi Li, Muyi Sun, Min Ren, and Zhenan Sun. 2022 · 2022
Closest in time.
Personalized Graph Neural Networks With Attention Mechanism for Session-Aware Recommendation
Mengqi Zhang, Shu Wu, Meng Gao, Xin Jiang, Ke Xu, and Liang Wang. 2022 · 2022
Closest in time.