Fetching the paper…
Reading the bibliography…
In sequence-to-sequence learning, e.g., natural language generation, the decoder relies on the attention mechanism to efficiently extract information from the encoder.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey E. Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 1958
Earlier work this paper cites.
Backpropagation Applied to Handwritten Zip Code Recognition
Yann LeCun, Bernhard E. Boser, John S. Denker, Donnie Henderson, Richard E. Howard, Wayne E. Hubbard, and Lawrence D. Jackel. 1989 · 1989
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
BLEU: A Method for Automatic Evaluation of Machine Translation. In ACL
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
ROUGE: A Package for Automatic Evaluation of Summaries. In ACL
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments. In ACL Workshop
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database. In CVPR
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Fei-Fei Li. 2009 · 2009
Earlier work this paper cites.
Generating Phrasal and Sentential Paraphrases: A Survey of Data-Driven Methods
Nitin Madnani and Bonnie J. Dorr. 2010 · 2010
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks. In NIPS
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E. Hinton. 2012 · 2012
Earlier work this paper cites.
YouTube2Text: Recognizing and Describing Arbitrary Activities Using Semantic Hierarchies and Zero-Shot Recognition. In ICCV
Sergio Guadarrama, Niveda Krishnamoorthy, Girish Malkarnenkar, Subhashini Venugopalan, Raymond J. Mooney, Trevor Darrell, and Kate Saenko. 2013 · 2013
Earlier work this paper cites.
Image Classification with the Fisher Vector: Theory and Practice
Jorge Sánchez, Florent Perronnin, Thomas Mensink, and Jakob J. Verbeek. 2013 · 2013
Earlier work this paper cites.
Large-Scale Video Classification with Convolutional Neural Networks. In CVPR
Andrej Karpathy, George Toderici, Sanketh Shetty, Thomas Leung, Rahul Sukthankar, and Fei-Fei Li. 2014 · 2014
Earlier work this paper cites.
Softening quantization in bag-of-audio-words. In ICASSP
Stephanie Pancoast and Murat Akbacak. 2014 · 2014
Earlier work this paper cites.
Neural Machine Translation by Jointly Learning to Align and Translate. In ICLR
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
Microsoft COCO Captions: Data Collection and Evaluation Server
Xinlei Chen, Hao Fang, Tsung-Yi Lin, Ramakrishna Vedantam, Saurabh Gupta, Piotr Dollár, and C. Lawrence Zitnick. 2015 · 2015
Earlier work this paper cites.
Teaching Machines to Read and Comprehend. In NIPS
Karl Moritz Hermann, Tomás Kociský, Edward Grefenstette, Lasse Espeholt, Will Kay, Mustafa Suleyman, and Phil Blunsom. 2015 · 2015
Earlier work this paper cites.
Deep visual-semantic alignments for generating image descriptions. In CVPR
Andrej Karpathy and Fei-Fei Li. 2015 · 2015
Earlier work this paper cites.
Effective Approaches to Attention-based Neural Machine Translation. In EMNLP
Thang Luong, Hieu Pham, and Christopher D. Manning. 2015 · 2015
Earlier work this paper cites.
U-Net: Convolutional Networks for Biomedical Image Segmentation. In MICCAI
Olaf Ronneberger, Philipp Fischer, and Thomas Brox. 2015 · 2015
Earlier work this paper cites.
Learning Spatiotemporal Features with 3D Convolutional Networks. In ICCV
Du Tran, Lubomir D. Bourdev, Rob Fergus, Lorenzo Torresani, and Manohar Paluri. 2015 · 2015
Earlier work this paper cites.
CIDEr: Consensus-based image description evaluation. In CVPR
Ramakrishna Vedantam, C. Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Earlier work this paper cites.
Show and tell: A neural image caption generator. In CVPR
Oriol Vinyals, Alexander Toshev, Samy Bengio, and Dumitru Erhan. 2015 · 2015
Earlier work this paper cites.
Show, Attend and Tell: Neural Image Caption Generation with Visual Attention. In ICML
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron C. Courville, Ruslan Salakhutdinov, Richard S. Zemel, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
Describing Videos by Exploiting Temporal Structure. In ICCV
Li Yao, Atousa Torabi, Kyunghyun Cho, Nicolas Ballas, Christopher J. Pal, Hugo Larochelle, and Aaron C. Courville. 2015 · 2015
Earlier work this paper cites.
SPICE: Semantic Propositional Image Caption Evaluation. In ECCV
Peter Anderson, Basura Fernando, Mark Johnson, and Stephen Gould. 2016 · 2016
Earlier work this paper cites.
Lei Jimmy Ba, Ryan Kiros, and Geoffrey E. Hinton. 2016 · 2016
Earlier work this paper cites.
Preparing a collection of radiology examinations for distribution and retrieval
Dina Demner-Fushman, Marc D. Kohli, Marc B. Rosenman, Sonya E. Shooshan, Laritza Rodriguez, Sameer K. Antani, George R. Thoma, and Clement J. McDonald. 2016 · 2016
Earlier work this paper cites.
Deep Residual Learning for Image Recognition. In CVPR
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Jointly Modeling Embedding and Translation to Bridge Video and Language. In CVPR
Yingwei Pan, Tao Mei, Ting Yao, Houqiang Li, and Yong Rui. 2016 · 2016
Earlier work this paper cites.
MSR-VTT: A Large Video Description Dataset for Bridging Video and Language. In CVPR
Jun Xu, Tao Mei, Ting Yao, and Yong Rui. 2016 · 2016
Earlier work this paper cites.
Video Paragraph Captioning Using Hierarchical Recurrent Neural Networks. In CVPR
Haonan Yu, Jiang Wang, Zhiheng Huang, Yi Yang, and Wei Xu. 2016 · 2016
Earlier work this paper cites.
Learning to Paraphrase for Question Answering. In EMNLP . 875–886
Li Dong, Jonathan Mallinson, Siva Reddy, and Mirella Lapata. 2017 · 2017
Cited alongside, same era.
CNN architectures for large-scale audio classification. In ICASSP
Shawn Hershey, Sourish Chaudhuri, Daniel P. W. Ellis, Jort F. Gemmeke, Aren Jansen, R. Channing Moore, Manoj Plakal, Devin Platt, Rif A. Saurous, Bryan Seybold, Malcolm Slaney, Ron J. Weiss, and Kevin W. Wilson. 2017 · 2017
Cited alongside, same era.
OpenNMT: Open-Source Toolkit for Neural Machine Translation. In ACL
Guillaume Klein, Yoon Kim, Yuntian Deng, Jean Senellart, and Alexander M. Rush. 2017 · 2017
Cited alongside, same era.
A Hierarchical Approach for Generating Descriptive Image Paragraphs. In CVPR
Jonathan Krause, Justin Johnson, Ranjay Krishna, and Li Fei-Fei. 2017 · 2017
Cited alongside, same era.
A Continuously Growing Dataset of Sentential Paraphrases. In EMNLP . 1224–1234
Wuwei Lan, Siyu Qiu, Hua He, and Wei Xu. 2017 · 2017
Cited alongside, same era.
Multimodal Recurrent Model with Attention for Automated Radiology Report Generation. In MICCAI
Yuan Xue, Tao Xu, L. Rodney Long, Zhiyun Xue, Sameer K. Antani, George R. Thoma, and Xiaolei Huang. 2018 · 2018
Later among the works it cites.
Exploring Visual Relationship for Image Captioning. In ECCV
Ting Yao, Yingwei Pan, Yehao Li, and Tao Mei. 2018 · 2018
Later among the works it cites.
End-to-End Dense Video Captioning With Masked Transformer. In CVPR
Luowei Zhou, Yingbo Zhou, Jason J. Corso, Richard Socher, and Caiming Xiong. 2018 · 2018
Later among the works it cites.
Spatio-Temporal Dynamics and Semantic Attribute Enriched Visual Encoding for Video Captioning. In CVPR
Nayyer Aafaq, Naveed Akhtar, Wei Liu, Syed Zulqarnain Gilani, and Ajmal Mian. 2019 · 2019
Later among the works it cites.
Dynamic Layer Aggregation for Neural Machine Translation with Routing-by-Agreement. In AAAI
Zi-Yi Dou, Zhaopeng Tu, Xing Wang, Longyue Wang, Shuming Shi, and Tong Zhang. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Recurrent Topic-Transition GAN for Visual Paragraph Generation. In ICCV
Xiaodan Liang, Zhiting Hu, Hao Zhang, Chuang Gan, and Eric P. Xing. 2017 · 2017
Cited alongside, same era.
Knowing When to Look: Adaptive Attention via a Visual Sentinel for Image Captioning. In CVPR
Jiasen Lu, Caiming Xiong, Devi Parikh, and Richard Socher. 2017 · 2017
Cited alongside, same era.
Self-Critical Sequence Training for Image Captioning. In CVPR
Steven J. Rennie, Etienne Marcheret, Youssef Mroueh, Jarret Ross, and Vaibhava Goel. 2017 · 2017
Cited alongside, same era.
Get To The Point: Summarization with Pointer-Generator Networks. In ACL . 1073–1083
Abigail See, Peter J. Liu, and Christopher D. Manning. 2017 · 2017
Cited alongside, same era.
Inception-v4, Inception-ResNet and the Impact of Residual Connections on Learning. In AAAI
Christian Szegedy, Sergey Ioffe, Vincent Vanhoucke, and Alexander A. Alemi. 2017 · 2017
Cited alongside, same era.
Attention is All you Need. In NIPS
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Show and Tell: Lessons Learned from the 2015 MSCOCO Image Captioning Challenge
Oriol Vinyals, Alexander Toshev, Samy Bengio, and Dumitru Erhan. 2017 · 2017
Cited alongside, same era.
Image Captioning: Transforming Objects into Words. In NeurIPS
Simao Herdade, Armin Kappeler, Kofi Boakye, and Joao Soares. 2019 · 2019
Later among the works it cites.
CheXpert: A Large Chest Radiograph Dataset with Uncertainty Labels and Expert Comparison. In AAAI
Jeremy Irvin, Pranav Rajpurkar, Michael Ko, Yifan Yu, Silviana Ciurea-Ilcus, Chris Chute, Henrik Marklund, Behzad Haghgoo, Robyn L. Ball, Katie S. Shpanskaya, Jayne Seekins, David A. Mong, Safwan S. Halabi, Jesse K. Sandberg, Ricky Jones, David B. Larson, Curtis P. Langlotz, Bhavik N. Patel, Matthew P. Lungren, and Andrew Y. Ng. 2019 · 2019
Later among the works it cites.
Show, Describe and Conclude: On Exploiting the Structure Information of Chest X-ray Reports. In ACL
Baoyu Jing, Zeya Wang, and Eric P. Xing. 2019 · 2019
Later among the works it cites.
MIMIC-CXR: A large publicly available database of labeled chest radiographs
Alistair E. W. Johnson, Tom J. Pollard, Seth J. Berkowitz, Nathaniel R. Greenbaum, Matthew P. Lungren, Chih-ying Deng, Roger G. Mark, and Steven Horng. 2019 · 2019
Later among the works it cites.
fairseq: A Fast, Extensible Toolkit for Sequence Modeling. In NAACL-HLT
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
The Evolved Transformer. In ICML
David R. So, Quoc V. Le, and Chen Liang. 2019 · 2019
Later among the works it cites.
Pay Less Attention with Lightweight and Dynamic Convolutions. In ICLR
Felix Wu, Angela Fan, Alexei Baevski, Yann N. Dauphin, and Michael Auli. 2019 · 2019
Later among the works it cites.
Auto-Encoding Scene Graphs for Image Captioning. In CVPR
Xu Yang, Kaihua Tang, Hanwang Zhang, and Jianfei Cai. 2019 · 2019
Later among the works it cites.
Automatic Radiology Report Generation Based on Multi-view Image Fusion and Medical Concept Enrichment. In MICCAI
Jianbo Yuan, Haofu Liao, Rui Luo, and Jiebo Luo. 2019 · 2019
Later among the works it cites.
Fixup Initialization: Residual Learning Without Normalization. In ICLR
Hongyi Zhang, Yann N. Dauphin, and Tengyu Ma. 2019 · 2019
Later among the works it cites.
Generating Radiology Reports via Memory-driven Transformer. In EMNLP
Zhihong Chen, Yan Song, Tsung-Hui Chang, and Xiang Wan. 2020 · 2020
Closest in time.
Meshed-Memory Transformer for Image Captioning. In CVPR
Marcella Cornia, Matteo Stefanini, Lorenzo Baraldi, and Rita Cucchiara. 2020 · 2020
Closest in time.
Solving high-dimensional parameter inference: marginal posterior densities & Moment Networks. In NeurIPS (Workshop)
Niall Jeffrey and Benjamin D. Wandelt. 2020 · 2020
Closest in time.
Pre-training via Paraphrasing. In NeurIPS
Mike Lewis, Marjan Ghazvininejad, Gargi Ghosh, Armen Aghajanyan, Sida Wang, and Luke Zettlemoyer. 2020 · 2020
Closest in time.
Neuron Interaction Based Representation Composition for Neural Machine Translation. In AAAI
Jian Li, Xing Wang, Baosong Yang, Shuming Shi, Michael R. Lyu, and Zhaopeng Tu. 2020 · 2020
Closest in time.
Prophet Attention: Predicting Attention with Future Attention. In NeurIPS
Fenglin Liu, Xuancheng Ren, Xian Wu, Shen Ge, Wei Fan, Yuexian Zou, and Xu Sun. 2020 · 2020
Closest in time.
STAT: Spatial-Temporal Attention Mechanism for Video Captioning
Chenggang Yan, Yunbin Tu, Xingzheng Wang, Yongbing Zhang, Xinhong Hao, Yongdong Zhang, and Qionghai Dai. 2020 · 2020
Closest in time.
Syntax-Aware Action Targeting for Video Captioning. In CVPR
Qi Zheng, Chaoyue Wang, and Dacheng Tao. 2020 · 2020
Closest in time.
Cross-modal Memory Networks for Radiology Report Generation. In ACL/IJCNLP
Zhihong Chen, Yaling Shen, Yan Song, and Xiang Wan. 2021 · 2021
Closest in time.
Exploring and Distilling Posterior and Prior Knowledge for Radiology Report Generation. In CVPR
Fenglin Liu, Xian Wu, Shen Ge, Wei Fan, and Yuexian Zou. 2021 · 2021
Closest in time.
Attention Calibration for Transformer in Neural Machine Translation. In ACL/IJCNLP
Yu Lu, Jiali Zeng, Jiajun Zhang, Shuangzhi Wu, and Mu Li. 2021 · 2021
Closest in time.
Prevent the Language Model from being Overconfident in Neural Machine Translation. In ACL/IJCNLP
Mengqi Miao, Fandong Meng, Yijin Liu, Xiao-Hua Zhou, and Jie Zhou. 2021 · 2021
Closest in time.
Improving Factual Completeness and Consistency of Image-to-Text Radiology Report Generation. In NAACL-HLT
Yasuhide Miura, Yuhao Zhang, Emily Bao Tsai, Curtis P. Langlotz, and Dan Jurafsky. 2021 · 2021
Closest in time.
Semantic Grouping Network for Video Captioning. In AAAI
Hobin Ryu, Sunghun Kang, Haeyong Kang, and Chang D. Yoo. 2021 · 2021
Closest in time.
RSTNet: Captioning With Adaptive Attention on Visual and Non-Visual Words. In CVPR
Xuying Zhang, Xiaoshuai Sun, Yunpeng Luo, Jiayi Ji, Yiyi Zhou, Yongjian Wu, Feiyue Huang, and Rongrong Ji. 2021 · 2021
Closest in time.