Fetching the paper…
Reading the bibliography…
Large scale pretrained language models have demonstrated state-of-the-art performance in language understanding tasks.
Vl-bert: Pre-training of generic visual-linguistic representations
Su, W., Zhu, X., Cao, Y., Li, B., Lu, L., Wei, F., and Dai, J · 1908
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
Lin, C.-Y · 2004
Earlier work this paper cites.
Overview of duc 2005
Dang, H. T · 2005
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Fei-Fei, L · 2009
Earlier work this paper cites.
Meteor universal: Language specific translation evaluation for any target language
Denkowski, M. and Lavie, A · 2014
Earlier work this paper cites.
Semi-supervised sequence learning
Dai, A. M. and Le, Q. V · 2015
Earlier work this paper cites.
On using monolingual corpora in neural machine translation
Gulcehre, C., Firat, O., Xu, K., Cho, K., Barrault, L., Lin, H.-C., Bougares, F., Schwenk, H., and Bengio, Y · 2015
Earlier work this paper cites.
Moddrop: adaptive multi-modal gesture recognition
Neverova, N., Wolf, C., Taylor, G., and Nebout, F · 2015
Earlier work this paper cites.
Cider: Consensus-based image description evaluation
Vedantam, R., Lawrence Zitnick, C., and Parikh, D · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Earlier work this paper cites.
Cold fusion: Training seq2seq models together with language models
Sriram, A., Jun, H., Satheesh, S., and Coates, A · 2017
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2018
Cited alongside, same era.
Universal language model fine-tuning for text classification
Howard, J. and Ruder, S · 2018
Cited alongside, same era.
Pre-trained language model representations for language generation
Edunov, S., Baevski, A., and Auli, M · 2019
Cited alongside, same era.
Large-scale transfer learning for natural language generation
Golovanov, S., Kurbanov, R., Nikolenko, S., Truskovskyi, K., Tselousov, A., and Wolf, T · 2019
Language models are unsupervised multitask learners
Radford, A., Wu, J., Child, R., Luan, D., Amodei, D., and Sutskever, I · 2019
Later among the works it cites.
Mass: Masked sequence to sequence pre-training for language generation
Song, K., Tan, X., Qin, T., Lu, J., and Liu, T.-Y · 2019
Later among the works it cites.
Huggingface’s transformers: State-of-the-art natural language processing
Wolf, T., Debut, L., Sanh, V., Chaumond, J., Delangue, C., Moi, A., Cistac, P., Rault, T., Louf, R., Funtowicz, M., et al · 2019
Later among the works it cites.
Simple and effective noisy channel modeling for neural machine translation
Yee, K., Ng, N., Dauphin, Y. N., and Auli, M · 2019
Later among the works it cites.
Pretraining-based natural language generation for text summarization
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Ctrl: A conditional transformer language model for controllable generation
Keskar, N. S., McCann, B., Varshney, L. R., Xiong, C., and Socher, R · 2019
Cited alongside, same era.
Supervised multimodal bitransformers for classifying images and text
Kiela, D., Bhooshan, S., Firooz, H., and Testuggine, D · 2019
Cited alongside, same era.
Visualbert: A simple and performant baseline for vision and language
Li, L. H., Yatskar, M., Yin, D., Hsieh, C.-J., and Chang, K.-W · 2019
Cited alongside, same era.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Lu, J., Batra, D., Parikh, D., and Lee, S · 2019
Cited alongside, same era.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A., Gross, S., Massa, F., Lerer, A., Bradbury, J., Chanan, G., Killeen, T., Lin, Z., Gimelshein, N., Antiga, L., et al · 2019
Cited alongside, same era.
Generalizing question answering system with pre-trained language model fine-tuning
Su, D., Xu, Y., Winata, G. I., Xu, P., Kim, H., Liu, Z., and Fung, P
Cited in the paper.
Zhang, H., Xu, J., and Wang, J · 2019
Later among the works it cites.
Encoder-agnostic adaptation for conditional language generation
Ziegler, Z. M., Melas-Kyriazi, L., Gehrmann, S., and Rush, A. M · 2019
Later among the works it cites.
Language models are few-shot learners
Brown, T. B., Mann, B., Ryder, N., Subbiah, M., Kaplan, J., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al · 2020
Later among the works it cites.
Uniter: Universal image-text representation learning
Chen, Y.-C., Li, L., Yu, L., El Kholy, A., Ahmed, F., Gan, Z., Cheng, Y., and Liu, J · 2020
Later among the works it cites.
Unicoder-vl: A universal encoder for vision and language by cross-modal pre-training
Li, G., Duan, N., Fang, Y., Gong, M., and Jiang, D · 2020
Later among the works it cites.
Fashion captioning: Towards generating accurate descriptions with semantic rewards
Yang, X., Zhang, H., Jin, D., Liu, Y., Wu, C.-H., Tan, J., Xie, D., Wang, J., and Wang, X · 2020
Later among the works it cites.