Fetching the paper…
Reading the bibliography…
With direct access to human-written reference as memory, retrieval-augmented generation has achieved much progress in a wide range of text generation tasks.
Statistical significance tests for machine translation evaluation
Philipp Koehn · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin · 2004
Earlier work this paper cites.
A smorgasbord of features for statistical machine translation
Franz Josef Och, Daniel Gildea, Sanjeev Khudanpur, Anoop Sarkar, Kenji Yamada, Alexander M. Fraser, Shankar Kumar, Libin Shen, David Smith, Katherine Eng, Viren Jain, Zhen Jin, and Dragomir R. Radev · 2004
Earlier work this paper cites.
Discriminative reranking for machine translation
Libin Shen, Anoop Sarkar, and Franz Josef Och · 2004
Earlier work this paper cites.
Coarse-to-fine n-best parsing and maxent discriminative reranking
Eugene Charniak and Mark Johnson · 2005
Earlier work this paper cites.
Discriminative reranking for natural language parsing
Michael Collins and Terry Koo · 2005
Earlier work this paper cites.
Re-evaluating the role of Bleu in machine translation research
Chris Callison-Burch, Miles Osborne, and Philipp Koehn · 2006
Earlier work this paper cites.
The jrc-acquis: A multilingual aligned parallel corpus with 20+ languages
Ralf Steinberger, Bruno Pouliquen, Anna Widiger, Camelia Ignat, Tomaz Erjavec, Dan Tufis, and Dániel Varga · 2006
Earlier work this paper cites.
The probabilistic relevance framework: BM25 and beyond
Stephen E. Robertson and Hugo Zaragoza · 2009
Earlier work this paper cites.
Phrase-based machine translation in a computer-assisted translation environment
Michel Simard and Pierre Isabelle · 2009
Earlier work this paper cites.
The effect of translation memory databases on productivity
Masaru Yamada · 2011
Earlier work this paper cites.
Locally training the log-linear model for SMT
Lemao Liu, Hailong Cao, Taro Watanabe, Tiejun Zhao, Mo Yu, and Conghui Zhu · 2012
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio · 2015
Earlier work this paper cites.
A neural conversational model
Oriol Vinyals and Quoc V. Le · 2015
Earlier work this paper cites.
Ask me anything: Dynamic memory networks for natural language processing
Ankit Kumar, Ozan Irsoy, Peter Ondruska, Mohit Iyyer, James Bradbury, Ishaan Gulrajani, Victor Zhong, Romain Paulus, and Richard Socher · 2016
Earlier work this paper cites.
A diversity-promoting objective function for neural conversation models
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan · 2016
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch · 2016
Earlier work this paper cites.
Two are better than one: An ensemble of retrieval- and generation-based dialog systems
Yiping Song, Rui Yan, Xiang Li, Dongyan Zhao, and Ming Zhang · 2016
Earlier work this paper cites.
Reading wikipedia to answer open-domain questions
Danqi Chen, Adam Fisch, Jason Weston, and Antoine Bordes · 2017
Earlier work this paper cites.
Ensemble and reranking: Using multiple models in the NICT-2 neural machine translation system at WAT2017
Kenji Imamura and Eiichiro Sumita · 2017
Earlier work this paper cites.
Dailydialog: A manually labelled multi-turn dialogue dataset
Yanran Li, Hui Su, Xiaoyu Shen, Wenjie Li, Ziqiang Cao, and Shuzi Niu · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Sogou neural machine translation systems for WMT17
Yuguang Wang, Shanbo Cheng, Liyang Jiang, Jiajun Yang, Wei Chen, Muze Li, Lin Shi, Yanfeng Wang, and Hongtao Yang · 2017
Earlier work this paper cites.
Encoding gated translation memory into neural machine translation
Qian Cao and Deyi Xiong · 2018
Earlier work this paper cites.
Search engine guided neural machine translation
Jiatao Gu, Yong Wang, Kyunghyun Cho, and Victor O. K. Li · 2018
Earlier work this paper cites.
A retrieve-and-edit framework for predicting structured outputs
Tatsunori B. Hashimoto, Kelvin Guu, Yonatan Oren, and Percy Liang · 2018
Earlier work this paper cites.
A comparable study on model averaging, ensembling and reranking in NMT
Yuchen Liu, Long Zhou, Yining Wang, Yang Zhao, Jiajun Zhang, and Chengqing Zong · 2018
Earlier work this paper cites.
Don’t give me the details, just the summary! topic-aware convolutional neural networks for extreme summarization
Shashi Narayan, Shay B. Cohen, and Mirella Lapata · 2018
Earlier work this paper cites.
A call for clarity in reporting BLEU scores
Matt Post · 2018
Earlier work this paper cites.
Adafactor: Adaptive learning rates with sublinear memory cost
Noam Shazeer and Mitchell Stern · 2018
Earlier work this paper cites.
Retrieve and refine: Improved sequence generation models for dialogue
Jason Weston, Emily Dinan, and Alexander H. Miller · 2018
Earlier work this paper cites.
Guiding neural machine translation with retrieved translation pieces
Jingyi Zhang, Masao Utiyama, Eiichiro Sumita, Graham Neubig, and Satoshi Nakamura · 2018
Earlier work this paper cites.
Skeleton-to-response: Dialogue generation guided by retrieval memory
Deng Cai, Yan Wang, Wei Bi, Zhaopeng Tu, Xiaojiang Liu, Wai Lam, and Shuming Shi · 2019
Earlier work this paper cites.
Retrieval-guided dialogue response generation via a matching-to-generation framework
Deng Cai, Yan Wang, Wei Bi, Zhaopeng Tu, Xiaojiang Liu, and Shuming Shi · 2019
Earlier work this paper cites.
Implicit deep latent variable models for text generation
Le Fang, Chunyuan Li, Jianfeng Gao, Wen Dong, and Changyou Chen · 2019
Earlier work this paper cites.
Text summarization with pretrained encoders
Yang Liu and Mirella Lapata · 2019
Cited alongside, same era.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov · 2019
Cited alongside, same era.
Text generation with exemplar-based adaptive decoding
Hao Peng, Ankur P. Parikh, Manaal Faruqui, Bhuwan Dhingra, and Dipanjan Das · 2019
Cited alongside, same era.
BIGPATENT: A large-scale dataset for abstractive and coherent summarization
Eva Sharma, Chen Li, and Lu Wang · 2019
Cited alongside, same era.
Response generation by context-aware prototype editing
Yu Wu, Furu Wei, Shaohan Huang, Yunli Wang, Zhoujun Li, and Ming Zhou · 2019
Cited alongside, same era.
Graph based translation memory for neural machine translation
Mengzhou Xia, Guoping Huang, Lemao Liu, and Shuming Shi · 2019
Adaptive semiparametric language models
Dani Yogatama, Cyprien de Masson d’Autume, and Lingpeng Kong · 2021
Later among the works it cites.
Stylized dialogue response generation using stylized unpaired texts
Yinhe Zheng, Zikai Chen, Rongsheng Zhang, Shilei Huang, Xiaoxi Mao, and Minlie Huang · 2021
Later among the works it cites.
In-context examples selection for machine translation
Sweta Agrawal, Chunting Zhou, Mike Lewis, Luke Zettlemoyer, and Marjan Ghazvininejad · 2022
Later among the works it cites.
Dialogved: A pre-trained latent variable encoder-decoder model for dialog response generation
Wei Chen, Yeyun Gong, Song Wang, Bolun Yao, Weizhen Qi, Zhongyu Wei, Xiaowu Hu, Bartuer Zhou, Yi Mao, Weizhu Chen, Biao Cheng, and Nan Duan · 2022
Later among the works it cites.
Target-aware abstractive related work generation with contrastive learning
Xiuying Chen, Hind Alamro, Mingzhe Li, Shen Gao, Rui Yan, Xin Gao, and Xiangliang Zhang · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
PLATO: pre-trained dialogue generation model with discrete latent variable
Siqi Bao, Huang He, Fan Wang, Hua Wu, and Haifeng Wang · 2020
Cited alongside, same era.
Unsupervised cross-lingual representation learning at scale
Alexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary, Guillaume Wenzek, Francisco Guzmán, Edouard Grave, Myle Ott, Luke Zettlemoyer, and Veselin Stoyanov · 2020
Cited alongside, same era.
Residual energy-based models for text generation
Yuntian Deng, Anton Bakhtin, Myle Ott, Arthur Szlam, and Marc’Aurelio Ranzato · 2020
Cited alongside, same era.
Retrieval augmented language model pre-training
Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Ming-Wei Chang · 2020
Cited alongside, same era.
Paraphrase generation by learning how to edit from samples
Amirhossein Kazemnejad, Mohammadreza Salehi, and Mahdieh Soleymani Baghshah · 2020
Cited alongside, same era.
Generalization through memorization: Nearest neighbor language models
Urvashi Khandelwal, Omer Levy, Dan Jurafsky, Luke Zettlemoyer, and Mike Lewis · 2020
Cited alongside, same era.
Xiuying Chen, Mingzhe Li, Xin Gao, and Xiangliang Zhang · 2022
Later among the works it cites.
Neural machine translation with contrastive translation memories
Xin Cheng, Shen Gao, Lemao Liu, Dongyan Zhao, and Rui Yan · 2022
Later among the works it cites.
CORE: A retrieve-then-edit framework for counterfactual data generation
Tanay Dixit, Bhargavi Paranjape, Hannaneh Hajishirzi, and Luke Zettlemoyer · 2022
Later among the works it cites.
Reciprocal learning of knowledge retriever and response ranker for knowledge-grounded conversations
Jiazhan Feng, Chongyang Tao, Zhen Li, Chang Liu, Tao Shen, and Dongyan Zhao · 2022
Later among the works it cites.
There are a thousand hamlets in a thousand people’s eyes: Enhancing knowledge-grounded dialogue with personal memory
Tingchen Fu, Xueliang Zhao, Chongyang Tao, Ji-Rong Wen, and Rui Yan · 2022
Later among the works it cites.
Controllable dialogue simulation with in-context learning
Zekun Li, Wenhu Chen, Shiyang Li, Hong Wang, Jing Qian, and Xifeng Yan · 2022
Later among the works it cites.
Few-shot learning with multilingual generative language models
Xi Victoria Lin, Todor Mihaylov, Mikel Artetxe, Tianlu Wang, Shuohui Chen, Daniel Simig, Myle Ott, Naman Goyal, Shruti Bhosale, Jingfei Du, Ramakanth Pasunuru, Sam Shleifer, Punit Singh Koura, Vishrav Chaudhary, Brian O’Horo, Jeff Wang, Luke Zettlemoyer, Zornitsa Kozareva, Mona T. Diab, Veselin Stoyanov, and Xian Li · 2022
Later among the works it cites.
What makes good in-context examples for gpt-3?
Jiachang Liu, Dinghan Shen, Yizhe Zhang, Bill Dolan, Lawrence Carin, and Weizhu Chen · 2022
Later among the works it cites.
BRIO: bringing order to abstractive summarization
Yixin Liu, Pengfei Liu, Dragomir R. Radev, and Graham Neubig · 2022
Later among the works it cites.
Retrieval augmented classification for long-tail visual recognition
Alexander Long, Wei Yin, Thalaiyasingam Ajanthan, Vu Nguyen, Pulak Purkait, Ravi Garg, Alan Blair, Chunhua Shen, and Anton van den Hengel · 2022
Later among the works it cites.
Learning confidence for transformer-based neural machine translation
Yu Lu, Jiali Zeng, Jiajun Zhang, Shuangzhi Wu, and Mu Li · 2022
Later among the works it cites.
Summareranker: A multi-task mixture-of-experts re-ranking framework for abstractive summarization
Mathieu Ravaut, Shafiq R. Joty, and Nancy F. Chen · 2022
Later among the works it cites.
Towards summary candidates fusion
Mathieu Ravaut, Shafiq R. Joty, and Nancy F. Chen · 2022
Later among the works it cites.
Joint generator-ranker learning for natural language generation
Weizhou Shen, Yeyun Gong, Yelong Shen, Song Wang, Xiaojun Quan, Nan Duan, and Weizhu Chen · 2022
Later among the works it cites.
Training data is more valuable than you think: A simple and effective method by retrieving from training data
Shuohang Wang, Yichong Xu, Yuwei Fang, Yang Liu, Siqi Sun, Ruochen Xu, Chenguang Zhu, and Michael Zeng · 2022
Later among the works it cites.
Retrieval-augmented multimodal language modeling
Michihiro Yasunaga, Armen Aghajanyan, Weijia Shi, Rich James, Jure Leskovec, Percy Liang, Mike Lewis, Luke Zettlemoyer, and Wen-tau Yih · 2022
Later among the works it cites.
Generate-and-retrieve: Use your predictions to improve retrieval for semantic parsing
Yury Zemlyanskiy, Michiel de Jong, Joshua Ainslie, Panupong Pasupat, Peter Shaw, Linlu Qiu, Sumit Sanghai, and Fei Sha · 2022
Later among the works it cites.
Frequency-aware contrastive learning for neural machine translation
Tong Zhang, Wei Ye, Baosong Yang, Long Zhang, Xingzhang Ren, Dayiheng Liu, Jinan Sun, Shikun Zhang, Haibo Zhang, and Wen Zhao · 2022
Later among the works it cites.
Towards efficient dialogue pre-training with transferable and interpretable latent structure
Xueliang Zhao, Lemao Liu, Tingchen Fu, Shuming Shi, Dongyan Zhao, and Rui Yan · 2022
Later among the works it cites.
Training language models with memory augmentation
Zexuan Zhong, Tao Lei, and Danqi Chen · 2022
Later among the works it cites.
A topic-aware summarization framework with different modal side information
Xiuying Chen, Mingzhe Li, Shen Gao, Xin Cheng, Qiang Yang, Qishen Zhang, Xin Gao, and Xiangliang Zhang · 2023
Closest in time.
Xin Cheng, Shen Gao, Yuchi Zhang, Yongliang Wang, Xiuying Chen, Mingzhe Li, Dongyan Zhao, and Rui Yan · 2023
Closest in time.
Decouple knowledge from paramters for plug-and-play language modeling
Xin Cheng, Yankai Lin, Xiuying Chen, Dongyan Zhao, and Rui Yan · 2023
Closest in time.
Scale: Synergized collaboration of asymmetric language translation engines, 2023
Xin Cheng, Xun Wang, Tao Ge, Si-Qing Chen, Furu Wei, Dongyan Zhao, and Rui Yan · 2023
Closest in time.
GPT-4 technical report
OpenAI · 2023
Closest in time.
Replug: Retrieval-augmented black-box language models, 2023
Weijia Shi, Sewon Min, Michihiro Yasunaga, Minjoon Seo, Rich James, Mike Lewis, Luke Zettlemoyer, and Wen tau Yih · 2023
Closest in time.
REPLUG: retrieval-augmented black-box language models
Weijia Shi, Sewon Min, Michihiro Yasunaga, Minjoon Seo, Rich James, Mike Lewis, Luke Zettlemoyer, and Wen-tau Yih · 2023
Closest in time.
In-context learning as maintaining coherency: A study of on-the-fly machine translation using large language models
Suzanna Sia and Kevin Duh · 2023
Closest in time.
Prompting palm for translation: Assessing strategies and performance, 2023
David Vilar, Markus Freitag, Colin Cherry, Jiaming Luo, Viresh Ratnakar, and George Foster · 2023
Closest in time.
Generate rather than retrieve: Large language models are strong context generators, 2023
Wenhao Yu, Dan Iter, Shuohang Wang, Yichong Xu, Mingxuan Ju, Soumya Sanyal, Chenguang Zhu, Michael Zeng, and Meng Jiang · 2023
Closest in time.