Fetching the paper…
Reading the bibliography…
As Large Language Models (LLMs) become popular, there emerged an important trend of using multimodality to augment the LLMs' generation ability, which enables LLMs to better interact with the world.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Using local knowledge graph construction to scale seq2seq models to multi-document inputs
Angela Fan, Claire Gardent, Chloé Braud, and Antoine Bordes. 2019 · 1910
Earlier work this paper cites.
Worldly wise (WoW) - cross-lingual knowledge fusion for fact-based visual spoken-question answering
Kiran Ramnath, Leda Sari, Mark Hasegawa-Johnson, and Chang Yoo. 2021 · 1919
Earlier work this paper cites.
Generating synthetic speech from SpokenVocab for speech translation
Jinming Zhao, Gholamreza Haffari, and Ehsan Shareghi. 2023a · 1981
Earlier work this paper cites.
Retrieval, analogy, and composition: A framework for compositional generalization in image captioning
Zhan Shi, Hui Liu, Martin Renqiang Min, Christopher Malon, Li Erran Li, and Xiaodan Zhu. 2021 · 2000
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oğuz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020a · 2004
Earlier work this paper cites.
Code complete
Steve McConnell. 2004 · 2004
Earlier work this paper cites.
The isabelle framework
Makarius Wenzel, Lawrence C. Paulson, and Tobias Nipkow. 2008 · 2008
Earlier work this paper cites.
Yuma Koizumi, Yasunori Ohishi, Daisuke Niizumi, Daiki Takeuchi, and Masahiro Yasuda. 2020 · 2012
Earlier work this paper cites.
Mining of Massive Datasets, 2nd Ed
Jure Leskovec, Anand Rajaraman, and Jeffrey D. Ullman. 2014 · 2014
Earlier work this paper cites.
Do the fix ingredients already exist? an empirical inquiry into the redundancy assumptions of program repair approaches
Matias Martinez, Westley Weimer, and Monperrus Martin. 2014 · 2014
Earlier work this paper cites.
The strength of random search on automated program repair
Yuhua Qi, Xiaoguang Mao, Yan Lei, Ziying Dai, and Chengsong Wang. 2014 · 2014
Earlier work this paper cites.
Modeling and discovering vulnerabilities with code property graphs
Fabian Yamaguchi, Nico Golde, Daniel Arp, and Konrad Rieck. 2014 · 2014
Earlier work this paper cites.
Combining relevance language modeling and clarity measure for extractive speech summarization
Shih-Hung Liu, Kuan-Yu Chen, Berlin Chen, Hsin-Min Wang, Hsu-Chun Yen, and Wen-Lian Hsu. 2015 · 2015
Earlier work this paper cites.
Jointly modeling deep video and compositional text to bridge vision and language in a unified framework
Ran Xu, Caiming Xiong, Wei Chen, and Jason Corso. 2015 · 2015
Earlier work this paper cites.
Ambient search: A document retrieval system for speech streams
Benjamin Milde, Jonas Wacker, Stefan Radomski, Max Mühlhäuser, and Chris Biemann. 2016a · 2016
Earlier work this paper cites.
Demonstrating ambient search: Implicit document retrieval for speech streams
Benjamin Milde, Jonas Wacker, Stefan Radomski, Max Mühlhäuser, and Chris Biemann. 2016b · 2016
Earlier work this paper cites.
Grad-cam: Visual explanations from deep networks via gradient-based localization
Ramprasaath R. Selvaraju, Michael Cogswell, Abhishek Das, Ramakrishna Vedantam, Devi Parikh, and Dhruv Batra. 2017 · 2017
Earlier work this paper cites.
What do developers search for on the web?
Xin Xia, Lingfeng Bao, David Lo, Pavneet Singh Kochhar, Ahmed E. Hassan, and Zhenchang Xing. 2017 · 2017
Earlier work this paper cites.
Multimodal machine learning: A survey and taxonomy
Tadas Baltrušaitis, Chaitanya Ahuja, and Louis-Philippe Morency. 2018 · 2018
Earlier work this paper cites.
Search engine guided neural machine translation
Jiatao Gu, Yong Wang, Kyunghyun Cho, and Victor OK Li. 2018 · 2018
Earlier work this paper cites.
Retrieve and re-rank: A simple and effective IR approach to simple question answering over knowledge graphs
Vishal Gupta, Manoj Chinnakotla, and Manish Shrivastava. 2018 · 2018
Earlier work this paper cites.
A retrieve-and-edit framework for predicting structured outputs
Tatsunori B Hashimoto, Kelvin Guu, Yonatan Oren, and Percy S Liang. 2018 · 2018
Earlier work this paper cites.
Retrieval-based neural code generation
Shirley Anugrah Hayati, Raphael Olivier, Pravalika Avvaru, Pengcheng Yin, Anthony Tomasic, and Graham Neubig. 2018 · 2018
Earlier work this paper cites.
Shaping program repair space with existing patches and similar code
Jiajun Jiang, Yingfei Xiong, Hongyu Zhang, Qing Gao, and Xiangqun Chen. 2018 · 2018
Earlier work this paper cites.
Mining stackoverflow for program repair
Xuliang Liu and Hao Zhong. 2018 · 2018
Earlier work this paper cites.
Video captioning with multi-faceted attention
Xiang Long, Chuang Gan, and Gerard De Melo. 2018 · 2018
Earlier work this paper cites.
Learning joint embedding with multimodal cues for cross-modal video-text retrieval
Niluthpol Chowdhury Mithun, Juncheng Li, Florian Metze, and Amit K Roy-Chowdhury. 2018 · 2018
Earlier work this paper cites.
Game-based video-context dialogue
Ramakanth Pasunuru and Mohit Bansal. 2018 · 2018
Earlier work this paper cites.
Retrieve and refine: Improved sequence generation models for dialogue
Jason Weston, Emily Dinan, and Alexander Miller. 2018 · 2018
Earlier work this paper cites.
Incorporating background knowledge into video description generation
Spencer Whitehead, Heng Ji, Mohit Bansal, Shih-Fu Chang, and Clare Voss. 2018 · 2018
Earlier work this paper cites.
Guiding neural machine translation with retrieved translation pieces
Jingyi Zhang, Masao Utiyama, Eiichro Sumita, Graham Neubig, and Satoshi Nakamura. 2018 · 2018
Earlier work this paper cites.
Skeleton-to-response: Dialogue generation guided by retrieval memory
Deng Cai, Yan Wang, Wei Bi, Zhaopeng Tu, Xiaojiang Liu, Wai Lam, and Shuming Shi. 2019a · 2019
Earlier work this paper cites.
Retrieval-guided dialogue response generation via a matching-to-generation framework
Deng Cai, Yan Wang, Wei Bi, Zhaopeng Tu, Xiaojiang Liu, and Shuming Shi. 2019b · 2019
Earlier work this paper cites.
Extract, transform and filling: A pipeline model for question paraphrasing based on template
Yunfan Gu, Yang Yuqiao, and Zhongyu Wei. 2019 · 2019
Earlier work this paper cites.
TextGraphs 2019 shared task on multi-hop inference for explanation regeneration
Peter Jansen and Dmitry Ustalov. 2019 · 2019
Earlier work this paper cites.
Knowledge-driven encode, retrieve, paraphrase for medical image report generation
Christy Y. Li, Xiaodan Liang, Zhiting Hu, and Eric P. Xing. 2019 · 2019
Earlier work this paper cites.
Text generation with exemplar-based adaptive decoding
Hao Peng, Ankur Parikh, Manaal Faruqui, Bhuwan Dhingra, and Dipanjan Das. 2019 · 2019
Earlier work this paper cites.
Videobert: A joint model for video and language representation learning
Chen Sun, Austin Myers, Carl Vondrick, Kevin Murphy, and Cordelia Schmid. 2019 · 2019
Earlier work this paper cites.
Multimodal transformer for unaligned multimodal language sequences
Yao-Hung Hubert Tsai, Shaojie Bai, Paul Pu Liang, J Zico Kolter, Louis-Philippe Morency, and Ruslan Salakhutdinov. 2019 · 2019
Earlier work this paper cites.
Sorting and transforming program repair ingredients via deep learning code similarities
Martin White, Michele Tufano, Matias Martinez, Monperrus Martin, and Denys Poshyvanyk. 2019 · 2019
Earlier work this paper cites.
Response generation by context-aware prototype editing
Yu Wu, Furu Wei, Shaohan Huang, Yunli Wang, Zhoujun Li, and Ming Zhou. 2019 · 2019
Earlier work this paper cites.
Modeling graph structure in transformer for better AMR-to-text generation
Jie Zhu, Junhui Li, Muhua Zhu, Longhua Qian, Min Zhang, and Guodong Zhou. 2019 · 2019
Earlier work this paper cites.
AirConcierge: Generating task-oriented dialogue via efficient large-scale knowledge retrieval
Chieh-Yang Chen, Pei-Hsin Wang, Shih-Chieh Chang, Da-Cheng Juan, Wei Wei, and Jia-Yu Pan. 2020 · 2020
Earlier work this paper cites.
ENT-DESC: Entity description generation by exploring knowledge graph
Liying Cheng, Dekun Wu, Lidong Bing, Yan Zhang, Zhanming Jie, Wei Lu, and Luo Si. 2020 · 2020
Earlier work this paper cites.
A survey on deep learning for multimodal data fusion
Jing Gao, Peng Li, Zhikui Chen, and Jianing Zhang. 2020 · 2020
Earlier work this paper cites.
Filtering before iteratively referring for knowledge-grounded response selection in retrieval-based chatbots
Jia-Chen Gu, Zhenhua Ling, Quan Liu, Zhigang Chen, and Xiaodan Zhu. 2020 · 2020
Earlier work this paper cites.
AttnIO: Knowledge Graph Exploration with In-and-Out Attention Flow for Knowledge-Grounded Dialogue
Jaehun Jung, Bokyung Son, and Sungwon Lyu. 2020 · 2020
Earlier work this paper cites.
Dense passage retrieval for open-domain question answering
Vladimir Karpukhin, Barlas Oguz, Sewon Min, Patrick Lewis, Ledell Wu, Sergey Edunov, Danqi Chen, and Wen-tau Yih. 2020b · 2020
Earlier work this paper cites.
BiST: Bi-directional spatio-temporal reasoning for video-grounded dialogues
Hung Le, Doyen Sahoo, Nancy Chen, and Steven C.H. Hoi. 2020 · 2020
Earlier work this paper cites.
TVQA+: Spatio-temporal grounding for video question answering
Jie Lei, Licheng Yu, Tamara Berg, and Mohit Bansal. 2020 · 2020
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al. 2020 · 2020
Earlier work this paper cites.
A case study of NLG from multimedia data sources: Generating architectural landmark descriptions
Simon Mille, Spyridon Symeonidis, Maria Rousi, Montserrat Marimon Felipe, Klearchos Stavrothanasopoulos, Petros Alvanitopoulos, Roberto Carlini Salguero, Jens Grivolla, Georgios Meditskos, Stefanos Vrochidis, and Leo Wanner. 2020 · 2020
Earlier work this paper cites.
Deep composer: Deep neural hashing and retrieval approach to automatic music generation
Brandon Royal, Kien Hua, and Brenton Zhang. 2020 · 2020
Earlier work this paper cites.
Improving knowledge-aware dialogue response generation by using human-written prototype dialogues
Sixing Wu, Ying Li, Dawei Zhang, and Zhonghai Wu. 2020a · 2020
Earlier work this paper cites.
Boosting neural machine translation with similar translations
Jitao Xu, Josep-Maria Crego, and Jean Senellart. 2020 · 2020
Earlier work this paper cites.
Retrieval-based neural source code summarization
Jian Zhang, Xu Wang, Hongyu Zhang, Hailong Sun, and Xudong Liu. 2020 · 2020
Cited alongside, same era.
Unified vision-language pre-training for image captioning and vqa
Luowei Zhou, Hamid Palangi, Lei Zhang, Houdong Hu, Jason Corso, and Jianfeng Gao. 2020 · 2020
Cited alongside, same era.
Explanations for CommonsenseQA: New Dataset and Models
Shourya Aggarwal, Divyanshu Mandowara, Vishwajeet Agrawal, Dinesh Khandelwal, Parag Singla, and Dinesh Garg. 2021 · 2021
Cited alongside, same era.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, et al. 2021 · 2021
Cited alongside, same era.
Long-range modeling of source code files with ewash: Extended window access by syntax hierarchy
Colin B Clement, Shuai Lu, Xiaoyu Liu, Michele Tufano, Dawn Drain, Nan Duan, Neel Sundaresan, and Alexey Svyatkovskiy. 2021 · 2021
Repair is nearly generation: Multilingual program repair with llms
Harshit Joshi, José Cambronero, Sumit Gulwani, Vu Le, Ivan Radicek, and Gust Verbruggen. 2022 · 2022
Later among the works it cites.
Prompting visual-language models for efficient video understanding
Chen Ju, Tengda Han, Kunhao Zheng, Ya Zhang, and Weidi Xie. 2022 · 2022
Later among the works it cites.
Audio retrieval with natural language queries: A benchmark study
A Sophia Koepke, Andreea-Maria Oncescu, Joao Henriques, Zeynep Akata, and Samuel Albanie. 2022 · 2022
Later among the works it cites.
Vgnmn: Video-grounded neural module networks for video-grounded dialogue systems
Hung Le, Nancy Chen, and Steven Hoi. 2022 · 2022
Later among the works it cites.
Knowledge-grounded dialogue generation with a unified knowledge representation
Yu Li, Baolin Peng, Yelong Shen, Yi Mao, Lars Liden, Zhou Yu, and Jianfeng Gao. 2022c · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Toward faithful case-based reasoning through learning prototypes in a nearest neighbor-friendly space
Seyed Omid Davoudi and Majid Komeili. 2021 · 2021
Cited alongside, same era.
Neural path hunter: Reducing hallucination in dialogue systems via path grounding
Nouha Dziri, Andrea Madotto, Osmar Zaïane, and Avishek Joey Bose. 2021 · 2021
Cited alongside, same era.
Augmenting transformers with KNN-based composite memory for dialog
Angela Fan, Claire Gardent, Chloé Braud, and Antoine Bordes. 2021 · 2021
Cited alongside, same era.
Space efficient context encoding for non-task-oriented dialogue generation with graph attention transformer
Fabian Galetzka, Jewgeni Rose, David Schlangen, and Jens Lehmann. 2021 · 2021
Cited alongside, same era.
Condenser: a pre-training architecture for dense retrieval
Luyu Gao and Jamie Callan. 2021 · 2021
Cited alongside, same era.
SimCSE: Simple contrastive learning of sentence embeddings
Tianyu Gao, Xingcheng Yao, and Danqi Chen. 2021 · 2021
Cited alongside, same era.
Fast and accurate neural machine translation with translation memory
Qiuxiang He, Guoping Huang, Qu Cui, Li Li, and Lemao Liu. 2021 · 2021
Cited alongside, same era.
Automating code review activities by large-scale pre-training
Zhiyu Li, Shuai Lu, Daya Guo, Nan Duan, Shailesh Jannu, Grant Jenks, Deep Majumder, Jared Green, Alexey Svyatkovskiy, Shengyu Fu, and Neel Sundaresan. 2022d · 2022
Later among the works it cites.
Retrieval augmented visual question answering with outside knowledge
Weizhe Lin and Bill Byrne. 2022b · 2022
Later among the works it cites.
Uni-parser: Unified semantic parser for question answering on knowledge base and database
Ye Liu, Semih Yavuz, Rui Meng, Dragomir Radev, Caiming Xiong, and Yingbo Zhou. 2022b · 2022
Later among the works it cites.
Audio-text retrieval in context
Siyu Lou, Xuenan Xu, Mengyue Wu, and Kai Yu. 2022 · 2022
Later among the works it cites.
Open-domain question answering via chain of reasoning over heterogeneous knowledge
Kaixin Ma, Hao Cheng, Xiaodong Liu, Eric Nyberg, and Jianfeng Gao. 2022 · 2022
Later among the works it cites.
Language models of code are few-shot commonsense learners
Aman Madaan, Shuyan Zhou, Uri Alon, Yiming Yang, and Graham Neubig. 2022 · 2022
Later among the works it cites.
HybriDialogue: An information-seeking dialogue dataset grounded on tabular and textual data
Kai Nakamura, Sharon Levy, Yi-Lin Tuan, Wenhu Chen, and William Yang Wang. 2022 · 2022
Later among the works it cites.
FeTaQA: Free-form table question answering
Linyong Nan, Chiachun Hsieh, Ziming Mao, Xi Victoria Lin, Neha Verma, Rui Zhang, Wojciech Kryściński, Hailey Schoelkopf, Riley Kong, Xiangru Tang, Mutethia Mutuma, Ben Rosand, Isabel Trindade, Renusree Bandaru, Jacob Cunningham, Caiming Xiong, Dragomir Radev, and Dragomir Radev. 2022 · 2022
Later among the works it cites.
Entailment tree explanations via iterative retrieval-generation reasoner
Danilo Neves Ribeiro, Shen Wang, Xiaofei Ma, Rui Dong, Xiaokai Wei, Henghui Zhu, Xinchi Chen, Peng Xu, Zhiheng Huang, Andrew Arnold, and Dan Roth. 2022 · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Later among the works it cites.
Dreamfusion: Text-to-3d using 2d diffusion
Ben Poole, Ajay Jain, Jonathan T Barron, and Ben Mildenhall. 2022 · 2022
Later among the works it cites.
Retrieval-augmented transformer for image captioning
Sara Sarto, Marcella Cornia, Lorenzo Baraldi, and Rita Cucchiara. 2022 · 2022
Later among the works it cites.
RACE: Retrieval-augmented commit message generation
Ensheng Shi, Yanlin Wang, Wei Tao, Lun Du, Hongyu Zhang, Shi Han, Dongmei Zhang, and Hongbin Sun. 2022 · 2022
Later among the works it cites.
TIARA: Multi-grained retrieval for robust question answering over large knowledge base
Yiheng Shu, Zhiwei Yu, Yuhan Li, Börje Karlsson, Tingting Ma, Yuzhong Qu, and Chin-Yew Lin. 2022 · 2022
Later among the works it cites.
TegTok: Augmenting text generation via task-specific and open-world knowledge
Chao-Hong Tan, Jia-Chen Gu, Chongyang Tao, Zhen-Hua Ling, Can Xu, Huang Hu, Xiubo Geng, and Daxin Jiang. 2022 · 2022
Later among the works it cites.
Lamda: Language models for dialog applications
Romal Thoppilan, Daniel De Freitas, Jamie Hall, Noam Shazeer, Apoorv Kulshreshtha, Heng-Tze Cheng, Alicia Jin, Taylor Bos, Leslie Baker, Yu Du, et al. 2022 · 2022
Later among the works it cites.
Plug-and-play VQA: Zero-shot VQA by conjoining large pretrained models with zero training
Anthony Meng Huat Tiong, Junnan Li, Boyang Li, Silvio Savarese, and Steven C.H. Hoi. 2022 · 2022
Later among the works it cites.
Interleaving retrieval with chain-of-thought reasoning for knowledge-intensive multi-step questions
Harsh Trivedi, Niranjan Balasubramanian, Tushar Khot, and Ashish Sabharwal. 2022 · 2022
Later among the works it cites.
Language models with image descriptors are strong few-shot video-language learners
Zhenhailong Wang, Manling Li, Ruochen Xu, Luowei Zhou, Jie Lei, Xudong Lin, Shuohang Wang, Ziyi Yang, Chenguang Zhu, Derek Hoiem, et al. 2022 · 2022
Later among the works it cites.
Chain of thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Ed Chi, Quoc Le, and Denny Zhou. 2022 · 2022
Later among the works it cites.
Z-LaVI: Zero-shot language solver fueled by visual imagination
Yue Yang, Wenlin Yao, Hongming Zhang, Xiaoyang Wang, Dong Yu, and Jianshu Chen. 2022a · 2022
Later among the works it cites.
LogicSolver: Towards interpretable math word problem solving with logical prompt-enhanced learning
Zhicheng Yang, Jinghui Qin, Jiaqi Chen, Liang Lin, and Xiaodan Liang. 2022c · 2022
Later among the works it cites.
Retrieval-augmented multimodal language modeling
Michihiro Yasunaga, Armen Aghajanyan, Weijia Shi, Rich James, Jure Leskovec, Percy Liang, Mike Lewis, Luke Zettlemoyer, and Wen-tau Yih. 2022 · 2022
Later among the works it cites.
The unreliability of explanations in few-shot prompting for textual reasoning
Xi Ye and Greg Durrett. 2022 · 2022
Later among the works it cites.
Scaling autoregressive models for content-rich text-to-image generation
Jiahui Yu, Yuanzhong Xu, Jing Yu Koh, Thang Luong, Gunjan Baid, Zirui Wang, Vijay Vasudevan, Alexander Ku, Yinfei Yang, Burcu Karagol Ayan, Ben Hutchinson, Wei Han, Zarana Parekh, Xin Li, Han Zhang, Jason Baldridge, and Yonghui Wu. 2022 · 2022
Later among the works it cites.
Socratic models: Composing zero-shot multimodal reasoning with language
Andy Zeng, Maria Attarian, Brian Ichter, Krzysztof Choromanski, Adrian Wong, Stefan Welker, Federico Tombari, Aveek Purohit, Michael Ryoo, Vikas Sindhwani, et al. 2022 · 2022
Later among the works it cites.
Focus! relevant and sufficient context selection for news image captioning
Mingyang Zhou, Grace Luo, Anna Rohrbach, and Zhou Yu. 2022a · 2022
Later among the works it cites.
Jointly training large autoregressive multimodal models
Emanuele Aiello, Lili Yu, Yixin Nie, Armen Aghajanyan, and Barlas Oguz. 2023 · 2023
Closest in time.
Characterizing attribution and fluency tradeoffs for retrieval-augmented large language models
Renat Aksitov, Chung-Ching Chang, David Reitter, Siamak Shakeri, and Yunhsuan Sung. 2023 · 2023
Closest in time.
Knowledge-augmented language model prompting for zero-shot knowledge graph question answering
Jinheon Baek, Alham Fikri Aji, and Amir Saffari. 2023 · 2023
Closest in time.
David M Chan, Shalini Ghosh, Ariya Rastrow, and Björn Hoffmeister. 2023 · 2023
Closest in time.
Binding language models in symbolic languages
Zhoujun Cheng, Tianbao Xie, Peng Shi, Chengzu Li, Rahul Nadkarni, Yushi Hu, Caiming Xiong, Dragomir Radev, Mari Ostendorf, Luke Zettlemoyer, Noah A. Smith, and Tao Yu. 2023 · 2023
Closest in time.
Palm-e: An embodied multimodal language model
Danny Driess, Fei Xia, Mehdi SM Sajjadi, Corey Lynch, Aakanksha Chowdhery, Brian Ichter, Ayzaan Wahid, Jonathan Tompson, Quan Vuong, Tianhe Yu, et al. 2023 · 2023
Closest in time.
Reveal: Retrieval-augmented visual-language pre-training with multi-source multimodal knowledge memory
Ziniu Hu, Ahmet Iscen, Chen Sun, Zirui Wang, Kai-Wei Chang, Yizhou Sun, Cordelia Schmid, David A Ross, and Alireza Fathi. 2023 · 2023
Closest in time.
Inferfix: End-to-end program repair with llms
Matthew Jin, Syed Shahriar, Michele Tufano, Xin Shi, Shuai Lu, Neel Sundaresan, and Alexey Svyatkovskiy. 2023 · 2023
Closest in time.
Knowledge graph-augmented language models for knowledge-grounded dialogue generation
Minki Kang, Jin Myung Kwak, Jinheon Baek, and Sung Ju Hwang. 2023 · 2023
Closest in time.
Prefix tuning for automated audio captioning
Minkyu Kim, Kim Sung-Bin, and Tae-Hyun Oh. 2023 · 2023
Closest in time.
FVQA 2.0: Introducing adversarial samples into fact-based visual question answering
Weizhe Lin, Zhilin Wang, and Bill Byrne. 2023 · 2023
Closest in time.
The StatCan dialogue dataset: Retrieving data tables through conversations with genuine intents
Xing Han Lu, Siva Reddy, and Harm de Vries. 2023 · 2023
Closest in time.
Faithful chain-of-thought reasoning
Qing Lyu, Shreya Havaldar, Adam Stein, Li Zhang, Delip Rao, Eric Wong, Marianna Apidianaki, and Chris Callison-Burch. 2023 · 2023
Closest in time.
Augmenting pre-trained language models with audio feature embedding for argumentation mining in political debates
Rafael Mestre, Stuart Middleton, Matt Ryan, Masood Gheasi, Timothy Norman, and Jiatong Zhu. 2023 · 2023
Closest in time.
Augmented language models: a survey
Grégoire Mialon, Roberto Dessì, Maria Lomeli, Christoforos Nalmpantis, Ram Pasunuru, Roberta Raileanu, Baptiste Rozière, Timo Schick, Jane Dwivedi-Yu, Asli Celikyilmaz, et al. 2023 · 2023
Closest in time.
Retrieval-based prompt selection for code-related few-shot learning
Noor Nashid, Mifta Sintaha, and Ali Mesbah. 2023 · 2023
Closest in time.
Gpt-4 technical report
OpenAI. 2023 · 2023
Closest in time.
Retrieval-augmented image captioning
Rita Ramos, Desmond Elliott, and Bruno Martins. 2023 · 2023
Closest in time.
Toolformer: Language models can teach themselves to use tools
Timo Schick, Jane Dwivedi-Yu, Roberto Dessì, Roberta Raileanu, Maria Lomeli, Luke Zettlemoyer, Nicola Cancedda, and Thomas Scialom. 2023 · 2023
Closest in time.
Yunhu Ye, Binyuan Hui, Min Yang, Binhua Li, Fei Huang, and Yongbin Li. 2023 · 2023
Closest in time.
RAMM: retrieval-augmented biomedical visual question answering with multi-modal pre-training
Zheng Yuan, Qiao Jin, Chuanqi Tan, Zhengyun Zhao, Hongyi Yuan, Fei Huang, and Songfang Huang. 2023 · 2023
Closest in time.
Style-aware contrastive learning for multi-style image captioning
Yucheng Zhou and Guodong Long. 2023 · 2023
Closest in time.
Visualize before you write: Imagination-guided open-ended text generation
Wanrong Zhu, An Yan, Yujie Lu, Wenda Xu, Xin Wang, Miguel Eckstein, and William Yang Wang. 2023 · 2023
Closest in time.