Fetching the paper…
Reading the bibliography…
LLMs have demonstrated commendable performance across diverse domains.
Hellaswag: Can a machine really finish your sentence?
Rowan Zellers, Ari Holtzman, Yonatan Bisk, Ali Farhadi, and Yejin Choi. 2019 · 1905
Earlier work this paper cites.
WINOGRANDE: an adversarial winograd schema challenge at scale
Keisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, and Yejin Choi. 2019 · 1907
Earlier work this paper cites.
Object oriented programming
Tim Rentsch. 1982 · 1982
Earlier work this paper cites.
Programming languages: design and implementation
Terrence W Pratt, Marvin V Zelkowitz, and Tadepalli V Gopal. 1984 · 1984
Earlier work this paper cites.
Introduction to the Theory of Computation
Michael Sipser. 1996 · 1996
Earlier work this paper cites.
Compilers: principles, techniques and tools
Alfred V Aho. 2007 · 2007
Earlier work this paper cites.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2021 · 2009
Earlier work this paper cites.
Programming Python: Powerful Object-Oriented Programming
Mark Lutz. 2010 · 2010
Earlier work this paper cites.
Markdown: Syntax
John Gruber. 2012 · 2012
Earlier work this paper cites.
Modern compiler design
Dick Grune, Kees Van Reeuwijk, Henri E Bal, et al. 2012 · 2012
Earlier work this paper cites.
Concepts of programming languages
Robert W Sebesta. 2012 · 2012
Earlier work this paper cites.
Foundations of JSON Schema
Felipe Pezoa, Juan L. Reutter, Fernando Suarez, et al. 2016 · 2016
Earlier work this paper cites.
Programming Languages: Types and Features
Chakray. 2018 · 2018
Earlier work this paper cites.
Think you have solved question answering? try arc, the ai2 reasoning challenge
Peter Clark, Isaac Cowhey, Oren Etzioni, Tushar Khot, Ashish Sabharwal, Carissa Schoenick, and Oyvind Tafjord. 2018 · 2018
Earlier work this paper cites.
An Introduction to Language
Victoria Fromkin, Robert Rodman, and Nina Hyams. 2018 · 2018
Earlier work this paper cites.
Training verifiers to solve math word problems
Karl Cobbe, Vineet Kosaraju, Mohammad Bavarian, Mark Chen, Heewoo Jun, Lukasz Kaiser, Matthias Plappert, Jerry Tworek, Jacob Hilton, Reiichiro Nakano, Christopher Hesse, and John Schulman. 2021 · 2021
Earlier work this paper cites.
A framework for few-shot language model evaluation
Leo Gao, Jonathan Tow, Stella Biderman, Sid Black, Anthony DiPofi, Charles Foster, Laurence Golding, Jeffrey Hsu, Kyle McDonell, Niklas Muennighoff, Jason Phang, Laria Reynolds, Eric Tang, Anish Thite, Ben Wang, Kevin Wang, and Andy Zou. 2021 · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Earlier work this paper cites.
LangChain: Building applications with LLMs through composability
Harrison Chase. 2022 · 2022
Earlier work this paper cites.
A complete introduction to prompt engineering for large language models
Mihail Eric. 2022 · 2022
Earlier work this paper cites.
Optimizing Prompts for Text-to-Image Generation
Yaru Hao, Zewen Chi, Li Dong, and Furu Wei. 2022 · 2022
Earlier work this paper cites.
Truthfulqa: Measuring how models mimic human falsehoods
Stephanie Lin, Jacob Hilton, and Owain Evans. 2022 · 2022
Cited alongside, same era.
Design Guidelines for Prompt Engineering Text-to-Image Generative Models
Vivian Liu and Lydia B Chilton. 2022 · 2022
Cited alongside, same era.
Crosslingual generalization through multitask finetuning
Niklas Muennighoff, Thomas Wang, Lintang Sutawika, Adam Roberts, Stella Biderman, Teven Le Scao, M Saiful Bari, Sheng Shen, Zheng-Xin Yong, Hailey Schoelkopf, et al. 2022 · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul F. Christiano, Jan Leike, and Ryan Lowe. 2022 · 2022
Cited alongside, same era.
Decorate the Examples: A Simple Method of Prompt Design for Biomedical Relation Extraction
The AI War: Mastering the Art of Prompt Engineering in the Era of Large Language Models
Shritam Kumar Mund. 2023 · 2023
Later among the works it cites.
ChatGPT3-Free-Prompt-List: A free guide for learning to create ChatGPT3 Prompts
Matt Nigh. 2023 · 2023
Later among the works it cites.
Generative Agents: Interactive Simulacra of Human Behavior
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, et al. 2023 · 2023
Later among the works it cites.
Automatic Prompt Optimization with "Gradient Descent" and Beam Search
Reid Pryzant, Dan Iter, Jerry Li, et al. 2023 · 2023
Later among the works it cites.
Stanford alpaca: An instruction-following llama model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B. Hashimoto. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hui-Syuan Yeh, Thomas Lavergne, and Pierre Zweigenbaum. 2022 · 2022
Cited alongside, same era.
Building the Next Generation of Open-Source and Bilingual LLMs
01-ai. 2023 · 2023
Cited alongside, same era.
Josh Achiam, Steven Adler, Sandhini Agarwal, et al. 2023 · 2023
Cited alongside, same era.
Jinze Bai, Shuai Bai, Yunfei Chu, et al. 2023 · 2023
Cited alongside, same era.
Open llm leaderboard
Edward Beeching, Clémentine Fourrier, Nathan Habib, Sheon Han, Nathan Lambert, Nazneen Rajani, Omar Sanseviero, Lewis Tunstall, and Thomas Wolf. 2023 · 2023
Cited alongside, same era.
Principled Instructions Are All You Need for Questioning LLaMA-1/2, GPT-3.5/4
Sondos Mahmoud Bsharat, Aidar Myrzakhan, and Zhiqiang Shen. 2023 · 2023
Cited alongside, same era.
A study on Prompt Design, Advantages and Limitations of ChatGPT for Deep Learning Program Repair
Jialun Cao, Meiziniu Li, Ming Wen, and Shing-chi Cheung. 2023 · 2023
Cited alongside, same era.
Unleashing the potential of prompt engineering in Large Language Models: a comprehensive review
Banghao Chen, Zhaofeng Zhang, Nicolas Langrené, and Shengxin Zhu. 2023 · 2023
Cited alongside, same era.
An Introduction to Large Language Models: Prompt Engineering and P-Tuning
Tanay Varshney and Annie Surla. 2023 · 2023
Later among the works it cites.
How to use prompt engineering with large language models
Zhenxuan Wang. 2023 · 2023
Later among the works it cites.
Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed Chi, Quoc Le, and Denny Zhou. 2023 · 2023
Later among the works it cites.
AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation
Qingyun Wu, Gagan Bansal, Jieyu Zhang, et al. 2023 · 2023
Later among the works it cites.
The Rise and Potential of Large Language Model Based Agents: A Survey
Zhiheng Xi, Wenxiang Chen, Xin Guo, et al. 2023 · 2023
Later among the works it cites.
Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf
Yuzhuang Xu, Shuo Wang, Peng Li, et al. 2023 · 2023
Later among the works it cites.
Large Language Model as Attributed Training Data Generator: A Tale of Diversity and Bias
Yue Yu, Yuchen Zhuang, Jieyu Zhang, et al. 2023 · 2023
Later among the works it cites.
Why johnny can’t prompt: How non-ai experts try (and fail) to design llm prompts
J.D. Zamfirescu-Pereira, Richmond Y. Wong, Bjoern Hartmann, and Qian Yang. 2023a · 2023
Later among the works it cites.
Why Johnny Can’t Prompt: How Non-AI Experts Try (and Fail) to Design LLM Prompts
J.D. Zamfirescu-Pereira, Richmond Y. Wong, Bjoern Hartmann, and Qian Yang. 2023b · 2023
Later among the works it cites.
Llama 3 model card
AI@Meta. 2024 · 2024
Closest in time.
Prompt Engineering Could Be the Hottest Programming Language of 2024 — Here’s Why
Nabil Alouani. 2023 · 2024
Closest in time.
Zheng Cai, Maosong Cao, Haojiong Chen, Kai Chen, Keyu Chen, Xin Chen, Xun Chen, Zehui Chen, Zhi Chen, Pei Chu, Xiaoyi Dong, Haodong Duan, Qi Fan, Zhaoye Fei, Yang Gao, Jiaye Ge, Chenya Gu, Yuzhe Gu, Tao Gui, Aijia Guo, Qipeng Guo, Conghui He, Yingfan Hu, Ting Huang, Tao Jiang, Penglong Jiao, Zhenjiang Jin, Zhikai Lei, Jiaxing Li, Jingwen Li, Linyang Li, Shuaibin Li, Wei Li, Yining Li, Hongwei Liu, Jiangning Liu, Jiawei Hong, Kaiwen Liu, Kuikun Liu, Xiaoran Liu, Chengqi Lv, Haijun Lv, Kai Lv, Li Ma, Runyuan Ma, Zerun Ma, Wenchang Ning, Linke Ouyang, Jiantao Qiu, Yuan Qu, Fukai Shang, Yunfan Shao, Demin Song, Zifan Song, Zhihao Sui, Peng Sun, Yu Sun, Huanze Tang, Bin Wang, Guoteng Wang, Jiaqi Wang, Jiayu Wang, Rui Wang, Yudong Wang, Ziyi Wang, Xingjian Wei, Qizhen Weng, Fan Wu, Yingtong Xiong, Chao Xu, Ruiliang Xu, Hang Yan, Yirong Yan, Xiaogui Yang, Haochen Ye, Huaiyuan Ying, Jia Yu, Jing Yu, Yuhang Zang, Chuyu Zhang, Li Zhang, Pan Zhang, Peng Zhang, Ruijie Zhang, Shuo Zhang, Songyang Zhang, Wenjian Zhang, Wenwei Zhang, Xingcheng Zhang, Xinyue Zhang, Hui Zhao, Qian Zhao, Xiaomeng Zhao, Fengzhe Zhou, Zaida Zhou, Jingming Zhuo, Yicheng Zou, Xipeng Qiu, Yu Qiao, and Dahua Lin. 2024 · 2024
Closest in time.
Scaling Instruction-Finetuned Language Models
Hyung Won Chung, Le Hou, Shayne Longpre, Barret Zoph, Yi Tay, William Fedus, Yunxuan Li, Xuezhi Wang, Mostafa Dehghani, Siddhartha Brahma, Albert Webson, Shixiang Shane Gu, Zhuyun Dai, Mirac Suzgun, Xinyun Chen, Aakanksha Chowdhery, Alex Castro-Ros, Marie Pellat, Kevin Robinson, Dasha Valter, Sharan Narang, Gaurav Mishra, Adams Yu, Vincent Zhao, Yanping Huang, Andrew Dai, Hongkun Yu, Slav Petrov, Ed H. Chi, Jeff Dean, Jacob Devlin, Adam Roberts, Denny Zhou, Quoc V. Le, and Jason Wei. 2024 · 2024
Closest in time.
Deepseek llm: Scaling open-source language models with longtermism
DeepSeek-AI. 2024 · 2024
Closest in time.
Llamafactory: Unified efficient fine-tuning of 100+ language models
Yaowei Zheng, Richong Zhang, Junhao Zhang, Yanhan Ye, Zheyan Luo, and Yongqiang Ma. 2024 · 2024
Closest in time.