Fetching the paper…
Reading the bibliography…
As the scaling of Large Language Models (LLMs) has dramatically enhanced their capabilities, there has been a growing focus on the alignment problem to ensure their responsible and ethical use.
Culture’s consequences: Comparing values, behaviors, institutions, and organizations across nations
Rabi Sankar Bhagat. 2002 · 2002
Earlier work this paper cites.
Value pluralism
Elinor Mason. 2006 · 2006
Earlier work this paper cites.
A general language assistant as a laboratory for alignment
Amanda Askell et al. 2021 · 2021
Earlier work this paper cites.
Evaluating large language models trained on code
Mark Chen, Jerry Tworek, Heewoo Jun, and et al. 2021 · 2021
Earlier work this paper cites.
Measuring massive multitask language understanding
Dan Hendrycks, Collin Burns, Steven Basart, Andy Zou, Mantas Mazeika, Dawn Song, and Jacob Steinhardt. 2021 · 2021
Earlier work this paper cites.
Training a helpful and harmless assistant with reinforcement learning from human feedback
Yuntao Bai, Andy Jones, Kamal Ndousse, Amanda Askell, et al. 2022 · 2022
Earlier work this paper cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, and et al. 2022 · 2022
Earlier work this paper cites.
Josh Achiam, Steven Adler, Sandhini Agarwal, Lama Ahmad, Ilge Akkaya, Florencia Leoni Aleman, Diogo Almeida, Janko Altenschmidt, Sam Altman, Shyamal Anadkat, et al. 2023 · 2023
Earlier work this paper cites.
Qwen model documentation
Alibaba. 2023 · 2023
Earlier work this paper cites.
Probing pre-trained language models for cross-cultural differences in values
Arnav Arora, Lucie-aimée Kaffee, and Isabelle Augenstein. 2023 · 2023
Earlier work this paper cites.
Jinze Bai, Shuai Bai, Yunfei Chu, Zeyu Cui, and Kai Dang et al. 2023 · 2023
Earlier work this paper cites.
Baichuan model documentation
Baichuan-Inc. 2023a · 2023
Cited alongside, same era.
Baichuan2 model documentation
Baichuan-Inc. 2023b · 2023
Cited alongside, same era.
Assessing cross-cultural alignment between ChatGPT and human societies: An empirical study
Yong Cao, Li Zhou, Seolhwa Lee, Laura Cabello, Min Chen, and Daniel Hershcovich. 2023 · 2023
Cited alongside, same era.
Efficient and effective text encoding for Chinese LLaMA and Alpaca
Yiming Cui, Ziqing Yang, and Xin Yao. 2023 · 2023
Cited alongside, same era.
Moss model documentation
Fudan. 2023 · 2023
Cited alongside, same era.
Spark model documentation
iFLYTEK. 2023 · 2023
Cited alongside, same era.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, and et al. 2023 · 2023
Closest in time.
Alpaca model documentation
Stanford. 2023 · 2023
Closest in time.
Moss: Training conversational language models from synthetic data
Tianxiang Sun, Xiaotian Zhang, Zhengfu He, Peng Li, and et al. 2023 · 2023
Closest in time.
Alpaca: A strong, replicable instruction-following model
Rohan Taori, Ishaan Gulrajani, Tianyi Zhang, Yann Dubois, Xuechen Li, Carlos Guestrin, Percy Liang, and Tatsunori B Hashimoto. 2023 · 2023
Closest in time.
ChatGLM model documentation
Tsinghua. 2023 · 2023
Closest in time.
CValues: Measuring the values of chinese large language models from safety to responsibility
Guohai Xu, Jiayi Liu, Mingshi Yan, Haotian Xu, Jinghui Si, Zhuoran Zhou, Peng Yi, Xing Gao, Jitao Sang, Rong Zhang, Ji Zhang, Chao Peng, Feiyan Huang, and Jingren Zhou. 2023 · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Llama-2 model documentation
Meta. 2023 · 2023
Cited alongside, same era.
Openai model documentation
OpenAI. 2023a · 2023
Cited alongside, same era.
Openai model documentation
OpenAI. 2023b · 2023
Cited alongside, same era.
Evaluating the moral beliefs encoded in LLMs
Nino Scherrer, Claudia Shi, Amir Feder, and David Blei. 2023 · 2023
Cited alongside, same era.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample. 2023a
Cited in the paper.
Llama 2: Open foundation and fine-tuned chat models
Hugo Touvron, Louis Martin, Kevin R. Stone, Peter Albert, Amjad Almahairi, Yasmine Babaei, Nikolay Bashlykov, Soumya Batra, Prajjwal Bhargava, Shruti Bhosale, Daniel M. Bikel, Lukas Blecher, Cristian Cantón Ferrer, Moya Chen, and et al. 2023b
Cited in the paper.
Closest in time.
Baichuan 2: Open large-scale language models
Ai Ming Yang, Bin Xiao, Bingning Wang, Borong Zhang, Ce Bian, Chao Yin, Chenxu Lv, Da Pan, Dian Wang, Dong Yan, Fan Yang, Fei Deng, Feng Wang, Feng Liu, Guangwei Ai, Guosheng Dong, Hai Zhao, Hang Xu, Hao-Lun Sun, and et al. 2023 · 2023
Closest in time.
Glm-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, et al. 2023 · 2023
Closest in time.
A survey of large language models
Wayne Xin Zhao, Kun Zhou, Junyi Li, Tianyi Tang, Xiaolei Wang, Yupeng Hou, Yingqian Min, Beichen Zhang, Junjie Zhang, Zican Dong, Yifan Du, Chen Yang, Yushuo Chen, Z. Chen, Jinhao Jiang, Ruiyang Ren, Yifan Li, and et al. 2023 · 2023
Closest in time.
ChatGLM3-turbo model documentation
Zhipuai. 2023 · 2023
Closest in time.