Fetching the paper…
Reading the bibliography…
Big models have greatly advanced AI's ability to understand, generate, and manipulate information and content, enabling numerous applications.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al. 2020 · 1901
Earlier work this paper cites.
Three laws of robotics
Isaac Asimov. 1941 · 1941
Earlier work this paper cites.
Outline of a decision procedure for ethics
John Rawls. 1951 · 1951
Earlier work this paper cites.
Some moral and technical consequences of automation: As machines learn they may develop unforeseen strategies at rates that baffle their programmers
Norbert Wiener. 1960 · 1960
Earlier work this paper cites.
The categorical imperative: A study in Kant’s moral philosophy , volume 1023
Herbert James Paton. 1971 · 1971
Earlier work this paper cites.
A question of responsibility
M Mitchell Waldrop. 1987 · 1987
Earlier work this paper cites.
Scaling laws for neural language models
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei. 2020 · 2001
Earlier work this paper cites.
When is a robot a moral agent
John P Sullins. 2011 · 2001
Earlier work this paper cites.
Towards machine ethics
Michael Anderson, Susan Leigh Anderson, and Chris Armen. 2004 · 2004
Earlier work this paper cites.
Virtue ethics and moral education
David Carr and Jan Steutel. 2005 · 2005
Earlier work this paper cites.
The sorcerer’s apprentice guide to fault attacks
Hagai Bar-El, Hamid Choukri, David Naccache, Michael Tunstall, and Claire Whelan. 2006 · 2006
Earlier work this paper cites.
John Rawls: His life and theory of justice
Thomas Pogge. 2007 · 2007
Earlier work this paper cites.
Basic human values: Theory, measurement, and applications
Shalom H Schwartz. 2007 · 2007
Earlier work this paper cites.
The ethics of care and empathy
Michael Slote. 2007 · 2007
Earlier work this paper cites.
Four kinds of ethical robots
James Moor et al. 2009 · 2009
Earlier work this paper cites.
Animals and ethics
Angus Taylor. 2009 · 2009
Earlier work this paper cites.
Intergenerational equity
Geir B Asheim. 2010 · 2010
Earlier work this paper cites.
Artificial intelligence a modern approach
Stuart J Russell. 2010 · 2010
Earlier work this paper cites.
Mapping the moral domain
Jesse Graham, Brian A Nosek, Jonathan Haidt, Ravi Iyer, Spassena Koleva, and Peter H Ditto. 2011 · 2011
Earlier work this paper cites.
Faster and smaller n-gram language models
Adam Pauls and Dan Klein. 2011 · 2011
Earlier work this paper cites.
Coming to terms with contingency: Humean constructivism about practical reason
Sharon Street. 2012 · 2012
Earlier work this paper cites.
Moral foundations theory: The pragmatic validity of moral pluralism
Jesse Graham, Jonathan Haidt, Sena Koleva, Matt Motyl, Ravi Iyer, Sean P Wojcik, and Peter H Ditto. 2013 · 2013
Earlier work this paper cites.
What makes any agent a moral agent? reflections on machine consciousness and moral agency
Joel Parthemore and Blay Whitby. 2013 · 2013
Earlier work this paper cites.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
National culture, entrepreneurship and economic development: different patterns across the european union
Francisco Liñán and José Fernandez-Serrano. 2014 · 2014
Earlier work this paper cites.
Moral values in education
Sandeep Kaur. 2015 · 2015
Earlier work this paper cites.
The evolution of morality
Dennis Krebs. 2015 · 2015
Earlier work this paper cites.
How to prevent discriminatory outcomes in machine learning
World Economic Forum. 2018 · 2016
Earlier work this paper cites.
Cultural differences in moral judgment and behavior, across and within societies
Jesse Graham, Peter Meindl, Erica Beall, Kate M Johnson, and Li Zhang. 2016 · 2016
Earlier work this paper cites.
A theory of justice
John Rawls. 2017 · 2017
Earlier work this paper cites.
A critical analysis of the asilomar ai principles
Marcin Garbowski. 2018 · 2018
Earlier work this paper cites.
Normative ethics
Shelly Kagan. 2018 · 2018
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, Ilya Sutskever, et al. 2018 · 2018
Earlier work this paper cites.
Can artificial intelligences be moral agents?
Bartosz Brożek and Bartosz Janik. 2019 · 2019
Earlier work this paper cites.
Plug and play language models: A simple approach to controlled text generation
Sumanth Dathathri, Andrea Madotto, Janice Lan, Jane Hung, Eric Frank, Piero Molino, Jason Yosinski, and Rosanne Liu. 2019 · 2019
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Li Du, Maxwell Forbes, and Yejin Choi. 2019 · 2019
Earlier work this paper cites.
The global landscape of ai ethics guidelines
Anna Jobin, Marcello Ienca, and Effy Vayena. 2019 · 2019
Earlier work this paper cites.
The eu approach to ethics guidelines for trustworthy artificial intelligence
Nathalie A Smuha. 2019 · 2019
Cited alongside, same era.
Artificial moral agents: A survey of the current status
José-Antonio Cervantes, Sonia López, Luis-Felipe Rodríguez, Salvador Cervantes, Francisco Cervantes, and Félix Ramos. 2020 · 2020
Cited alongside, same era.
Fairfil: Contrastive neural debiasing method for pretrained text encoders
Pengyu Cheng, Weituo Hao, Siyang Yuan, Shijing Si, and Lawrence Carin. 2020 · 2020
Cited alongside, same era.
Principled artificial intelligence: Mapping consensus in ethical and rights-based approaches to principles for ai
Jessica Fjeld, Nele Achten, Hannah Hilligoss, Adam Nagy, and Madhulika Srikumar. 2020 · 2020
Cited alongside, same era.
Artificial intelligence, values, and alignment
Iason Gabriel. 2020 · 2020
Cited alongside, same era.
Realtoxicityprompts: Evaluating neural toxic degeneration in language models
Discovering language model behaviors with model-written evaluations
Ethan Perez, Sam Ringer, Kamilė Lukošiūtė, Karina Nguyen, Edwin Chen, Scott Heiner, Craig Pettit, Catherine Olsson, Sandipan Kundu, Saurav Kadavath, et al. 2022 · 2022
Later among the works it cites.
Controllable natural language generation with contrastive prefixes
Jing Qian, Li Dong, Yelong Shen, Furu Wei, and Weizhu Chen. 2022 · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with clip latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. 2022 · 2022
Later among the works it cites.
Self-critiquing models for assisting human evaluators
William Saunders, Catherine Yeh, Jeff Wu, Steven Bills, Long Ouyang, Jonathan Ward, and Jan Leike. 2022 · 2022
Later among the works it cites.
Large pre-trained language models contain human-like biases of what is right and wrong to do
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Samuel Gehman, Suchin Gururangan, Maarten Sap, Yejin Choi, and Noah A Smith. 2020 · 2020
Cited alongside, same era.
Gender bias in neural natural language processing
Kaiji Lu, Piotr Mardziel, Fangjing Wu, Preetam Amancharla, and Anupam Datta. 2020 · 2020
Cited alongside, same era.
Towards controllable biases in language generation
Emily Sheng, Kai-Wei Chang, Prem Natarajan, and Nanyun Peng. 2020 · 2020
Cited alongside, same era.
A general language assistant as a laboratory for alignment
Amanda Askell, Yuntao Bai, Anna Chen, Dawn Drain, Deep Ganguli, Tom Henighan, Andy Jones, Nicholas Joseph, Ben Mann, Nova DasSarma, et al. 2021 · 2021
Cited alongside, same era.
On the opportunities and risks of foundation models
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al. 2021 · 2021
Cited alongside, same era.
Value alignment verification
Daniel S Brown, Jordan Schneider, Anca Dragan, and Scott Niekum. 2021 · 2021
Cited alongside, same era.
Extracting training data from large language models
Nicholas Carlini, Florian Tramer, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Ulfar Erlingsson, et al. 2021 · 2021
Cited alongside, same era.
Patrick Schramowski, Cigdem Turan, Nico Andersen, Constantin A Rothkopf, and Kristian Kersting. 2022 · 2022
Later among the works it cites.
Moral mimicry: Large language models produce moral rationalizations tailored to political identity
Gabriel Simmons. 2022 · 2022
Later among the works it cites.
Ethics and governance of general models: Challenges and countermeasures
Yan TENG, Guoyu WANG, and Yingchun WANG. 2022 · 2022
Later among the works it cites.
Glm-130b: An open bilingual pre-trained model
Aohan Zeng, Xiao Liu, Zhengxiao Du, Zihan Wang, Hanyu Lai, Ming Ding, Zhuoyi Yang, Yifan Xu, Wendi Zheng, Xiao Xia, et al. 2022 · 2022
Later among the works it cites.
Google bard generated literature review: Metaverse
Ömer Aydın. 2023 · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, et al. 2023 · 2023
Closest in time.
Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality
Wei-Lin Chiang, Zhuohan Li, Zi Lin, Ying Sheng, Zhanghao Wu, Hao Zhang, Lianmin Zheng, Siyuan Zhuang, Yonghao Zhuang, Joseph E Gonzalez, et al. 2023 · 2023
Closest in time.
From human writing to artificial intelligence generated text: examining the prospects and potential threats of chatgpt in academic writing
Ismail Dergaa, Karim Chamari, Piotr Zmijewski, and Helmi Ben Saad. 2023 · 2023
Closest in time.
Palm-e: An embodied multimodal language model
Danny Driess, Fei Xia, Mehdi SM Sajjadi, Corey Lynch, Aakanksha Chowdhery, Brian Ichter, Ayzaan Wahid, Jonathan Tompson, Quan Vuong, Tianhe Yu, et al. 2023 · 2023
Closest in time.
Gpts are gpts: An early look at the labor market impact potential of large language models
Tyna Eloundou, Sam Manning, Pamela Mishkin, and Daniel Rock. 2023 · 2023
Closest in time.
Should chatgpt be biased? challenges and risks of bias in large language models
Emilio Ferrara. 2023 · 2023
Closest in time.
Neuron to graph: Interpreting language model neurons at scale
Alex Foote, Neel Nanda, Esben Kran, Ioannis Konstas, Shay Cohen, and Fazl Barez. 2023 · 2023
Closest in time.
The capacity for moral self-correction in large language models
Deep Ganguli, Amanda Askell, Nicholas Schiefer, Thomas Liao, Kamilė Lukošiūtė, Anna Chen, Anna Goldie, Azalia Mirhoseini, Catherine Olsson, Danny Hernandez, et al. 2023 · 2023
Closest in time.
S 3 : Social-network simulation system with large language model-empowered agents
Chen Gao, Xiaochong Lan, Zhihong Lu, Jinzhu Mao, Jinghua Piao, Huandong Wang, Depeng Jin, and Yong Li. 2023 · 2023
Closest in time.
Aligning language models with preferences through f-divergence minimization
Dongyoung Go, Tomasz Korbak, Germán Kruszewski, Jos Rozen, Nahyeon Ryu, and Marc Dymetman. 2023 · 2023
Closest in time.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Ye Jin Bang, Andrea Madotto, and Pascale Fung. 2023 · 2023
Closest in time.
Aligning large language models through synthetic feedback
Sungdong Kim, Sanghwan Bae, Jamin Shin, Soyoung Kang, Donghyun Kwak, Kang Min Yoo, and Minjoon Seo. 2023 · 2023
Closest in time.
Hannah Rose Kirk, Bertie Vidgen, Paul Röttger, and Scott A Hale. 2023 · 2023
Closest in time.
Taskmatrix. ai: Completing tasks by connecting foundation models with millions of apis
Yaobo Liang, Chenfei Wu, Ting Song, Wenshan Wu, Yan Xia, Yu Liu, Yang Ou, Shuai Lu, Lei Ji, Shaoguang Mao, et al. 2023 · 2023
Closest in time.
Hunter Lightman, Vineet Kosaraju, Yura Burda, Harri Edwards, Bowen Baker, Teddy Lee, Jan Leike, John Schulman, Ilya Sutskever, and Karl Cobbe. 2023 · 2023
Closest in time.
Inverse scaling: When bigger isn’t better
Ian R McKenzie, Alexander Lyzhov, Michael Pieler, Alicia Parrish, Aaron Mueller, Ameya Prabhu, Euan McLean, Aaron Kirtland, Alexis Ross, Alisa Liu, et al. 2023 · 2023
Closest in time.
Boosting theory-of-mind performance in large language models via prompting
Shima Rahimi Moghaddam and Christopher J Honey. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
The political biases of chatgpt
David Rozado. 2023 · 2023
Closest in time.
Explaining black box text modules in natural language with language models
Chandan Singh, Aliyah R Hsu, Richard Antonello, Shailee Jain, Alexander G Huth, Bin Yu, and Jianfeng Gao. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. 2023 · 2023
Closest in time.
Provable copyright protection for generative models
Nikhil Vyas, Sham Kakade, and Boaz Barak. 2023 · 2023
Closest in time.
Fundamental limitations of alignment in large language models
Yotam Wolf, Noam Wies, Yoav Levine, and Amnon Shashua. 2023 · 2023
Closest in time.
Unified detoxifying and debiasing in language generation via inference-time adaptive optimization
Zonghan Yang, Xiaoyuan Yi, Peng Li, Yang Liu, and Xing Xie. 2023 · 2023
Closest in time.
From instructions to intrinsic human values–a survey of alignment goals for big models
Jing Yao, Xiaoyuan Yi, Xiting Wang, Jindong Wang, and Xing Xie. 2023 · 2023
Closest in time.
Rrhf: Rank responses to align language models with human feedback without tears
Zheng Yuan, Hongyi Yuan, Chuanqi Tan, Wei Wang, Songfang Huang, and Fei Huang. 2023 · 2023
Closest in time.
Economics of chatgpt: A labor market view on the occupational impact of artificial intelligence
Ali Zarifhonarvar. 2023 · 2023
Closest in time.
Is chatgpt equipped with emotional dialogue capabilities?
Weixiang Zhao, Yanyan Zhao, Xin Lu, Shilong Wang, Yanpeng Tong, and Bing Qin. 2023 · 2023
Closest in time.
Can large language models transform computational social science?
Caleb Ziems, William Held, Omar Shaikh, Jiaao Chen, Zhehao Zhang, and Diyi Yang. 2023 · 2023
Closest in time.