Fetching the paper…
Reading the bibliography…
Discussion of AI alignment (alignment between humans and AI systems) has focused on value alignment, broadly referring to creating AI systems that share human values.
Two dogmas of empiricism
Willard V. O. Quine · 1951
Earlier work this paper cites.
Systems of reference and horizontal–vertical coordinates
J Piaget and B Inhelder · 1967
Earlier work this paper cites.
Ontological relativity
W.V.O. Quine · 1968
Earlier work this paper cites.
Natural categories
Eleanor H Rosch · 1973
Earlier work this paper cites.
The essential tension
Thomas S. Kuhn · 1978
Earlier work this paper cites.
Minds, brains, and programs
John R Searle · 1980
Earlier work this paper cites.
Rationality
Charles Taylor · 1982
Earlier work this paper cites.
Protocol Statements
Otto Neurath · 1983
Earlier work this paper cites.
The Structure of Empirical Knowledge
Laurence BonJour · 1985
Earlier work this paper cites.
The symbol grounding problem
Stevan Harnad · 1990
Earlier work this paper cites.
Conceptual pacts and lexical choice in conversation
Susan E Brennan and Herbert H Clark · 1996
Earlier work this paper cites.
Lexical entrainment in spontaneous dialog
Susan E Brennan et al · 1996
Earlier work this paper cites.
Syntactic priming in language production
Martin J Pickering and Holly P Branigan · 1999
Earlier work this paper cites.
The foundations of mind: Origins of conceptual thought
Jean Matter Mandler · 2004
Earlier work this paper cites.
Toward a mechanistic psychology of dialogue
Martin J Pickering and Simon Garrod · 2004
Earlier work this paper cites.
Representational similarity analysis-connecting the branches of systems neuroscience
Nikolaus Kriegeskorte, Marieke Mur, and Peter A Bandettini · 2008
Earlier work this paper cites.
Maximizing information exchange between complex networks
Bruce J West, Elvis L Geneston, and Paolo Grigolini · 2008
Earlier work this paper cites.
The Origin of Concepts
Susan Carey · 2009
Earlier work this paper cites.
Behavior matching in multimodal communication is synchronized
Max M Louwerse, Rick Dale, Ellen G Bard, and Patrick Jeuniaux · 2012
Earlier work this paper cites.
The self-organization of human interaction
Rick Dale, Riccardo Fusaroli, Nicholas D Duran, and Daniel C Richardson · 2013
Earlier work this paper cites.
Conversation, coupling and complexity: Matching scaling laws predict performance in a joint decision task
Riccardo Fusaroli, Drew H Abney, Bahador Bahrami, Christopher T Kello, and Kristian Tylén · 2013
Earlier work this paper cites.
How to improve on quinian bootstrapping-a response to nativist objections
Zoltan Jakab · 2013
Earlier work this paper cites.
Complexity matching in dyadic conversation
Drew H Abney, Alexandra Paxton, Rick Dale, and Christopher T Kello · 2014
Earlier work this paper cites.
Perspective-taking in dialogue as self-organization under social constraints
Nicholas D Duran and Rick Dale · 2014
Cited alongside, same era.
Dialog as interpersonal synergy
Riccardo Fusaroli, Joanna Rączaszek-Leonardi, and Kristian Tylén · 2014
Cited alongside, same era.
Deep inside convolutional networks: visualising image classification models and saliency maps
K Simonyan, A Vedaldi, and A Zisserman · 2014
Cited alongside, same era.
Movement dynamics reflect a functional role for weak coupling and role structure in dyadic problem solving
Drew H Abney, Alexandra Paxton, Rick Dale, and Christopher T Kello · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Cited alongside, same era.
Beyond the isolated brain: The promise and challenge of interacting minds
Thalia Wheatley, Adam Boncz, Ivan Toni, and Arjen Stolk · 2019
Later among the works it cites.
Fine-tuning language models from human preferences
Daniel M Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B Brown, Alec Radford, Dario Amodei, Paul Christiano, and Geoffrey Irving · 2019
Later among the works it cites.
A simple framework for contrastive learning of visual representations
Ting Chen, Simon Kornblith, Mohammad Norouzi, and Geoffrey Hinton · 2020
Later among the works it cites.
The Alignment Problem: Machine Learning and Human Values
B. Christian · 2020
Later among the works it cites.
Artificial intelligence, values, and alignment
Iason Gabriel · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ekaterina Vylomova, Laura Rimell, Trevor Cohn, and Timothy Baldwin · 2015
Cited alongside, same era.
Cooperative inverse reinforcement learning
Dylan Hadfield-Menell, Stuart J Russell, Pieter Abbeel, and Anca Dragan · 2016
Cited alongside, same era.
Adversarial machine learning at scale
Alexey Kurakin, Ian Goodfellow, and Samy Bengio · 2016
Cited alongside, same era.
How google’s ai viewed the move no human could understand
Cade Metz · 2016
Cited alongside, same era.
Social coordination of verbal and nonverbal behaviours
Alexandra Paxton, Rick Dale, and Daniel C Richardson · 2016
Cited alongside, same era.
" why should i trust you?" explaining the predictions of any classifier
Marco Tulio Ribeiro, Sameer Singh, and Carlos Guestrin · 2016
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Cited alongside, same era.
Dan Hendrycks, Collin Burns, Steven Basart, Andrew Critch, Jerry Li, Dawn Song, and Jacob Steinhardt · 2020
Later among the works it cites.
Parallelograms revisited: Exploring the limitations of vector space models for simple analogies. cognition, 205, article 104440, 2020
JC Peterson, D Chen, and TL Griffiths · 2020
Later among the works it cites.
Alignment in multimodal interaction: An integrative framework
Marlou Rasenberg, Asli Özyürek, and Mark Dingemanse · 2020
Later among the works it cites.
Value alignment verification
Daniel S Brown, Jordan Schneider, Anca Dragan, and Scott Niekum · 2021
Later among the works it cites.
Facebook apologizes after a.i. puts ‘primates’ label on video of black men
Ryan Mac · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Zero-shot text-to-image generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever · 2021
Later among the works it cites.
ChatGPT, OpenAI
https://openai.com/blog/chatgpt/ , 2022 · 2022
Later among the works it cites.
Shades of confusion: Lexical uncertainty modulates ad hoc coordination in an interactive communication task
Sonia K Murthy, Thomas L Griffiths, and Robert D Hawkins · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily L Denton, Kamyar Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, et al · 2022
Later among the works it cites.
Scaling autoregressive models for content-rich text-to-image generation
Jiahui Yu, Yuanzhong Xu, Jing Yu Koh, Thang Luong, Gunjan Baid, Zirui Wang, Vijay Vasudevan, Alexander Ku, Yinfei Yang, Burcu Karagol Ayan, et al · 2022
Later among the works it cites.
Palm-e: An embodied multimodal language model
Danny Driess, Fei Xia, Mehdi SM Sajjadi, Corey Lynch, Aakanksha Chowdhery, Brian Ichter, Ayzaan Wahid, Jonathan Tompson, Quan Vuong, Tianhe Yu, et al · 2023
Later among the works it cites.
From partners to populations: A hierarchical bayesian account of coordination and convention
Robert D Hawkins, Michael Franke, Michael C Frank, Adele E Goldberg, Kenny Smith, Thomas L Griffiths, and Noah D Goodman · 2023
Later among the works it cites.
Can we stop runaway a.i.?
Matthew Hutson · 2023
Later among the works it cites.
Hidden differences in phenomenal experience
Gary Lupyan, Ryutaro Uchiyama, Bill Thompson, and Daniel Casasanto · 2023
Later among the works it cites.
from risk to reward: the role of ai alignment in shaping a positive future
SingularityGroup · 2023
Later among the works it cites.
The clock and the pizza: Two stories in mechanistic explanation of neural networks
Ziqian Zhong, Ziming Liu, Max Tegmark, and Jacob Andreas · 2023
Later among the works it cites.