Fetching the paper…
Reading the bibliography…
For humanity to maintain and expand its agency into the future, the most powerful systems we create must be those which act to align the future with the will of humanity.
Fine-tuning language models from human preferences, 2019
Daniel M. Ziegler, Nisan Stiennon, Jeffrey Wu, Tom B. Brown, Alec Radford, Dario Amodei, Paul Christiano, and Geoffrey Irving · 1909
Earlier work this paper cites.
United Nations Charter
United Nations · 1945
Earlier work this paper cites.
The possibility of a social welfare function
James S Coleman · 1966
Earlier work this paper cites.
On taxation and the control of externalities
William J Baumol · 1972
Earlier work this paper cites.
Liberalism Against Populism: A Confrontation Between the Theory of Democracy and the Theory of Social Choice
W.H. Riker · 1982
Earlier work this paper cites.
A learning algorithm for boltzmann machines
David H Ackley, Geoffrey E Hinton, and Terrence J Sejnowski · 1985
Earlier work this paper cites.
Connectionist ai, symbolic ai, and the brain
Paul Smolensky · 1987
Earlier work this paper cites.
Toward a theory of the universal content and structure of values: Extensions and cross-cultural replications
Shalom H Schwartz and Wolfgang Bilsky · 1990
Earlier work this paper cites.
Conflict resolution applications to peace studies
Louis Kriesberg · 1991
Earlier work this paper cites.
Institutions
Douglass C North · 1991
Earlier work this paper cites.
Reinforcement learning: A survey, 1996
L. P. Kaelbling, M. L. Littman, and A. W. Moore · 1996
Earlier work this paper cites.
Institutions and their design
Robert E Goodin · 1996
Earlier work this paper cites.
Deliberative Democracy: Essays on Reason and Politics
James Bohman and William Rehg · 1997
Earlier work this paper cites.
Recommender systems
Paul Resnick and Hal R Varian · 1997
Earlier work this paper cites.
Collective will-formation: The missing dimension in public administration
Lance deHaven Smith · 1998
Earlier work this paper cites.
Survey article: The coming of age of deliberative democracy
James Bohman · 1998
Earlier work this paper cites.
The possibility of social choice
Amartya Sen · 1999
Earlier work this paper cites.
International peacebuilding: A theoretical and quantitative analysis
Michael W. Doyle and Nicholas Sambanis · 2000
Earlier work this paper cites.
Artificial intelligence, values and alignment
Iason Gabriel · 2001
Earlier work this paper cites.
Pca versus lda
Aleix M Martinez and Avinash C Kak · 2001
Earlier work this paper cites.
A functional approach to instrumental and terminal values and the value-attitude-behaviour system of consumer choice
Michael W Allen, Sik Hung Ng, and Marc Wilson · 2002
Earlier work this paper cites.
Social choice theory and deliberative democracy: a reconciliation
John S Dryzek and Christian List · 2003
Earlier work this paper cites.
Deliberative democratic theory
Simone Chambers · 2003
Earlier work this paper cites.
A gentle introduction to the universal algorithmic agent aixi, 2003
Marcus Hutter · 2003
Earlier work this paper cites.
Coherent Extrapolated Volition
Eliezer Yudkowsky · 2004
Earlier work this paper cites.
Why deliberative democracy?
Amy Gutmann and Dennis Thompson · 2004
Earlier work this paper cites.
What is artificial intelligence?
John McCarthy · 2004
Earlier work this paper cites.
A bayesian truth serum for subjective data
Drazen Prelec · 2004
Earlier work this paper cites.
Ethnic diversity and economic performance
Alberto Alesina and Eliana La Ferrara · 2005
Earlier work this paper cites.
Experimenting with a democratic ideal: Deliberative polling and public opinion
James S Fishkin and Robert C Luskin · 2005
Earlier work this paper cites.
What are institutions?
Geoffrey M Hodgson · 2006
Earlier work this paper cites.
Demand, innovation, and the dynamics of market structure: The role of experimental users and diverse preferences
Franco Malerba, Richard Nelson, Luigi Orsenigo, and Sidney Winter · 2007
Earlier work this paper cites.
Machine learning , volume 1
Tom Michael Mitchell et al · 2007
Earlier work this paper cites.
A survey of image classification methods and techniques for improving classification performance
Dengsheng Lu and Qihao Weng · 2007
Earlier work this paper cites.
Probabilistic matrix factorization
Andriy Mnih and Russ R Salakhutdinov · 2007
Earlier work this paper cites.
Visualizing data using t-sne
Laurens Van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
A large-scale hidden semi-markov model for anomaly detection on user browsing behaviors
Yi Xie and Shun-Zheng Yu · 2008
Earlier work this paper cites.
Dimensionality reduction: a comparative
Laurens Van Der Maaten, Eric Postma, Jaap Van den Herik, et al · 2009
Earlier work this paper cites.
Navigating the Power Dynamics between Institutions and Their Communities
Byron P White · 2009
Earlier work this paper cites.
The network of global corporate control
Stefania Vitali, James B Glattfelder, and Stefano Battiston · 2011
Earlier work this paper cites.
Visions of the Social: Society as a Political Project in France, 1750-1950
Jean Terrier · 2011
Earlier work this paper cites.
Small Area Estimation Using a Reweighting Algorithm
Robert Tanton, Yogi Vidyattama, Binod Nepal, and Justine McNamara · 2011
Earlier work this paper cites.
The walk in the woods: A step-by-step method for facilitating interest-based negotiation and conflict resolution
Leonard J Marcus, Barry C Dorn, and Eric J McNulty · 2012
Earlier work this paper cites.
Using the theory of satisficing to evaluate the quality of survey data
Scott Barge and Hunter Gehlbach · 2012
Earlier work this paper cites.
Social choice theory
Christian List · 2013
Earlier work this paper cites.
Creating truth-telling incentives with the bayesian truth serum
Ray Weaver and Drazen Prelec · 2013
Earlier work this paper cites.
Essai sur l’application de l’analyse à la probabilité des décisions rendues à la pluralité des voix
Nicolas De Condorcet · 2014
Earlier work this paper cites.
Sentiment analysis algorithms and applications: A survey
Walaa Medhat, Ahmed Hassan, and Hoda Korashy · 2014
Earlier work this paper cites.
Wiki surveys: Open and quantifiable social data collection
Matthew J. Salganik and Karen E. C. Levy · 2015
Earlier work this paper cites.
Probabilistic backpropagation for scalable learning of bayesian neural networks
José Miguel Hernández-Lobato and Ryan Adams · 2015
Earlier work this paper cites.
A novel neural topic model and its supervised extension
Ziqiang Cao, Sujian Li, Yang Liu, Wenjie Li, and Heng Ji · 2015
Earlier work this paper cites.
Hate speech detection with comment embeddings
Nemanja Djuric, Jing Zhou, Robin Morris, Mihajlo Grbovic, Vladan Radosavljevic, and Narayan Bhamidipati · 2015
Earlier work this paper cites.
The ai alignment problem: why it is hard, and where to start
Eliezer Yudkowsky · 2016
Earlier work this paper cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al · 2016
Earlier work this paper cites.
Practical secure aggregation for federated learning on user-held data
Keith Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone, H Brendan McMahan, Sarvar Patel, Daniel Ramage, Aaron Segal, and Karn Seth · 2016
Earlier work this paper cites.
Understanding institutions
Francesco Guala · 2016
Earlier work this paper cites.
Concrete problems in ai safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, and Dan Mané · 2016
Cited alongside, same era.
The political economy of heterogeneity and conflict
Enrico Spolaore and Romain Wacziarg · 2017
Cited alongside, same era.
Report of the joint committee on the eighth amendment of the constitution, 2017
Tithe an Oireachtais · 2017
Cited alongside, same era.
Deep reinforcement learning: An overview
Yuxi Li · 2017
Cited alongside, same era.
Implicit regularization in matrix factorization
Suriya Gunasekar, Blake E Woodworth, Srinadh Bhojanapalli, Behnam Neyshabur, and Nati Srebro · 2017
Cited alongside, same era.
Deep reinforcement learning from human preferences
Commercialization of fusion power plants
Hanni Lux, Dan Wolff, and Jack Foster · 2022
Later among the works it cites.
Anti-corporate sentiment in u.s. is now widespread in both parties, 2022
Amina Dunn and Andy Cerda · 2022
Later among the works it cites.
Elicitation inference optimization for multi-principal-agent alignment
Andrew Konya, Yeping Lina Qiu, Michael P Varga, and Aviv Ovadya · 2022
Later among the works it cites.
Fine-tuning language models to find agreement among humans with diverse preferences
Michiel A. Bakker, Martin J Chadwick, Hannah Sheahan, Michael Henry Tessler, Lucy Campbell-Gillingham, Jan Balaguer, Nat McAleese, Amelia Glaese, John Aslanides, Matthew Botvinick, and Christopher Summerfield · 2022
Later among the works it cites.
Democratizing behavioral economics
Zachary D. Liscow and Daniel Markovits · 2022
Later among the works it cites.
Negotiations in heterogeneous societies: Ratifying a peace agreement in israel
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Cited alongside, same era.
Review on the progress in nuclear fission—experimental methods and theoretical descriptions
Karl-Heinz Schmidt and Beatriz Jurado · 2018
Cited alongside, same era.
A deliberative theory of interest representation
Jane J Mansbridge · 2018
Cited alongside, same era.
The simple but ingenious system taiwan uses to crowdsource its laws
Chris Horton · 2018
Cited alongside, same era.
URL https://2016-2018.citizensassembly.ie/en/The-Eighth-Amendment-of-the-Constitution/Final-Report-on-the-Eighth-Amendment-of-the-Constitution/Final-Report-incl-Appendix-A-D.pdf
First report and recommendations of the citizens’ assembly; the eight ammendment of the constitution, 2017 · 2018
Cited alongside, same era.
Irish abortion referendum: Ireland overturns abortion ban, 2018
BBC · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Cited alongside, same era.
Ofer Zalzberg and Roie Ravitzky · 2022
Later among the works it cites.
Representation with incomplete votes, 2022
Daniel Halpern, Ariel D. Procaccia, Gregory Kehne, Jamie Tucker-Foltz, and Manuel Wültrich · 2022
Later among the works it cites.
Scott Reed, Konrad Zolna, Emilio Parisotto, Sergio Gomez Colmenarejo, Alexander Novikov, Gabriel Barth-Maron, Mai Gimenez, Yury Sulsky, Jackie Kay, Jost Tobias Springenberg, et al · 2022
Later among the works it cites.
Rethinking the role of demonstrations: What makes in-context learning work?
Sewon Min, Xinxi Lyu, Ari Holtzman, Mikel Artetxe, Mike Lewis, Hannaneh Hajishirzi, and Luke Zettlemoyer · 2022
Later among the works it cites.
Google engineer claims ai chatbot is sentient: Why that matters
Leonardo De Cosmo · 2022
Later among the works it cites.
Topic modeling by clustering language model embeddings: Human validation on an industry dataset
Anton Eklund and Mona Forsman · 2022
Later among the works it cites.
Bertopic: Neural topic modeling with a class-based tf-idf procedure, 2022
Maarten Grootendorst · 2022
Later among the works it cites.
Towards privacy-preserving and verifiable federated matrix factorization
Xicheng Wan, Yifeng Zheng, Qun Li, Anmin Fu, Mang Su, and Yansong Gao · 2022
Later among the works it cites.
Robust aggregation for federated learning
Krishna Pillutla, Sham M Kakade, and Zaid Harchaoui · 2022
Later among the works it cites.
Deep learning, reinforcement learning, and world models
Yutaka Matsuo, Yann LeCun, Maneesh Sahani, Doina Precup, David Silver, Masashi Sugiyama, Eiji Uchibe, and Jun Morimoto · 2022
Later among the works it cites.
A path towards autonomous machine intelligence version 0.9. 2, 2022-06-27
Yann LeCun · 2022
Later among the works it cites.
Measuring progress on scalable oversight for large language models
Samuel R Bowman, Jeeyoon Hyun, Ethan Perez, Edwin Chen, Craig Pettit, Scott Heiner, Kamile Lukosuite, Amanda Askell, Andy Jones, Anna Chen, et al · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
Learning reward functions from diverse sources of human feedback: Optimally integrating demonstrations and preferences
Erdem Bıyık, Dylan P Losey, Malayandi Palan, Nicholas C Landolfi, Gleb Shevchuk, and Dorsa Sadigh · 2022
Later among the works it cites.
Jury learning: Integrating dissenting voices into machine learning models
Mitchell L. Gordon, Michelle S. Lam, Joon Sung Park, Kayur Patel, Jeffrey T. Hancock, Tatsunori Hashimoto, and Michael S. Bernstein · 2022
Later among the works it cites.
Out of one, many: Using language models to simulate human samples, 2022
Lisa P. Argyle, Ethan C. Busby, Nancy Fulda, Joshua Gubler, Christopher Rytting, and David Wingate · 2022
Later among the works it cites.
Social simulacra: Creating populated prototypes for social computing systems, 2022
Joon Sung Park, Lindsay Popowski, Carrie J. Cai, Meredith Ringel Morris, Percy Liang, and Michael S. Bernstein · 2022
Later among the works it cites.
Controllable face editing for video reconstruction in human digital twins
Chengde Lin and Shengwu Xiong · 2022
Later among the works it cites.
Congress and the public, 2023
Gallop · 2023
Closest in time.
Uk government approval rating, 2023
Statista · 2023
Closest in time.
Natural selection favors ais over humans
Dan Hendrycks · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with gpt-4
Sébastien Bubeck, Varun Chandrasekaran, Ronen Eldan, Johannes Gehrke, Eric Horvitz, Ece Kamar, Peter Lee, Yin Tat Lee, Yuanzhi Li, Scott Lundberg, Harsha Nori, Hamid Palangi, Marco Tulio Ribeiro, and Yi Zhang · 2023
Closest in time.
Democratic policy development using collective dialogues and ai, 2023
Andrew Konya, Lisa Schirch, Colin Irwin, and Aviv Ovadya · 2023
Closest in time.
Collective constitutional ai: Aligning a language model with public input, 2023
Deep Ganguli, Saffron Huang, Liane Lovitt, Divya Siddarth, Thomas Liao, and Esin Durmus · 2023
Closest in time.
A research agenda for the production of a flourishing civilization: Ai objectives institute whitepaper, 2023a
Deger Turan, Peter Eckersley, Max Shron, Tushant Jha, Divya Siddarth, Brittney Gallagher, Carroll Wainwright, Joel Lehman, and Brian Christian · 2023
Closest in time.
The open agency model, 2023
Eric Drexler · 2023
Closest in time.
Rohan Anil, Andrew M Dai, Orhan Firat, Melvin Johnson, Dmitry Lepikhin, Alexandre Passos, Siamak Shakeri, Emanuel Taropa, Paige Bailey, Zhifeng Chen, et al · 2023
Closest in time.
Auto-gpt
Significant-Gravitas · 2023
Closest in time.
Opportunities and risks of llms for scalable deliberation with polis, 2023
Christopher T. Small, Ivan Vendrov, Esin Durmus, Hadjar Homaei, Elizabeth Barry, Julien Cornebise, Ted Suzman, Deep Ganguli, and Colin Megill · 2023
Closest in time.
Benefits, limits, and risks of gpt-4 as an ai chatbot for medicine
Peter Lee, Sebastien Bubeck, and Joseph Petro · 2023
Closest in time.
How good are gpt models at machine translation? a comprehensive evaluation
Amr Hendy, Mohamed Abdelrehim, Amr Sharaf, Vikas Raunak, Mohamed Gabr, Hitokazu Matsushita, Young Jin Kim, Mohamed Afify, and Hany Hassan Awadalla · 2023
Closest in time.
Is chatgpt a good translator? yes with gpt-4 as the engine
Wenxiang Jiao, WX Wang, JT Huang, Xing Wang, and ZP Tu · 2023
Closest in time.
Preference transformer: Modeling human preferences using transformers for rl, 2023
Changyeon Kim, Jongjin Park, Jinwoo Shin, Honglak Lee, Pieter Abbeel, and Kimin Lee · 2023
Closest in time.
Reinforcement learning for topic models, 2023
Jeremy Costello and Marek Z. Reformat · 2023
Closest in time.
Large-scale text analysis using generative language models: A case study in discovering public value expressions in ai patents, 2023
Sergio Pelaez, Gaurav Verma, Barbara Ribeiro, and Philip Shapira · 2023
Closest in time.
Generative social choice, 2023
Sara Fish, Paul Gölz, David C. Parkes, Ariel D. Procaccia, Gili Rusak, Itai Shapira, and Manuel Wüthrich · 2023
Closest in time.
Ai-assisted coding: Experiments with gpt-4, 2023
Russell A Poldrack, Thomas Lu, and Gašper Beguš · 2023
Closest in time.
Is gpt-4 a good data analyst?, 2023
Liying Cheng, Xingxuan Li, and Lidong Bing · 2023
Closest in time.
Beyond generating code: Evaluating gpt on a data visualization course, 2023
Zhutian Chen, Chenyang Zhang, Qianwen Wang, Jakob Troidl, Simon Warchol, Johanna Beyer, Nils Gehlenborg, and Hanspeter Pfister · 2023
Closest in time.
Can ai moderate online communities?
Henrik Axelsen, Johannes Rude Jensen, Sebastian Axelsen, Valdemar Licht, and Omri Ross · 2023
Closest in time.
Sandra Mitrović, Davide Andreoletti, and Omran Ayoub · 2023
Closest in time.
Proof of humanity, 2023
Kleros · 2023
Closest in time.
A verifiable and privacy-preserving framework for federated recommendation system
Fei Gao, Hanlin Zhang, Jie Lin, Hansong Xu, Fanyu Kong, and Guoqiang Yang · 2023
Closest in time.
Validating the integrity of convolutional neural network predictions based on zero-knowledge proof
Yongkai Fan, Binyuan Xu, Linlin Zhang, Jinbao Song, Albert Zomaya, and Kuan-Ching Li · 2023
Closest in time.
From word models to world models: Translating from natural language to the probabilistic language of thought, 2023
Lionel Wong, Gabriel Grand, Alexander K. Lew, Noah D. Goodman, Vikash K. Mansinghka, Jacob Andreas, and Joshua B. Tenenbaum · 2023
Closest in time.
Evaluating the logical reasoning ability of chatgpt and gpt-4
Hanmeng Liu, Ruoxi Ning, Zhiyang Teng, Jian Liu, Qiji Zhou, and Yue Zhang · 2023
Closest in time.
Instruction tuning with gpt-4, 2023
Baolin Peng, Chunyuan Li, Pengcheng He, Michel Galley, and Jianfeng Gao · 2023
Closest in time.
Principled reinforcement learning with human feedback from pairwise or k k -wise comparisons
Banghua Zhu, Jiantao Jiao, and Michael I Jordan · 2023
Closest in time.
Chatgpt sets record for fastest-growing user base, 2023
Krystal Hu · 2023
Closest in time.
Ego, fear and money: How the a.i. fuse was lit
Cade Metz, Karen Weise, Nico Grant, and Mike Isaac · 2023
Closest in time.
Societal divides as a taxable negative externality of digital platforms, 2023
Helena Puig · 2023
Closest in time.