Fetching the paper…
Reading the bibliography…
We are introducing Aligned, a platform for global governance and alignment of frontier models, and eventually superintelligence.
Toward trustworthy ai development: Mechanisms for supporting verifiable claims
Miles Brundage, Shahar Avin, Jasmine Wang, Haydn Belfield, Gretchen Krueger, et al · 2004
Earlier work this paper cites.
Stefan Wojcik, Sophie Hilgard, Nick Judd, Delia Mocanu, Stephen Ragain, M. B. Fallin Hunzaker, Keith Coleman, and Jay Baxter · 2022
Earlier work this paper cites.
Fine-tuning language models to find agreement among humans with diverse preferences
Michiel A. Bakker, Martin J. Chadwick, Hannah R. Sheahan, Michael Henry Tessler, Lucy Campbell-Gillingham, Jan Balaguer, Nat McAleese, Amelia Glaese, John Aslanides, Matthew M. Botvinick, and Christopher Summerfield · 2022
Earlier work this paper cites.
Democratic inputs to ai, 2023
Wojciech Zaremba, Arka Dhar, Lama Ahmad, Tyna Eloundou, Shibani Santurkar, Sandhini Agarwal, and Jade Leung · 2023
Cited alongside, same era.
Governance of superintelligence, 2023
Sam Altman, Greg Brockman, and Ilya Sutskever · 2023
Cited alongside, same era.
Collective constitutional ai: Aligning a language model with public input, 2023
Anthropic · 2023
Cited alongside, same era.
How are community notes ranked?
X · 2023
Closest in time.
Using gpt-4 for content moderation
Lilian Weng, Vik Goel, and Andrea Vallone · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…