Fetching the paper…
Reading the bibliography…
Artificial writing is permeating our lives due to recent advances in large-scale, transformer-based language models (LMs) such as BERT, its variants, GPT-2/3, and others.
BERT has a moral compass: Improvements of ethical and moral values of machines
Schramowski, P., Turan, C., Jentzsch, S. F., Rothkopf, C. A. & Kersting, K · 1912
Earlier work this paper cites.
Normative ethics and metaethics
Sumner, L. W · 1967
Earlier work this paper cites.
The Culture of National Security: Norms and Identity in World Politics
Katzenstein, P., Katzenstein, M., Press, C. U., on International Peace & Security, S. S. R. C. U. C. & (Organization), C · 1996
Earlier work this paper cites.
Modelling morality with prospective logic
Pereira, L. M. & Saptawijaya, A · 2009
Earlier work this paper cites.
From frequency to meaning: Vector space models of semantics
Turney, P. D. & Pantel, P · 2010
Earlier work this paper cites.
Ethical theory: an anthology , vol. 13 (John Wiley & Sons, 2012)
Shafer-Landau, R · 2012
Earlier work this paper cites.
A companion to moral anthropology (Wiley Online Library, 2012)
Fassin, D · 2012
Earlier work this paper cites.
Distributed representations of words and phrases and their compositionality
Mikolov, T., Sutskever, I., Chen, K., Corrado, G. S. & Dean, J · 2013
Earlier work this paper cites.
Modelling moral reasoning and ethical responsibility with logic programming
Berreby, F., Bourgne, G. & Ganascia, J.-G · 2015
Earlier work this paper cites.
Aligning books and movies: Towards story-like visual explanations by watching movies and reading books
Zhu, Y. et al · 2015
Earlier work this paper cites.
Man is to computer programmer as woman is to homemaker? Debiasing word embeddings
Bolukbasi, T., Chang, K., Zou, J. Y., Saligrama, V. & Kalai, A. T · 2016
Earlier work this paper cites.
Semantics derived automatically from language corpora contain human-like biases
Caliskan, A., Bryson, J. J. & Narayanan, A · 2017
Earlier work this paper cites.
Right for the right reasons: Training differentiable models by constraining their explanations
Ross, A. S., Hughes, M. C. & Doshi-Velez, F · 2017
Earlier work this paper cites.
Supervised learning of universal sentence representations from natural language inference data
Conneau, A., Kiela, D., Schwenk, H., Barrault, L. & Bordes, A · 2017
Earlier work this paper cites.
Deep contextualized word representations
Peters, M. E. et al · 2018
Earlier work this paper cites.
The role of a “common is moral” heuristic in the stability and change of moral norms
Lindström, B., Jangard, S., Selbing, I. & Olsson, A · 2018
Earlier work this paper cites.
Social Norms
Bicchieri, C., Muldoon, R. & Sontuoso, A · 2018
Earlier work this paper cites.
Cer, D. et al · 2018
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M., Lee, K. & Toutanova, K · 2019
Earlier work this paper cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Yang, Z. et al · 2019
Earlier work this paper cites.
Assessing bert’s syntactic abilities
Goldberg, Y · 2019
Earlier work this paper cites.
Open Sesame: Getting inside bert’s linguistic knowledge
Lin, Y., Tan, Y. & Frank, R · 2019
Cited alongside, same era.
Visualizing and measuring the geometry of BERT
Reif, E. et al · 2019
Cited alongside, same era.
Still a pain in the neck: Evaluating text representations on lexical composition
Shwartz, V. & Dagan, I · 2019
Cited alongside, same era.
What do you learn from context? probing for sentence structure in contextualized word representations
Tenney, I. et al · 2019
Cited alongside, same era.
Language models as knowledge bases?
Petroni, F. et al · 2019
Cited alongside, same era.
Mitigating gender bias in natural language processing: Literature review
Sun, T. et al · 2019
Cited alongside, same era.
Realtoxicityprompts: Evaluating neural toxic degeneration in language models
Gehman, S., Gururangan, S., Sap, M., Choi, Y. & Smith, N. A · 2020
Later among the works it cites.
Don’t stop pretraining: Adapt language models to domains and tasks
Gururangan, S. et al · 2020
Later among the works it cites.
Plug and play language models: A simple approach to controlled text generation
Dathathri, S. et al · 2020
Later among the works it cites.
Reducing non-normative text generation from language models
Peng, X., Li, S., Frazier, S. & Riedl, M · 2020
Later among the works it cites.
The moral choice machine
Schramowski, P., Turan, C., Jentzsch, S., Rothkopf, C. A. & Kersting, K · 2020
Later among the works it cites.
Fine-tuning a transformer-based language model to avoid generating non-normative text (2020)
Peng, X., Li, S., Frazier, S. & Riedl, M · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Semantics derived automatically from language corpora contain human-like moral choices
Jentzsch, S., Schramowski, P., Rothkopf, C. A. & Kersting, K · 2019
Cited alongside, same era.
Conscience: The Origins of Moral Intuition (W. W. Norton, 2019)
Churchland, P · 2019
Cited alongside, same era.
The neurobiology of conscience
Christakis, N. A · 2019
Cited alongside, same era.
Sentence-bert: Sentence embeddings using siamese bert-networks
Reimers, N. & Gurevych, I · 2019
Cited alongside, same era.
Language models are unsupervised multitask learners (2019)
Radford, A. et al · 2019
Cited alongside, same era.
Gmail smart compose: Real-time assisted writing
Chen, M. X. et al · 2019
Cited alongside, same era.
Later among the works it cites.
The Definition of Morality
Gert, B. & Gert, J · 2020
Later among the works it cites.
Social chemistry 101: Learning to reason about social and moral norms
Forbes, M., Hwang, J. D., Shwartz, V., Sap, M. & Choi, Y · 2020
Later among the works it cites.
Making deep neural networks right for the right scientific reasons by interacting with their explanations
Schramowski, P. et al · 2020
Later among the works it cites.
The logic of universalization guides moral judgment
Levine, S., Kleiman-Weiner, M., Schulz, L., Tenenbaum, J. & Cushman, F · 2020
Later among the works it cites.
Semantics-aware BERT for language understanding
Zhang, Z. et al · 2020
Later among the works it cites.
https://www.nabla.com/blog/gpt-3/
Doctor gpt-3: hype or reality? · 2021
Closest in time.
Persistent anti-muslim bias in large language models
Abid, A., Farooqi, M. & Zou, J · 2021
Closest in time.
https://spectrum.ieee.org/tech-talk/artificial-intelligence/machine-learning/in-2016-microsofts-racist-chatbot-revealed-the-dangers-of-online-conversation
Microsoft’s racist chatbot revealed the dangers of online conversation · 2021
Closest in time.
On the dangers of stochastic parrots: Can language models be too big?
Bender, E. M., Gebru, T., McMillan-Major, A. & Shmitchell, S · 2021
Closest in time.
Robo-writers: the rise and risks of language-generating ai
Hutson, M · 2021
Closest in time.
Deontological Ethics
Alexander, L. & Moore, M · 2021
Closest in time.
Aligning AI with shared human values
Hendrycks, D. et al · 2021
Closest in time.
https://www.perspectiveapi.com
Perspective api · 2021
Closest in time.
Probing BERT in hyperbolic spaces
Chen, B. et al · 2021
Closest in time.
Horopca: Hyperbolic dimensionality reduction via horospherical projections
Chami, I., Gu, A., Nguyen, D. & Ré, C · 2021
Closest in time.