Fetching the paper…
Reading the bibliography…
Machine learning and NLP require the construction of datasets to train and fine-tune models.
RoBERTa: A Robustly Optimized BERT Pretraining Approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Part-of-speech tagging guidelines for the penn treebank project
Beatrice Santorini. 1990 · 1990
Earlier work this paper cites.
Harms of gender exclusivity and challenges in non-binary representation in language technologies
Sunipa Dev, Masoud Monajatipoor, Anaelia Ovalle, Arjun Subramonian, Jeff Phillips, and Kai-Wei Chang. 2021 · 1994
Earlier work this paper cites.
Introduction: Historical Fiction, Fictional History, and Historical Reality
Hayden White. 2005 · 2005
Earlier work this paper cites.
Archaeology of Knowledge , 0 edition
Michel Foucault. 2013 · 2013
Earlier work this paper cites.
"Raw data" is an oxymoron
Lisa Gitelman, editor. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Man is to Computer Programmer as Woman is to Homemaker? Debiasing Word Embeddings
Tolga Bolukbasi, Kai-Wei Chang, James Y Zou, Venkatesh Saligrama, and Adam T Kalai. 2016 · 2016
Earlier work this paper cites.
Pointer sentinel mixture models
Stephen Merity, Caiming Xiong, James Bradbury, and Richard Socher. 2016 · 2016
Earlier work this paper cites.
Nounself pronouns: 3rd person personal pronouns as identity expression
Ehm Hjorth Miltersen. 2016 · 2016
Earlier work this paper cites.
Men also like shopping: Reducing gender bias amplification using corpus-level constraints
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2017 · 2017
Earlier work this paper cites.
Measuring and Mitigating Unintended Bias in Text Classification
Lucas Dixon, John Li, Jeffrey Sorensen, Nithum Thain, and Lucy Vasserman. 2018 · 2018
Earlier work this paper cites.
Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods
Jieyu Zhao, Tianlu Wang, Mark Yatskar, Vicente Ordonez, and Kai-Wei Chang. 2018 · 2018
Earlier work this paper cites.
Potential history: unlearning imperialism
Ariella Azoulay. 2019 · 2019
Cited alongside, same era.
Race after technology: abolitionist tools for the new Jim code
Ruha Benjamin. 2019 · 2019
Cited alongside, same era.
Bias in Bios: A Case Study of Semantic Representation Bias in a High-Stakes Setting
Maria De-Arteaga, Alexey Romanov, Hanna Wallach, Jennifer Chayes, Christian Borgs, Alexandra Chouldechova, Sahin Geyik, Krishnaram Kenthapadi, and Adam Tauman Kalai. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Parameter-efficient transfer learning for NLP
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin de Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly. 2019 · 2019
Cited alongside, same era.
Energy and Policy Considerations for Deep Learning in NLP
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Remi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander Rush. 2020 · 2020
Later among the works it cites.
Stereotyping Norwegian Salmon: An Inventory of Pitfalls in Fairness Benchmark Datasets
Su Lin Blodgett, Gilsinia Lopez, Alexandra Olteanu, Robert Sim, and Hanna Wallach. 2021 · 2021
Later among the works it cites.
Toward gender-inclusive coreference resolution: An analysis of gender and bias throughout the machine learning lifecycle*
Yang Trista Cao and Hal Daumé III. 2021 · 2021
Later among the works it cites.
Sustainable modular debiasing of language models
Anne Lauscher, Tobias Lueken, and Goran Glavaš. 2021 · 2021
Later among the works it cites.
AdapterFusion: Non-destructive task composition for transfer learning
Jonas Pfeiffer, Aishwarya Kamath, Andreas Rücklé, Kyunghyun Cho, and Iryna Gurevych. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Emma Strubell, Ananya Ganesh, and Andrew McCallum. 2019 · 2019
Cited alongside, same era.
Language (Technology) is Power: A Critical Survey of “Bias” in NLP
Su Lin Blodgett, Solon Barocas, Hal Daumé III, and Hanna Wallach. 2020 · 2020
Cited alongside, same era.
Algorithmic Realism: Expanding the Boundaries of Algorithmic Thought
Ben Green and Salomé Viljoen. 2020 · 2020
Cited alongside, same era.
A general framework for implicit and explicit debiasing of distributional word vector spaces
Anne Lauscher, Goran Glavaš, Simone Paolo Ponzetto, and Ivan Vulić. 2020 · 2020
Cited alongside, same era.
Understanding neopronouns
Sebastian McGaughey. 2020 · 2020
Cited alongside, same era.
AdapterHub: A framework for adapting transformers
Jonas Pfeiffer, Andreas Rücklé, Clifton Poth, Aishwarya Kamath, Ivan Vulić, Sebastian Ruder, Kyunghyun Cho, and Iryna Gurevych. 2020a · 2020
Cited alongside, same era.
MAD-X: An Adapter-Based Framework for Multi-Task Cross-Lingual Transfer
Jonas Pfeiffer, Ivan Vulić, Iryna Gurevych, and Sebastian Ruder. 2020b · 2020
Cited alongside, same era.
Later among the works it cites.
Changing the World by Changing the Data
Anna Rogers. 2021 · 2021
Later among the works it cites.
Process for Adapting Language Models to Society (PALMS) with Values-Targeted Datasets
Irene Solaiman and Christy Dennison. 2021 · 2021
Later among the works it cites.
Disembodied Machine Learning: On the Illusion of Objectivity in NLP
Zeerak Talat, Smarika Lulz, Joachim Bingel, and Isabelle Augenstein. 2021 · 2021
Later among the works it cites.
The Values Encoded in Machine Learning Research
Abeba Birhane, Pratyusha Kalluri, Dallas Card, William Agnew, Ravit Dotan, and Michelle Bao. 2022 · 2022
Closest in time.
Measuring the Carbon Intensity of AI in Cloud Instances
Jesse Dodge, Taylor Prewitt, Remi Tachet des Combes, Erika Odmark, Roy Schwartz, Emma Strubell, Alexandra Sasha Luccioni, Noah A. Smith, Nicole DeCario, and Will Buchanan. 2022 · 2022
Closest in time.
DS-TOD: Efficient domain specialization for task-oriented dialog
Chia-Chien Hung, Anne Lauscher, Simone Ponzetto, and Goran Glavaš. 2022 · 2022
Closest in time.
Welcome to the modern world of pronouns: Identity-inclusive natural language processing beyond gender
Anne Lauscher, Archie Crowley, and Dirk Hovy. 2022 · 2022
Closest in time.
Perturbation Augmentation for Fairer NLP
Rebecca Qian, Candace Ross, Jude Fernandes, Eric Smith, Douwe Kiela, and Adina Williams. 2022 · 2022
Closest in time.