Fetching the paper…
Reading the bibliography…
Pre-trained language models (PLMs) have outperformed other NLP models on a wide range of tasks.
Roberta: A robustly optimized bert pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
Why gender and age prediction from tweets is hard: Lessons from a crowdsourcing experiment
Dong Nguyen, Dolf Trieschnigg, A. Seza Doğruöz, Rilana Gravel, Mariët Theune, Theo Meder, and Franciska de Jong. 2014 · 1961
Earlier work this paper cites.
The imperative of responsibility: In search of an ethics for the technological age
Hans Jonas. 1984 · 1984
Earlier work this paper cites.
Sociolinguistics: An introduction to language and society
Peter Trudgill. 2000 · 2000
Earlier work this paper cites.
Measuring and reducing gendered correlations in pre-trained models
Kellie Webster, Xuezhi Wang, Ian Tenney, Alex Beutel, Emily Pitler, Ellie Pavlick, Jilin Chen, Ed Chi, and Slav Petrov. 2020 · 2010
Earlier work this paper cites.
Discriminating gender on Twitter
John D. Burger, John Henderson, George Kim, and Guido Zarrella. 2011 · 2011
Earlier work this paper cites.
Age prediction in blogs: A study of style, content, and online behavior in pre- and post-social media generations
Sara Rosenthal and Kathleen McKeown. 2011 · 2011
Earlier work this paper cites.
Language and gender
Penelope Eckert and Sally McConnell-Ginet. 2013 · 2013
Earlier work this paper cites.
Exploring demographic language variations to improve multilingual sentiment analysis in social media
Svitlana Volkova, Theresa Wilson, and David Yarowsky. 2013 · 2013
Earlier work this paper cites.
Demographic factors improve classification performance
Dirk Hovy. 2015 · 2015
Earlier work this paper cites.
Cross-lingual syntactic variation over age and gender
Anders Johannsen, Dirk Hovy, and Anders Søgaard. 2015 · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Yukun Zhu, Ryan Kiros, Rich Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler. 2015 · 2015
Earlier work this paper cites.
Fine-grained analysis of sentence embeddings using auxiliary prediction tasks
Yossi Adi, Einat Kermany, Yonatan Belinkov, Ofer Lavi, and Yoav Goldberg. 2016 · 2016
Earlier work this paper cites.
Demographic dialectal variation in social media: A case study of African-American English
Su Lin Blodgett, Lisa Green, and Brendan O’Connor. 2016 · 2016
Earlier work this paper cites.
Probing for semantic evidence of composition by means of simple classification tasks
Allyson Ettinger, Ahmed Elgohary, and Philip Resnik. 2016 · 2016
Earlier work this paper cites.
Multitask learning for mental health conditions with limited social media data
Adrian Benton, Margaret Mitchell, and Dirk Hovy. 2017 · 2017
Earlier work this paper cites.
Language-independent gender prediction on Twitter
Nikola Ljubešić, Darja Fišer, and Tomaž Erjavec. 2017 · 2017
Earlier work this paper cites.
Human centered NLP with user-factor adaptation
Veronica Lynn, Youngseo Son, Vivek Kulkarni, Niranjan Balasubramanian, and H. Andrew Schwartz. 2017 · 2017
Earlier work this paper cites.
Overcoming language variation in sentiment analysis with social attention
Yi Yang and Jacob Eisenstein. 2017 · 2017
Earlier work this paper cites.
What you can cram into a single $&!#* vector: Probing sentence embeddings for linguistic properties
Alexis Conneau, German Kruszewski, Guillaume Lample, Loïc Barrault, and Marco Baroni. 2018 · 2018
Earlier work this paper cites.
Towards robust and privacy-preserving text representations
Yitong Li, Timothy Baldwin, and Trevor Cohn. 2018 · 2018
Earlier work this paper cites.
Reusable workflows for gender prediction
Matej Martinc and Senja Pollak. 2018 · 2018
Earlier work this paper cites.
Sentence encoders on stilts: Supplementary training on intermediate labeled-data tasks
Jason Phang, Thibault Févry, and Samuel R Bowman. 2018 · 2018
Earlier work this paper cites.
RtGender: A corpus for studying differential responses to gender
Rob Voigt, David Jurgens, Vinodkumar Prabhakaran, Dan Jurafsky, and Yulia Tsvetkov. 2018 · 2018
Earlier work this paper cites.
GLUE: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel Bowman. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Women’s syntactic resilience and men’s grammatical luck: Gender-bias in part-of-speech tagging and dependency parsing
Aparna Garimella, Carmen Banea, Dirk Hovy, and Rada Mihalcea. 2019 · 2019
Cited alongside, same era.
Designing and interpreting probes with control tasks
John Hewitt and Percy Liang. 2019 · 2019
Cited alongside, same era.
A structural probe for finding syntax in word representations
John Hewitt and Christopher D. Manning. 2019 · 2019
Cited alongside, same era.
Intrinsic probing through dimension selection
Lucas Torroba Hennigen, Adina Williams, and Ryan Cotterell. 2020 · 2020
Later among the works it cites.
Information-theoretic probing with minimum description length
Elena Voita and Ivan Titov. 2020 · 2020
Later among the works it cites.
Probing pretrained language models for lexical semantics
Ivan Vulić, Edoardo Maria Ponti, Robert Litschko, Goran Glavaš, and Anna Korhonen. 2020 · 2020
Later among the works it cites.
Learning which features matter: RoBERTa acquires a preference for linguistic generalizations (eventually)
Alex Warstadt, Yian Zhang, Xiaocheng Li, Haokun Liu, and Samuel R. Bowman. 2020 · 2020
Later among the works it cites.
Probing task-oriented dialogue representation from language models
Chien-Sheng Wu and Caiming Xiong. 2020 · 2020
Later among the works it cites.
Probing pre-trained language models for semantic attributes and their values
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Qiao Jin, Bhuwan Dhingra, William Cohen, and Xinghua Lu. 2019 · 2019
Cited alongside, same era.
Are we consistently biased? multidimensional analysis of biases in distributional word vectors
Anne Lauscher and Goran Glavaš. 2019 · 2019
Cited alongside, same era.
On measuring social biases in sentence encoders
Chandler May, Alex Wang, Shikha Bordia, Samuel R. Bowman, and Rachel Rudinger. 2019 · 2019
Cited alongside, same era.
Language models as knowledge bases?
Fabio Petroni, Tim Rocktäschel, Sebastian Riedel, Patrick Lewis, Anton Bakhtin, Yuxiang Wu, and Alexander Miller. 2019 · 2019
Cited alongside, same era.
Probing multilingual sentence representations with X-probe
Vinit Ravishankar, Lilja Øvrelid, and Erik Velldal. 2019 · 2019
Cited alongside, same era.
Energy and policy considerations for deep learning in NLP
Emma Strubell, Ananya Ganesh, and Andrew McCallum. 2019 · 2019
Cited alongside, same era.
BERT rediscovers the classical NLP pipeline
Ian Tenney, Dipanjan Das, and Ellie Pavlick. 2019 · 2019
Cited alongside, same era.
Meriem Beloucif and Chris Biemann. 2021 · 2021
Later among the works it cites.
Low-complexity probing via finding subnetworks
Steven Cao, Victor Sanh, and Alexander Rush. 2021 · 2021
Later among the works it cites.
A multilingual benchmark for probing negation-awareness with minimal pairs
Mareike Hartmann, Miryam de Lhoneux, Daniel Hershcovich, Yova Kementchedjhieva, Lukas Nielsen, Chen Qiu, and Anders Søgaard. 2021 · 2021
Later among the works it cites.
Language models as knowledge bases: On entity representations, storage capacity, and paraphrased queries
Benjamin Heinzerling and Kentaro Inui. 2021 · 2021
Later among the works it cites.
Probing image-language transformers for verb understanding
Lisa Anne Hendricks and Aida Nematzadeh. 2021 · 2021
Later among the works it cites.
The importance of modeling social factors of language: Theory and practice
Dirk Hovy and Diyi Yang. 2021 · 2021
Later among the works it cites.
Discourse probing of pretrained language models
Fajri Koto, Jey Han Lau, and Timothy Baldwin. 2021 · 2021
Later among the works it cites.
Probing multilingual language models for discourse
Murathan Kurfalı and Robert Östling. 2021 · 2021
Later among the works it cites.
Sustainable modular debiasing of language models
Anne Lauscher, Tobias Lueken, and Goran Glavaš. 2021 · 2021
Later among the works it cites.
Probing across time: What does RoBERTa know and when?
Zeyu Liu, Yizhong Wang, Jungo Kasai, Hannaneh Hajishirzi, and Noah A. Smith. 2021 · 2021
Later among the works it cites.
Probing for bridging inference in transformer language models
Onkar Pandit and Yufang Hou. 2021 · 2021
Later among the works it cites.
How much pretraining data do language models need to learn syntax?
Laura Pérez-Mayos, Miguel Ballesteros, and Leo Wanner. 2021 · 2021
Later among the works it cites.
A Bayesian framework for information-theoretic probing
Tiago Pimentel and Ryan Cotterell. 2021 · 2021
Later among the works it cites.
Probing the probing paradigm: Does probing accuracy entail task relevance?
Abhilasha Ravichander, Yonatan Belinkov, and Eduard Hovy. 2021 · 2021
Later among the works it cites.
A multilabel approach to morphosyntactic probing
Naomi Shapiro, Amandalynne Paullada, and Shane Steinert-Threlkeld. 2021 · 2021
Later among the works it cites.
Sociolectal analysis of pretrained language models
Sheng Zhang, Xin Zhang, Weiming Zhang, and Anders Søgaard. 2021a · 2021
Later among the works it cites.
Factual probing is [MASK]: Learning vs. learning to recall
Zexuan Zhong, Dan Friedman, and Danqi Chen. 2021 · 2021
Later among the works it cites.
Chia-Chien Hung, Anne Lauscher, Dirk Hovy, Simone Paolo Ponzetto, and Goran Glavaš. 2022 · 2022
Closest in time.
Welcome to the modern world of pronouns: Identity-inclusive natural language processing beyond gender
Anne Lauscher, Archie Crowley, and Dirk Hovy. 2022 · 2022
Closest in time.