Large-scale methods for distributionally robust optimization
Daniel Levy, Yair Carmon, John C. Duchi, and Aaron Sidford. 2020 · 2020
Later among the works it cites.
XGLUE: A new benchmark datasetfor cross-lingual pre-training, understanding and generation
Yaobo Liang, Nan Duan, Yeyun Gong, Ning Wu, Fenfei Guo, Weizhen Qi, Ming Gong, Linjun Shou, Daxin Jiang, Guihong Cao, Xiaodong Fan, Ruofei Zhang, Rahul Agrawal, Edward Cui, Sining Wei, Taroon Bharti, Ying Qiao, Jiun-Hung Chen, Winnie Wu, Shuguang Liu, Fan Yang, Daniel Campos, Rangan Majumder, and Ming Zhou. 2020 · 2020
Later among the works it cites.
Decolonial AI: Decolonial theory as sociotechnical foresight in artificial intelligence
Shakir Mohamed, Marie-Therese Png, and William Isaac. 2020 · 2020
Later among the works it cites.
Universal Dependencies v2: An evergrowing multilingual treebank collection
Joakim Nivre, Marie-Catherine de Marneffe, Filip Ginter, Jan Hajič, Christopher D. Manning, Sampo Pyysalo, Sebastian Schuster, Francis Tyers, and Daniel Zeman. 2020 · 2020
Later among the works it cites.
XCOPA: A multilingual dataset for causal commonsense reasoning
Edoardo Maria Ponti, Goran Glavaš, Olga Majewska, Qianchu Liu, Ivan Vulić, and Anna Korhonen. 2020 · 2020
Later among the works it cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2020 · 2020
Later among the works it cites.
Distributionally Robust Neural Networks
Shiori Sagawa, Pang Wei Koh, Tatsunori B. Hashimoto, and Percy Liang. 2020 · 2020
Later among the works it cites.
Commonsense reasoning for natural language processing
Maarten Sap, Vered Shwartz, Antoine Bosselut, Yejin Choi, and Dan Roth. 2020 · 2020
Later among the works it cites.
Xiaomingbot: A Multilingual Robot News Reporter
Runxin Xu, Jun Cao, Mingxuan Wang, Jiaze Chen, Hao Zhou, Ying Zeng, Yuping Wang, Li Chen, Xiang Yin, Xijin Zhang, Songcheng Jiang, Yuxuan Wang, and Lei Li. 2020 · 2020
Later among the works it cites.
Natural language processing for similar languages, varieties, and dialects: A survey
Marcos Zampieri, Preslav Nakov, and Yves Scherrer. 2020 · 2020
Later among the works it cites.
Dziribert: a pre-trained language model for the Algerian dialect
Original
Amine Abdaoui, Mohamed Berrimi, Mourad Oussalah, and Abdelouahab Moussaoui. 2021 · 2021
Later among the works it cites.
Toward an atlas of cultural commonsense for machine reasoning
Anurag Acharya, Kartik Talamadupula, and Mark A. Finlayson. 2021 · 2021
Later among the works it cites.
MasakhaNER: Named entity recognition for African languages
Original
David Ifeoluwa Adelani, Jade Z. Abbott, Graham Neubig, Daniel D’souza, Julia Kreutzer, Constantine Lignos, Chester Palen-Michel, Happy Buzaaba, Shruti Rijhwani, Sebastian Ruder, Stephen Mayhew, Israel Abebe Azime, Shamsuddeen Hassan Muhammad, Chris Chinenye Emezue, Joyce Nakatumba-Nabende, Perez Ogayo, Aremu Anuoluwapo, Catherine Gitau, Derguene Mbaye, Jesujoba O. Alabi, Seid Muhie Yimam, Tajuddeen Gwadabe, Ignatius Ezeani, Rubungo Andre Niyongabo, Jonathan Mukiibi, Verrah Otiende, Iroro Orife, Davis David, Samba Ngom, Tosin P. Adewumi, Paul Rayson, Mofetoluwa Adeyemi, Gerald Muriuki, Emmanuel Anebi, Chiamaka Chukwuneke, Nkiruka Odu, Eric Peter Wairagala, Samuel Oyerinde, Clemencia Siro, Tobius Saul Bateesa, Temilola Oloyede, Yvonne Wambui, Victor Akinode, Deborah Nabagereka, Maurice Katusiime, Ayodele Awokoya, Mouhamadane Mboup, Dibora Gebreyohannes, Henok Tilaye, Kelechi Nwaike, Degaga Wolde, Abdoulaye Faye, Blessing Sibanda, Orevaoghene Ahia, Bonaventure F. P. Dossou, Kelechi Ogueji, Thierno Ibrahima Diop, Abdoulaye Diallo, Adewale Akinfaderin, Tendai Marengereke, and Salomey Osei. 2021 · 2021
Later among the works it cites.
Sexism in the judiciary: The importance of bias definition in NLP and in our courts
Noa Baker Gillis. 2021 · 2021
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big?
Emily M. Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell. 2021 · 2021
Later among the works it cites.
Towards decolonising computational sciences
Original
Abeba Birhane and Olivia Guest. 2021 · 2021
Later among the works it cites.
On the opportunities and risks of foundation models
Original
Rishi Bommasani, Drew A Hudson, Ehsan Adeli, Russ Altman, Simran Arora, Sydney von Arx, Michael S Bernstein, Jeannette Bohg, Antoine Bosselut, Emma Brunskill, et al. 2021 · 2021
Later among the works it cites.
Evaluating the evaluation metrics for style transfer: A case study in multilingual formality transfer
Eleftheria Briakou, Sweta Agrawal, Joel Tetreault, and Marine Carpuat. 2021a · 2021
Later among the works it cites.
A review of human evaluation for style transfer
Eleftheria Briakou, Sweta Agrawal, Ke Zhang, Joel Tetreault, and Marine Carpuat. 2021b · 2021
Later among the works it cites.
Beyond noise: Mitigating the impact of fine-grained semantic divergences on neural machine translation
Eleftheria Briakou and Marine Carpuat. 2021 · 2021
Later among the works it cites.
Dealing with disagreements: Looking beyond the majority vote in subjective annotations
Original
Aida Mostafazadeh Davani, Mark Díaz, and Vinodkumar Prabhakaran. 2021 · 2021
Later among the works it cites.
Documenting the English colossal clean crawled corpus
Original
Jesse Dodge, Maarten Sap, Ana Marasovic, William Agnew, Gabriel Ilharco, Dirk Groeneveld, and Matt Gardner. 2021 · 2021
Later among the works it cites.
Ethnologue: Languages of the world
David M. Eberhard, Gary F. Simons, and Charles D. Fennig. 2021 · 2021
Later among the works it cites.
The importance of modeling social factors of language: Theory and practice
Dirk Hovy and Diyi Yang. 2021 · 2021
Later among the works it cites.
Delphi: Towards machine ethics and norms
Original
Liwei Jiang, Jena D. Hwang, Chandra Bhagavatula, Ronan Le Bras, Maxwell Forbes, Jon Borchardt, Jenny Liang, Oren Etzioni, Maarten Sap, and Yejin Choi. 2021 · 2021
Later among the works it cites.
Evaluation of summarization systems across gender, age, and race
Anna Jørgensen and Anders Søgaard. 2021 · 2021
Later among the works it cites.
Multilingual LAMA: Investigating knowledge in multilingual pretrained language models
Nora Kassner, Philipp Dufter, and Hinrich Schütze. 2021 · 2021
Later among the works it cites.
Algorithmic monoculture and social welfare
Jon Kleinberg and Manish Raghavan. 2021 · 2021
Later among the works it cites.
On language models for creoles
Heather Lent, Emanuele Bugliarello, Miryam de Lhoneux, Chen Qiu, and Anders Søgaard. 2021 · 2021
Later among the works it cites.
Visually grounded reasoning across languages and cultures
Fangyu Liu, Emanuele Bugliarello, Edoardo Maria Ponti, Siva Reddy, Nigel Collier, and Desmond Elliott. 2021 · 2021
Later among the works it cites.
Concepts
Eric Margolis and Stephen Laurence. 2021 · 2021
Later among the works it cites.
A survey on bias and fairness in machine learning
Ninareh Mehrabi, Fred Morstatter, Nripsuta Saxena, Kristina Lerman, and Aram Galstyan. 2021 · 2021
Later among the works it cites.
Designing language technologies for social good: The road not taken
Original
Namrata Mukhija, Monojit Choudhury, and Kalika Bali. 2021 · 2021
Later among the works it cites.
Unpacking the expressed consequences of AI research in broader impact statements
Priyanka Nanayakkara, Jessica Hullman, and Nicholas Diakopoulos. 2021 · 2021
Later among the works it cites.
Adapting entities across languages and cultures
Denis Peskov, Viktor Hangya, Jordan Boyd-Graber, and Alexander Fraser. 2021 · 2021
Later among the works it cites.
Ethical issues across cultures: Managing the differing perspectives of China and the USA
Dennis A. Pitta, Fung Hung-Gay, and Steven Isberg. 1999 · 2021
Later among the works it cites.
Minimax and neyman–Pearson meta-learning for outlier languages
Edoardo Maria Ponti, Rahul Aralikatte, Disha Shrivastava, Siva Reddy, and Anders Søgaard. 2021 · 2021
Later among the works it cites.
On releasing annotator-level labels and information in datasets
Vinodkumar Prabhakaran, Aida Mostafazadeh Davani, and Mark Diaz. 2021 · 2021
Later among the works it cites.
Institutionalizing ethics in AI through broader impact requirements
Carina E. A. Prunkl, Carolyn Ashurst, Markus Anderljung, Helena Webb, Jan Leike, and Allan Dafoe. 2021 · 2021
Later among the works it cites.
XTREME-R: Towards more challenging and nuanced multilingual evaluation
Original
Sebastian Ruder, Noah Constant, Jan Botha, Aditya Siddhant, Orhan Firat, Jinlan Fu, Pengfei Liu, Junjie Hu, Graham Neubig, and Melvin Johnson. 2021 · 2021
Later among the works it cites.
It’s basically the same language anyway: the case for a nordic language model
Magnus Sahlgren, Fredrik Carlsson, Fredrik Olsson, and Love Börjeson. 2021 · 2021
Later among the works it cites.
COM2SENSE: A commonsense reasoning benchmark with complementary sentences
Shikhar Singh, Nuan Wen, Yu Hou, Pegah Alipoormolabashi, Te-lin Wu, Xuezhe Ma, and Nanyun Peng. 2021 · 2021
Later among the works it cites.
Process for adapting language models to society (PALMS) with values-targeted datasets
Irene Solaiman and Christy Dennison. 2021 · 2021
Later among the works it cites.
Cross-cultural similarity features for cross-lingual transfer learning of pragmatically motivated tasks
Jimin Sun, Hwijeen Ahn, Chan Young Park, Yulia Tsvetkov, and David R. Mortensen. 2021 · 2021
Later among the works it cites.
A word on machine ethics: A response to Jiang et al. (2021)
Zeerak Talat, Hagen Blix, Josef Valvoda, Maya Indira Ganesh, Ryan Cotterell, and Adina Williams. 2021 · 2021
Later among the works it cites.
Measuring and reducing gendered correlations in pre-trained models
Original
Kellie Webster, Xuezhi Wang, Ian Tenney, Alex Beutel, Emily Pitler, Ellie Pavlick, Jilin Chen, Ed Chi, and Slav Petrov. 2021 · 2021
Later among the works it cites.
mT5: A massively multilingual pre-trained text-to-text transformer
Linting Xue, Noah Constant, Adam Roberts, Mihir Kale, Rami Al-Rfou, Aditya Siddhant, Aditya Barua, and Colin Raffel. 2021 · 2021
Later among the works it cites.
Broaden the vision: Geo-diverse visual commonsense reasoning
Da Yin, Liunian Harold Li, Ziniu Hu, Nanyun Peng, and Kai-Wei Chang. 2021 · 2021
Later among the works it cites.
Sociolectal analysis of pretrained language models
Sheng Zhang, Xin Zhang, Weiming Zhang, and Anders Søgaard. 2021 · 2021
Later among the works it cites.
Distributionally robust multilingual machine translation
Original
Chunting Zhou, Daniel Levy, Xian Li, Marjan Ghazvininejad, and Graham Neubig. 2021 · 2021
Later among the works it cites.
Zero-shot dependency parsing with worst-case aware automated curriculum learning
Miryam de Lhoneux, Sheng Zhang, and Anders Søgaard. 2022 · 2022
Closest in time.
Jury learning: Integrating dissenting voices into machine learning models
Original
Mitchell L Gordon, Michelle S Lam, Joon Sung Park, Kayur Patel, Jeffrey T Hancock, Tatsunori Hashimoto, and Michael S Bernstein. 2022 · 2022
Closest in time.
Good night at 4 pm?! Time expressions in different cultures
Vered Shwartz. 2022 · 2022
Closest in time.