Fetching the paper…
Reading the bibliography…
Code-Switching, a common phenomenon in written text and conversation, has been studied over decades by the natural language processing (NLP) research community.
A survey of code-switched speech and language processing
Sunayana Sitaram, Khyathi Raghavi Chandu, Sai Krishna Rallabandi, and Alan W Black. 2019 · 1904
Earlier work this paper cites.
A study of lexical and prosodic cues to segmentation in a hindi-english code-switched discourse
Preeti Rao, Mugdha Pandya, Kamini Sabu, Kanhaiya Kumar, and Nandini Bondale. 2018 · 1922
Earlier work this paper cites.
Study of semi-supervised approaches to improving english-mandarin code-switching speech recognition
Pengcheng Guo, Haihua Xu, Lei Xie, and Eng Siong Chng. 2018 · 1932
Earlier work this paper cites.
Acoustic and textual data augmentation for improved asr of code-switching speech
E Yilmaz, H Heuvel, and DA van Leeuwen. 2018 · 1937
Earlier work this paper cites.
Acoustic and textual data augmentation for improved asr of code-switching speech
Emre Yılmaz, Henk van den Heuvel, and David van Leeuwen. 2018 · 1937
Earlier work this paper cites.
The role of cognate words, pos tags and entrainment in code-switching
Victor Soto, Nishmar Cestero, and Julia Hirschberg. 2018 · 1942
Earlier work this paper cites.
Homophone identification and merging for code-switched speech recognition
Brij Mohan Lal Srivastava and Sunayana Sitaram. 2018 · 1947
Earlier work this paper cites.
Code-switching in indic speech synthesisers
Anju Leela Thomas, Anusha Prakash, Arun Baby, and Hema Murthy. 2018a · 1952
Earlier work this paper cites.
Code-switching in indic speech synthesisers
Anju Leela Thomas, Anusha Prakash, Arun Baby, and Hema A Murthy. 2018b · 1952
Earlier work this paper cites.
A novel approach for effective recognition of the code-switched data on monolingual language model
Sreeram Ganji and Rohit Sinha. 2018 · 1957
Earlier work this paper cites.
Syntactic structure and social function of code-switching , volume 2
Shana Poplack. 1978 · 1978
Earlier work this paper cites.
Sometimes i’ll start a sentence in spanish y termino en espanol: toward a typology of code-switching
Shana Poplack. 1980 · 1980
Earlier work this paper cites.
Processing of sentences with intra-sentential code-switching
Aravind Joshi. 1982 · 1982
Earlier work this paper cites.
Estimating code-switching on twitter with a novel generalized word-level language detection technique
Shruti Rijhwani, Royal Sequiera, Monojit Choudhury, Kalika Bali, and Chandra Shekhar Maddila. 2017 · 1982
Earlier work this paper cites.
Seame: a mandarin-english code-switching speech corpus in south-east asia
Dau-Cheng Lyu, Tien-Ping Tan, Eng Siong Chng, and Haizhou Li. 2010b · 1989
Earlier work this paper cites.
Code switching and x-bar theory: The functional head constraint
Hedi M Belazi, Edward J Rubin, and Almeida Jacqueline Toribio. 1994 · 1994
Earlier work this paper cites.
Duelling languages: Grammatical structure in codeswitching
Carol Myers-Scotton. 1997 · 1997
Earlier work this paper cites.
The production of code-mixed discourse
David Sankoff. 1998 · 1998
Earlier work this paper cites.
Mixed language query disambiguation
Pascale Fung, Xiaohu Liu, and Chi-Shun Cheung. 1999 · 1999
Earlier work this paper cites.
The dynamics of code-copying in language encounters
Lars Johanson. 1999 · 1999
Earlier work this paper cites.
A global perspective on bilingualism and bilingual education
G Richard Tucker. 2001 · 1999
Earlier work this paper cites.
Model distance and it’s application on mixed language speech recognition system
Guokang Fu and Liqin Shen. 2000 · 2000
Earlier work this paper cites.
Code-switching and language attrition: Evidence from american finnish interview speech
Pekka Hirvonen and Timo Lauttamus. 2000 · 2000
Earlier work this paper cites.
Mixed-lingual spoken word recognition by using vq codebook sequences of variable length segments
Hiroaki Kojima and Kazuyo Tanaka. 2003 · 2003
Earlier work this paper cites.
A general approach to tts reading of mixed-language texts
Leonardo Badino, Claudia Barolo, and Silvia Quazza. 2004 · 2004
Earlier work this paper cites.
Detection of language boundary in code-switching utterances by bi-phone probabilities
Joyce YC Chan, PC Ching, Tan Lee, and Helen M Meng. 2004 · 2004
Earlier work this paper cites.
Multilingual e-mail text processing for speech synthesis
Daniela Oria and Akos Vetek. 2004 · 2004
Earlier work this paper cites.
Chinese-english mixed-lingual keyword spotting
Shan-Ruei You, Shih-Chieh Chien, Chih-Hsing Hsu, Ke-Shiu Chen, Jia-Jang Tu, Jeng Shien Lin, and Sen-Chia Chang. 2004 · 2004
Earlier work this paper cites.
Development of a cantonese-english code-mixing speech corpus
Joyce YC Chan, PC Ching, and Tan Lee. 2005 · 2005
Earlier work this paper cites.
A transformation-based learning approach to language identification for mixed-lingual text-to-speech synthesis
J. C. Marcadet, V. Fischer, and C. Waast-Richard. 2005 · 2005
Earlier work this paper cites.
Multiple voices: An introduction to bilingualism
Carol Myers-Scotton. 2005 · 2005
Earlier work this paper cites.
Mandarin/english mixed-lingual name recognition for mobile phone
Xiaolin Ren, Xin He, and Yaxin Zhang. 2005 · 2005
Earlier work this paper cites.
Phonetic labeling and segmentation of mixed-lingual prosody databases
Harald Romsdorfer and Beat Pfister. 2005 · 2005
Earlier work this paper cites.
Automatic speech recognition of cantonese-english code-mixing utterances
Joyce YC Chan, PC Ching, Tan Lee, and Houwei Cao. 2006 · 2006
Earlier work this paper cites.
Character stream parsing of mixed-lingual text
Harald Romsdorfer and Beat Pfister. 2006 · 2006
Earlier work this paper cites.
Learning to recognize code-switched speech without forgetting monolingual speech recognition
Sanket Shah, Basil Abraham, Sunayana Sitaram, Vikas Joshi, et al. 2020 · 2006
Earlier work this paper cites.
Language identification on code-switching speech
Chyng-Leei Chu, Dau-cheng Lyu, and Ren-yuan Lyu. 2007 · 2007
Earlier work this paper cites.
An hmm-based bilingual (mandarin-english) tts
Hui Liang, Yao Qian, and Frank K Soong. 2007 · 2007
Earlier work this paper cites.
A tagging algorithm for mixed language identification in a noisy domain
Mike Rosner and Paulseph-John Farrugia. 2007 · 2007
Earlier work this paper cites.
Prosodic variation in cantonese-english code-mixed speech
Wen-Tao Gu, Tan Lee, and P. C. Ching. 2008 · 2008
Earlier work this paper cites.
Language identification on code-switching utterances using multiple cues
Dau-Cheng Lyu and Ren-Yuan Lyu. 2008 · 2008
Earlier work this paper cites.
Accent identification in the presence of code-mixing
Thomas Niesler and Febe de Wet. 2008 · 2008
Earlier work this paper cites.
Hmm-based mixed-language (mandarin-english) speech synthesis
Yao Qian, Houwei Cao, and Frank K Soong. 2008 · 2008
Earlier work this paper cites.
Learning to predict code-switching points
Thamar Solorio and Yang Liu. 2008a · 2008
Earlier work this paper cites.
Part-of-speech tagging for english-spanish code-switched text
Thamar Solorio and Yang Liu. 2008b · 2008
Earlier work this paper cites.
An investigation of acoustic models for multilingual code-switching
Christopher M White, Sanjeev Khudanpur, and James K Baker. 2008 · 2008
Earlier work this paper cites.
Prosody modification on mixed-language speech synthesis
Yi Zhang and Jian-Hua Tao. 2008 · 2008
Earlier work this paper cites.
Effects of language mixing for automatic recognition of cantonese-english code-mixing utterances
Houwei Cao, P. C. Ching, and Tan Lee. 2009 · 2009
Earlier work this paper cites.
Automatic recognition of cantonese-english code-mixing speech
Joyce YC Chan, Houwei Cao, PC Ching, and Tan Lee. 2009 · 2009
Earlier work this paper cites.
Language attrition and code-switching among us americans in germany
Inke Du Bois. 2009 · 2009
Earlier work this paper cites.
Benglish verbs: A case of code-mixing in bengali
Shishir Bhattacharja. 2010 · 2010
Earlier work this paper cites.
Towards mixed language speech recognition systems
David Imseng, Hervé Bourlard, and Mathew Magimai Doss. 2010 · 2010
Earlier work this paper cites.
Hmm based tts for mixed language text
Zhiwei Shuang, Shiyin Kang, Yong Qin, Lirong Dai, and Lianhong Cai. 2010 · 2010
Earlier work this paper cites.
Language identification of code switching malay-english words using syllable structure information
Yin-Lai Yeong and Tien-Ping Tan. 2010 · 2010
Earlier work this paper cites.
Feasibility of leveraging crowd sourcing for the creation of a large scale annotated resource for hindi english code switched data: A pilot annotation
Mona Diab and Ankit Kamboj. 2011 · 2011
Earlier work this paper cites.
Subword-level language identification for intra-word code-switching
Manuel Mager, Özlem Çetinoğlu, and Katharina Kann. 2019 · 2011
Earlier work this paper cites.
Unknown word extraction from multilingual code-switching sentences
Yi-Lun Wu, Chaio-Wen Hsieh, Wei-Hsuan Lin, Chun-Yi Liu, and Liang-Chih Yu. 2011 · 2011
Earlier work this paper cites.
Token level identification of linguistic code switching
Heba Elfardy and Mona Diab. 2012 · 2012
Earlier work this paper cites.
Turning a monolingual speaker into multilingual for a mixed-language tts
Ji He, Yao Qian, Frank K Soong, and Sheng Zhao. 2012 · 2012
Earlier work this paper cites.
Code-switch language model with inversion constraints for mixed language speech recognition
Ying Li and Pascale Fung. 2012 · 2012
Earlier work this paper cites.
A mandarin-english code-switching corpus
Ying Li, Yue Yu, and Pascale Fung. 2012 · 2012
Earlier work this paper cites.
Pattern matching refinements to dictionary-based code-switching point detection
Nathaniel Oco and Rachel Edita Roxas. 2012 · 2012
Earlier work this paper cites.
An introduction to conditional random fields
Charles Sutton, Andrew McCallum, et al. 2012 · 2012
Earlier work this paper cites.
Integration of language identification into a recognition system for spoken conversations containing code-switches
Jochen Weiner, Ngoc Thang Vu, Dominic Telaar, Florian Metze, Tanja Schultz, Dau-Cheng Lyu, Eng-Siong Chng, and Haizhou Li. 2012b · 2012
Earlier work this paper cites.
A language modeling approach to identifying code-switched sentences and words
Liang-Chih Yu, Wei-Cheng He, and Wei-Nan Chien. 2012 · 2012
Earlier work this paper cites.
Combination of recurrent neural networks and factored language models for code-switching language modeling
Heike Adel, Ngoc Thang Vu, and Tanja Schultz. 2013 · 2013
Earlier work this paper cites.
Language modeling for mixed language speech recognition using weighted phrase extraction
Ying Li and Pascale Fung. 2013 · 2013
Earlier work this paper cites.
Code-switching event detection based on delta-bic using phonetic eigenvoice models
Wei-Bin Liang, Chung-Hsien Wu, and Chun-Shan Hsu. 2013 · 2013
Earlier work this paper cites.
Features for factored language models for code-switching speech
Heike Adel, Katrin Kirchhoff, Dominic Telaar, Ngoc Thang Vu, Tim Schlippe, and Tanja Schultz. 2014a · 2014
Earlier work this paper cites.
Comparing approaches to convert recurrent neural networks into backoff language models for efficient decoding
Heike Adel, Katrin Kirchhoff, Ngoc Thang Vu, Dominic Telaar, and Tanja Schultz. 2014b · 2014
Earlier work this paper cites.
“i am borrowing ya mixing?" an analysis of english-hindi code mixing in facebook
Kalika Bali, Jatin Sharma, Monojit Choudhury, and Yogarshi Vyas. 2014 · 2014
Earlier work this paper cites.
The tel aviv university system for the code-switching workshop shared task
Kfir Bar and Nachum Dershowitz. 2014 · 2014
Earlier work this paper cites.
Mixed language and code-switching in the canadian hansard
Marine Carpuat. 2014 · 2014
Earlier work this paper cites.
Word-level language identification using crf: Code-switching shared task report of msr india system
Gokul Chittaranjan, Yogarshi Vyas, Kalika Bali, and Monojit Choudhury. 2014 · 2014
Earlier work this paper cites.
Identifying languages at the word level in code-mixed indian social media text
Amitava Das and Björn Gambäck. 2014 · 2014
Earlier work this paper cites.
A hindi-english code-switching corpus
Anik Dey and Pascale Fung. 2014 · 2014
Earlier work this paper cites.
Aida: Identifying code switching in informal arabic text
Heba Elfardy, Mohamed Al-Badrashiny, and Mona Diab. 2014 · 2014
Earlier work this paper cites.
Foreign words and the automatic processing of arabic social media text written in roman script
Ramy Eskander, Mohamed Al-Badrashiny, Nizar Habash, and Owen Rambow. 2014 · 2014
Earlier work this paper cites.
Language identification of individual words with joint sequence models
Oluwapelumi Giwa and Marelie H. Davel. 2014 · 2014
Earlier work this paper cites.
Improving word alignment using linguistic code switching data
Fei Huang and Alexander Yates. 2014 · 2014
Earlier work this paper cites.
Language identification in code-switching scenario
Naman Jain and Riyaz Ahmad Bhat. 2014 · 2014
Earlier work this paper cites.
Word-level language identification in bi-lingual code-switched texts
Harsh Jhamtani, Suleep Kumar Bhogi, and Vaskar Raychoudhury. 2014 · 2014
Earlier work this paper cites.
Twitter users# codeswitch hashtags!# moltoimportante# wow
David Jurgens, Stefan Dimitrov, and Derek Ruths. 2014 · 2014
Earlier work this paper cites.
The iucl+ system: Word-level language identification via extended markov models
Levi King, Eric Baucom, Timur Gilmanov, Sandra Kübler, Daniel Whyatt, Wolfgang Maier, and Paul Rodrigues. 2014 · 2014
Earlier work this paper cites.
Language modeling with functional head constraint for code switching speech recognition
Ying Li and Pascale Fung. 2014 · 2014
Earlier work this paper cites.
The cmu submission for the shared task on language identification in code-switched data
Chu-Cheng Lin, Waleed Ammar, Lori Levin, and Chris Dyer. 2014 · 2014
Earlier work this paper cites.
Code-switching speech recognition for closely related languages
Tetyana Lyudovyk and Valeriy Pylypenko. 2014 · 2014
Earlier work this paper cites.
Modeling code-switching speech on under-resourced languages for language identification
Koena Ronny Mabokela, Madimetja Jonas Manamela, and Mabu Manaileng. 2014 · 2014
Earlier work this paper cites.
Predicting code-switching in multilingual communication for immigrant communities
Evangelos Papalexakis, Dong Nguyen, and A Seza Doğruöz. 2014 · 2014
Earlier work this paper cites.
Learning polylingual topic models from code-switched social media documents
Nanyun Peng, Yiming Wang, and Mark Dredze. 2014 · 2014
Earlier work this paper cites.
Prosodic cues to monolingual versus code-switching sentences in english and spanish
Page Piccinini and Marc Garellek. 2014 · 2014
Earlier work this paper cites.
Incremental n-gram approach for language identification in code-switched text
Prajwol Shrestha. 2014 · 2014
Earlier work this paper cites.
Overview for the first shared task on language identification in code-switched data
Thamar Solorio, Elizabeth Blair, Suraj Maharjan, Steven Bethard, Mona Diab, Mahmoud Ghoneim, Abdelati Hawwari, Fahad AlGhamdi, Julia Hirschberg, Alison Chang, et al. 2014 · 2014
Earlier work this paper cites.
Detecting code-switching in a multilingual alpine heritage corpus
Martin Volk and Simon Clematide. 2014 · 2014
Earlier work this paper cites.
Finding romanized arabic dialect in code-mixed tweets
Clare Voss, Stephen Tratz, Jamal Laoudi, and Douglas Briesch. 2014 · 2014
Earlier work this paper cites.
Exploration of the impact of maximum entropy in recurrent neural network language models for code-switching speech
Ngoc Thang Vu and Tanja Schultz. 2014 · 2014
Earlier work this paper cites.
Pos tagging of english-hindi code-mixed social media content
Yogarshi Vyas, Spandana Gella, Jatin Sharma, Kalika Bali, and Monojit Choudhury. 2014 · 2014
Earlier work this paper cites.
Language identification of code switching sentences and multilingual sentences of under-resourced languages by using multi structural word information
Yin-Lai Yeong and Tien-Ping Tan. 2014 · 2014
Earlier work this paper cites.
Part-of-speech tagging for code-mixed english-hindi twitter and facebook chat messages
Anupam Jamatia, Björn Gambäck, and Amitava Das. 2015 · 2015
Earlier work this paper cites.
Pos tagging of hindi-english code mixed text from social media: Some machine learning experiments
Royal Sequiera, Monojit Choudhury, and Kalika Bali. 2015 · 2015
Earlier work this paper cites.
Emotion detection in code-switching texts via bilingual and sentimental information
Zhongqing Wang, Sophia Lee, Shoushan Li, and Guodong Zhou. 2015 · 2015
Earlier work this paper cites.
The george washington university system for the code-switching workshop shared task 2016
Mohamed Al-Badrashiny and Mona Diab. 2016 · 2016
Earlier work this paper cites.
Part of speech tagging for code switched data
Fahad AlGhamdi, Giovanni Molina, Mona Diab, Thamar Solorio, Abdelati Hawwari, Victor Soto, and Julia Hirschberg. 2016 · 2016
Earlier work this paper cites.
Use of semantic knowledge base for enhancement of coherence of code-mixed topic-based aspect clusters
Kavita Asnani and Jyoti Pawar. 2016 · 2016
Earlier work this paper cites.
Part-of-speech tagging of code-mixed social media content: Pipeline, stacking and joint modelling
Utsab Barman, Joachim Wagner, and Jennifer Foster. 2016 · 2016
Earlier work this paper cites.
Functions of code-switching in tweets: An annotation framework and some initial experiments
Rafiya Begum, Kalika Bali, Monojit Choudhury, Koustav Rudra, and Niloy Ganguly. 2016 · 2016
Earlier work this paper cites.
A turkish-german code-switching corpus
Özlem Çetinoğlu. 2016 · 2016
Earlier work this paper cites.
Challenges of computational processing of code-switching
Özlem Çetinoğlu, Sarah Schulz, and Ngoc Thang Vu. 2016 · 2016
Earlier work this paper cites.
Columbia-jadavpur submission for emnlp 2016 code-switching workshop shared task: System description
Arunavha Chanda, Dipankar Das, and Chandan Mazumdar. 2016a · 2016
Earlier work this paper cites.
Creating a large multi-layered representational repository of linguistic code switched arabic data
Mona Diab, Mahmoud Ghoneim, Abdelati Hawwari, Fahad AlGhamdi, Nada Almarwani, and Mohamed Al-Badrashiny. 2016 · 2016
Earlier work this paper cites.
Comparing the level of code-switching in corpora
Björn Gambäck and Amitava Das. 2016 · 2016
Earlier work this paper cites.
Part-of-speech tagging of code-mixed social media text
Souvick Ghosh, Satanu Ghosh, and Dipankar Das. 2016 · 2016
Earlier work this paper cites.
Opinion mining in a code-mixed environment: A case study with government portals
Deepak Gupta, Ankit Lamba, Asif Ekbal, and Pushpak Bhattacharyya. 2016 · 2016
Earlier work this paper cites.
Simple tools for exploring variation in code-switching for linguists
Gualberto A Guzman, Jacqueline Serigos, Barbara Bullock, and Almeida Jacqueline Toribio. 2016 · 2016
Earlier work this paper cites.
A neural model for language identification in code-switched tweets
Aaron Jaech, George Mulcaire, Mari Ostendorf, and Noah A Smith. 2016 · 2016
Earlier work this paper cites.
Towards sub-word level compositions for sentiment analysis of hindi-english code mixed text
Aditya Joshi, Ameya Prabhu, Manish Shrivastava, and Vasudeva Varma. 2016 · 2016
Earlier work this paper cites.
Fasttext.zip: Compressing text classification models
Armand Joulin, Edouard Grave, Piotr Bojanowski, Matthijs Douze, Hérve Jégou, and Tomas Mikolov. 2016 · 2016
Earlier work this paper cites.
Overview for the second shared task on language identification in code-switched data
Giovanni Molina, Fahad AlGhamdi, Mahmoud Ghoneim, Abdelati Hawwari, Nicolas Rey-Villamizar, Mona Diab, and Thamar Solorio. 2016 · 2016
Earlier work this paper cites.
Word-level language identification and predicting codeswitching points in swahili-english language data
Mario Piergallini, Rouzbeh Shirvani, Gauri Shankar Gautam, and Mohamed Chouikha. 2016 · 2016
Earlier work this paper cites.
An arabic-moroccan darija code-switched corpus
Younes Samih and Wolfgang Maier. 2016 · 2016
Earlier work this paper cites.
Shallow parsing pipeline-hindi-english code-mixed social media text
Arnav Sharma, Sakshi Gupta, Raveesh Motlani, Piyush Bansal, Manish Shrivastava, Radhika Mamidi, and Dipti Misra Sharma. 2016 · 2016
Earlier work this paper cites.
The howard university system submission for the shared task in language identification in spanish-english codeswitching
Rouzbeh Shirvani, Mario Piergallini, Gauri Shankar Gautam, and Mohamed Chouikha. 2016 · 2016
Earlier work this paper cites.
Codeswitching detection via lexical features in conditional random fields
Prajwol Shrestha. 2016 · 2016
Earlier work this paper cites.
Language identification in code-switched text using conditional random fields and babelnet
Utpal Kumar Sikdar and Björn Gambäck. 2016 · 2016
Earlier work this paper cites.
Speech synthesis of code-mixed text
Sunayana Sitaram and Alan W Black. 2016 · 2016
Earlier work this paper cites.
Experiments with cross-lingual systems for synthesis of code-mixed text
Sunayana Sitaram, Sai Krishna Rallabandi, Shruti Rijhwani, and Alan W. Black. 2016 · 2016
Earlier work this paper cites.
En-es-cs: An english-spanish code-switching twitter corpus for multilingual sentiment analysis
David Vilares, Miguel A Alonso, and Carlos Gómez-Rodríguez. 2016 · 2016
Earlier work this paper cites.
A bilingual attention network for code-switched emotion prediction
Zhongqing Wang, Yue Zhang, Sophia Lee, Shoushan Li, and Guodong Zhou. 2016 · 2016
Earlier work this paper cites.
Codeswitching language identification using subword information enriched word vectors
Meng Xuan Xia. 2016 · 2016
Earlier work this paper cites.
Accurate pinyin-english codeswitched language identification
Meng Xuan Xia and Jackie Chi Kit Cheung. 2016 · 2016
Earlier work this paper cites.
A longitudinal bilingual frisian-dutch radio broadcast database designed for code-switching research
Emre Yilmaz, Maaike Andringa, Sigrid Kingma, Jelske Dijkstra, Frits van der Kuip, Hans Van de Velde, Frederik Kampstra, Jouke Algra, Henk van den Heuvel, and David van Leeuwen. 2016 · 2016
Earlier work this paper cites.
Open source speech and language resources for frisian
Emre Yılmaz, Henk van den Heuvel, Jelske Dijkstra, Hans Van de Velde, Frederik Kampstra, Jouke Algra, and David Van Leeuwen. 2016 · 2016
Earlier work this paper cites.
Addressing code-switching in french/algerian arabic speech
Djegdjiga Amazouz, Martine Adda-Decker, and Lori Lamel. 2017 · 2017
Earlier work this paper cites.
Joining hands: Exploiting monolingual treebanks for parsing of code-mixing data
Irshad Bhat, Riyaz Ahmad Bhat, Manish Shrivastava, and Dipti Misra Sharma. 2017 · 2017
Earlier work this paper cites.
Speech synthesis for mixed-language navigation instructions
Khyathi Raghavi Chandu, SaiKrishna Rallabandi, Sunayana Sitaram, and Alan W. Black. 2017 · 2017
Earlier work this paper cites.
Curriculum design for code-switching: Experiments with language identification and language modeling with deep neural networks
Monojit Choudhury, Kalika Bali, Sunayana Sitaram, and Ashutosh Baheti. 2017 · 2017
Earlier work this paper cites.
Word translation without parallel data
Alexis Conneau, Guillaume Lample, Marc’Aurelio Ranzato, Ludovic Denoyer, and Hervé Jégou. 2017 · 2017
Earlier work this paper cites.
Multilingual semantic parsing and code-switching
Long Duong, Hadi Afshar, Dominique Estival, Glen Pink, Philip R Cohen, and Mark Johnson. 2017 · 2017
Earlier work this paper cites.
Metrics for modeling code-switching across corpora
Gualberto A Guzmán, Joseph Ricard, Jacqueline Serigos, Barbara E Bullock, and Almeida Jacqueline Toribio. 2017 · 2017
Cited alongside, same era.
Metrics for modeling code-switching across corpora
Gualberto Guzmán, Joseph Ricard, Jacqueline Serigos, Barbara E. Bullock, and Almeida Jacqueline Toribio. 2017 · 2017
Cited alongside, same era.
Towards developing a phonetically balanced code-mixed speech corpus for hindi-english asr
Ayushi Pandey, Brij Mohan Lal Srivastava, and Suryakanth Gangashetty. 2017 · 2017
Cited alongside, same era.
Towards normalising konkani-english code-mixed social media text
Akshata Phadte and Gaurish Thakkar. 2017 · 2017
Cited alongside, same era.
Quantitative characterization of code switching patterns in complex multi-party conversations: A case study on hindi movie scripts
Adithya Pratapa and Monojit Choudhury. 2017 · 2017
Cited alongside, same era.
Malayalam-english code-switched: Grapheme to phoneme system
Sreeja Manghat, Sreeram Manghat, and Tanja Schultz. 2020 · 2020
Later among the works it cites.
One model, many languages: Meta-learning for multilingual text-to-speech
Tomáš Nekvinda and Ondřej Dušek. 2020 · 2020
Later among the works it cites.
Canvec-the canberra vietnamese-english code-switching natural speech corpus
Li Nguyen and Christopher Bryant. 2020 · 2020
Later among the works it cites.
Palomino-ochoa at semeval-2020 task 9: Robust system based on transformer for code-mixed sentiment classification
Daniel Palomino and José Ochoa-Luna. 2020 · 2020
Later among the works it cites.
Towards code-switched classification exploiting constituent language resources
Kartikey Pant and Tanvi Dadu. 2020 · 2020
Later among the works it cites.
Understanding linguistic accommodation in code-switched human-machine dialogues
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
SaiKrishna Rallabandi and Alan W. Black. 2017 · 2017
Cited alongside, same era.
Jee haan, i’d like both, por favor: Elicitation of a code-switched corpus of hindi–english and spanish–english human–machine dialog
Vikram Ramanarayanan and David Suendermann-Oeft. 2017 · 2017
Cited alongside, same era.
Crowdsourcing universal part-of-speech tags for code-switching
Victor Soto and Julia Hirschberg. 2017 · 2017
Cited alongside, same era.
Synthesising isizulu-english code-switch bigrams using word embeddings
Ewald van der Westhuizen and Thomas Niesler. 2017 · 2017
Cited alongside, same era.
Longitudinal speaker clustering and verification corpus with code-switching frisian-dutch speech
Emre Yılmaz, Jelske Dijkstra, Hans Van de Velde, Frederik Kampstra, Jouke Algra, Henk van den Heuvel, and David Van Leeuwen. 2017a · 2017
Cited alongside, same era.
Exploiting untranscribed broadcast data for improved code-switching detection
Emre Yılmaz, Henk van den Heuvel, and David Van Leeuwen. 2017b · 2017
Cited alongside, same era.
Improving neural network performance by injecting background knowledge: Detecting code-switching and borrowing in algerian texts
Wafia Adouane, Jean-Philippe Bernardy, and Simon Dobnik. 2018 · 2018
Cited alongside, same era.
Tanmay Parekh, Emily Ahn, Yulia Tsvetkov, and Alan W Black. 2020 · 2020
Later among the works it cites.
Irlab_daiict at semeval-2020 task 9: Machine learning and deep learning methods for sentiment analysis of code-mixed tweets
Apurva Parikh, Abhimanyu Singh Bisht, and Prasenjit Majumder. 2020 · 2020
Later among the works it cites.
Semeval-2020 task 9: Overview of sentiment analysis of code-mixed tweets
Parth Patwa, Gustavo Aguilar, Sudipta Kar, Suraj Pandey, Srinivas Pykl, Björn Gambäck, Tanmoy Chakraborty, Thamar Solorio, and Amitava Das. 2020 · 2020
Later among the works it cites.
Towards context-aware end-to-end code-switching speech recognition
Zimeng Qiu, Yiyuan Li, Xinjian Li, Florian Metze, and William M. Campbell. 2020 · 2020
Later among the works it cites.
A comparative study of different state-of-the-art hate speech detection methods in hindi-english code-mixed data
Priya Rani, Shardul Suryawanshi, Koustava Goswami, Bharathi Raja Chakravarthi, Theodorus Fransen, and John Philip McCrae. 2020 · 2020
Later among the works it cites.
Contextual embeddings for arabic-english code-switched data
Caroline Sabty, Mohamed Islam, and Slim Abdennadher. 2020 · 2020
Later among the works it cites.
Improving low resource code-switched asr using augmented code-switched tts
Yash Sharma, Basil Abraham, Karan Taneja, and Preethi Jyothi. 2020 · 2020
Later among the works it cites.
Sentiment analysis for hinglish code-mixed tweets by means of cross-lingual word embeddings
Pranaydeep Singh and Els Lefever. 2020 · 2020
Later among the works it cites.
Msr india at semeval-2020 task 9: Multilingual models can do code-mixing too
Anirudh Srinivasan. 2020 · 2020
Later among the works it cites.
Code-mixed parse trees and how to find them
Anirudh Srinivasan, Sandipan Dandapat, and Monojit Choudhury. 2020 · 2020
Later among the works it cites.
Hcms at semeval-2020 task 9: A neural approach to sentiment analysis for code-mixed texts
Aditya Srivastava and V Harsha Vardhan. 2020 · 2020
Later among the works it cites.
Iit gandhinagar at semeval-2020 task 9: Code-mixed sentiment classification using candidate sentence generation and selection
Vivek Srivastava and Mayank Singh. 2020a · 2020
Later among the works it cites.
Phinc: A parallel hinglish social media code-mixed corpus for machine translation
Vivek Srivastava and Mayank Singh. 2020b · 2020
Later among the works it cites.
Evaluating word embeddings for indonesian–english code-mixed text based on synthetic data
Sara Stymne et al. 2020 · 2020
Later among the works it cites.
Wessa at semeval-2020 task 9: Code-mixed sentiment analysis using transformers
Ahmed Sultan, Mahmoud Salim, Amina Gaber, and Islam El Hosary. 2020 · 2020
Later among the works it cites.
Exploring lexicon-free modeling units for end-to-end korean and korean-english code-switching speech recognition
Jisung Wang, Jihwan Kim, Sangki Kim, and Yeha Lee. 2020 · 2020
Later among the works it cites.
Semi-supervised acoustic modelling for five-lingual code-switched asr using automatically-segmented soap opera speech
Nick Wilkinson, Astik Biswas, Emre Yilmaz, Febe De Wet, Thomas Niesler, et al. 2020 · 2020
Later among the works it cites.
Meta-transfer learning for code-switched speech recognition
Genta Indra Winata, Samuel Cahyawijaya, Zhaojiang Lin, Zihan Liu, Peng Xu, and Pascale Fung. 2020 · 2020
Later among the works it cites.
Meistermorxrc at semeval-2020 task 9: Fine-tune bert and multitask learning for sentiment analysis of code-mixed tweets
Qi Wu, Peng Wang, and Chenghao Huang. 2020 · 2020
Later among the works it cites.
Csp: Code-switching pre-training for neural machine translation
Zhen Yang, Bojie Hu, Ambyera Han, Shen Huang, and Qi Ju. 2020 · 2020
Later among the works it cites.
A preliminary study on leveraging meta learning technique for code-switching speech recognition
Fu-Hao Yu and Kuan-Yu Chen. 2020 · 2020
Later among the works it cites.
Upb at semeval-2020 task 9: Identifying sentiment in code-mixed social media texts using transformers and multi-task learning
George-Eduard Zaharia, George-Alexandru Vlad, Dumitru-Clementin Cercel, Traian Rebedea, and Costin Chiru. 2020 · 2020
Later among the works it cites.
Monolingual data selection analysis for english-mandarin hybrid code-switching speech recognition
Haobo Zhang, Haihua Xu, Van Tung Pham, Hao Huang, and Eng Siong Chng. 2020 · 2020
Later among the works it cites.
Towards natural bilingual and code-switched speech synthesis based on mix of monolingual recordings and cross-lingual voice conversion
Shengkui Zhao, Trung Hieu Nguyen, Hao Wang, and Bin Ma. 2020 · 2020
Later among the works it cites.
Multi-encoder-decoder transformer for code-switching speech recognition
Xinyuan Zhou, Emre Yılmaz, Yanhua Long, Yijie Li, and Haizhou Li. 2020 · 2020
Later among the works it cites.
Yueying Zhu, Xiaobing Zhou, Hongling Li, and Kunjie Dong. 2020 · 2020
Later among the works it cites.
Humor generation and detection in code-mixed hindi-english
Kaustubh Agarwal and Rhythm Narula. 2021 · 2021
Later among the works it cites.
Towards code-mixed hinglish dialogue generation
Vibhav Agarwal, Pooja Rao, and Dinesh Babu Jayagopi. 2021 · 2021
Later among the works it cites.
Char2subword: Extending the subword embedding space using robust character compositionality
Gustavo Aguilar, Bryan McCann, Tong Niu, Nazneen Rajani, Nitish Shirish Keskar, and Thamar Solorio. 2021 · 2021
Later among the works it cites.
Arabic code-switching speech recognition using monolingual data
Ahmed Ali, Shammur Absar Chowdhury, Amir Hussein, and Yasser Hifny. 2021 · 2021
Later among the works it cites.
Judithjeyafreedaandrew@ dravidianlangtech-eacl2021: offensive language detection for dravidian code-mixed youtube comments
Judith Jeyafreeda Andrew. 2021 · 2021
Later among the works it cites.
Iitp-mt at calcs2021: English to hinglish neural machine translation using unsupervised synthetic code-mixed parallel corpus
Ramakrishna Appicharla, Kamal Kumar Gupta, Asif Ekbal, and Pushpak Bhattacharyya. 2021 · 2021
Later among the works it cites.
Mucs@ dravidianlangtech-eacl2021: Cooli-code-mixing offensive language identification
Fazlourrahman Balouchzahi, BK Aparna, and HL Shashirekha. 2021 · 2021
Later among the works it cites.
La-saco: A study of learning approaches for sentiments analysis incode-mixing texts
Fazlourrahman Balouchzahi and HL Shashirekha. 2021 · 2021
Later among the works it cites.
Slam: A unified encoder for speech and language modeling via speech-text joint pre-training
Ankur Bapna, Yu-an Chung, Nan Wu, Anmol Gulati, Ye Jia, Jonathan H Clark, Melvin Johnson, Jason Riesa, Alexis Conneau, and Yu Zhang. 2021 · 2021
Later among the works it cites.
Ssncse_nlp@ dravidianlangtech-eacl2021: Offensive language identification on multilingual code mixing text
B Bharathi et al. 2021 · 2021
Later among the works it cites.
Challenges in annotating and parsing spoken, code-switched, frisian-dutch data
Anouck Braggaar and Rob van der Goot. 2021 · 2021
Later among the works it cites.
Intrinsic evaluation of language models for code-switching
Sik Feng Cheong, Hai Leong Chieu, and Jing Lim. 2021 · 2021
Later among the works it cites.
dhivya-hope-detection@ lt-edi-eacl2021: multilingual hope speech detection for code-mixed and transliterated texts
Dhivya Chinnappa. 2021 · 2021
Later among the works it cites.
Switch point biased self-training: Re-purposing pretrained models for code-switching
Parul Chopra, Sai Krishna Rallabandi, Alan W Black, and Khyathi Raghavi Chandu. 2021 · 2021
Later among the works it cites.
Towards one model to rule all: Multilingual strategy for dialectal code-switching arabic asr
Shammur Absar Chowdhury, Amir Hussein, Ahmed Abdelali, and Ahmed Ali. 2021 · 2021
Later among the works it cites.
Irnlp_daiict@ dravidianlangtech-eacl2021: offensive language identification in dravidian languages using tf-idf char n-grams and muril
Bhargav Dave, Shripad Bhat, and Prasenjit Majumder. 2021 · 2021
Later among the works it cites.
Mucs 2021: Multilingual and code-switching asr challenges for low resource indian languages
Anuj Diwan, Rakesh Vaideeswaran, Sanket Shah, Ankita Singh, Srinivasa Raghavan, Shreya Khare, Vinit Unni, Saurabh Vyas, Akash Rajpuria, Chiranjeevi Yarra, Ashish Mittal, Prasanta Kumar Ghosh, Preethi Jyothi, Kalika Bali, Vivek Seshadri, Sunayana Sitaram, Samarth Bharadwaj, Jai Nanavati, Raoul Nanavati, and Karthik Sankaranarayanan. 2021 · 2021
Later among the works it cites.
A survey of code-switching: Linguistic and social perspectives for language technologies
A Seza Doğruöz, Sunayana Sitaram, Barbara Bullock, and Almeida Jacqueline Toribio. 2021 · 2021
Later among the works it cites.
Investigating code-mixed modern standard arabic-egyptian to english machine translation
AbdelRahim Elmadany, Muhammad Abdul-Mageed, et al. 2021 · 2021
Later among the works it cites.
Mipe: A metric independent pipeline for effective code-mixed nlg evaluation
Ayush Garg, Sammed Kagi, Vivek Srivastava, and Mayank Singh. 2021 · 2021
Later among the works it cites.
Training data augmentation for code-mixed translation
Abhirut Gupta, Aditya Vavre, and Sunita Sarawagi. 2021a · 2021
Later among the works it cites.
Nlp-cuet@ lt-edi-eacl2021: Multilingual code-mixed hope speech detection using cross-lingual representation learner
Eftekhar Hossain, Omar Sharif, and Mohammed Moshiul Hoque. 2021 · 2021
Later among the works it cites.
hub at semeval-2021 task 7: Fusion of albert and word frequency information detecting and rating humor and offense
Bo Huang and Yang Bai. 2021 · 2021
Later among the works it cites.
Much gracias: Semi-supervised code-switch detection for spanish-english: How far can we get?
Dana-Maria Iliescu, Rasmus Grand, Sara Qirko, and Rob van der Goot. 2021 · 2021
Later among the works it cites.
Exploring text-to-text transformers for english to hinglish machine translation with synthetic code-mixing
Ganesh Jawahar, Muhammad Abdul-Mageed, VS Laks Lakshmanan, et al. 2021 · 2021
Later among the works it cites.
Codemixednlp: An extensible and open nlp toolkit for code-mixing
Sai Muralidhar Jayanthi, Kavya Nerella, Khyathi Raghavi Chandu, and Alan W Black. 2021 · 2021
Later among the works it cites.
Towards developing a multilingual and code-mixed visual question answering system by knowledge distillation
Humair Raj Khan, Deepak Gupta, and Asif Ekbal. 2021 · 2021
Later among the works it cites.
The cstr system for multilingual and code-switching asr challenges for low resource indian languages
Ondřej Klejch, Electra Wallington, and Peter Bell. 2021 · 2021
Later among the works it cites.
Dual script e2e framework for multilingual and code-switching asr
Mari Ganesh Kumar, Jom Kuriakose, Anand Thyagachandran, Arun Kumar A, Ashish Seth, Lodagala V.S.V. Durga Prasad, Saish Jaiswal, Anusha Prakash, and Hema A. Murthy. 2021 · 2021
Later among the works it cites.
Codewithzichao@ dravidianlangtech-eacl2021: Exploring multimodal transformers for meme classification in tamil language
Zichao Li. 2021 · 2021
Later among the works it cites.
Exploiting low-resource code-switching data to mandarin-english speech recognition systems
Hou-An Lin and Chia-Ping Chen. 2021 · 2021
Later among the works it cites.
End-to-end language diarization for bilingual code-switching speech
Hexin Liu, Leibny Paola García Perera, Xinyi Zhang, Justin Dauwels, Andy W.H. Khong, Sanjeev Khudanpur, and Suzy J. Styles. 2021 · 2021
Later among the works it cites.
Towards phone number recognition for code switched algerian dialect
Khaled Lounnas, Mourad Abbas, and Mohamed Lichouri. 2021 · 2021
Later among the works it cites.
Ascend: A spontaneous chinese-english dataset for code-switching in multi-turn conversation
Holy Lovenia, Samuel Cahyawijaya, Genta Indra Winata, Peng Xu, Xu Yan, Zihan Liu, Rita Frieske, Tiezheng Yu, Wenliang Dai, Elham J Barezi, et al. 2021 · 2021
Later among the works it cites.
Sentiment classification of code-mixed tweets using bi-directional rnn and language tags
Sainik Mahata, Dipankar Das, and Sivaji Bandyopadhyay. 2021 · 2021
Later among the works it cites.
Sentiment analysis of dravidian code mixed data
Asrita Venkata Mandalam and Yashvardhan Sharma. 2021 · 2021
Later among the works it cites.
Gupshup: Summarizing open-domain code-switched conversations
Laiba Mehnaz, Debanjan Mahata, Rakesh Gosangi, Uma Sushmitha Gunturi, Riya Jain, Gauri Gupta, Amardeep Kumar, Isabelle G Lee, Anish Acharya, and Rajiv Shah. 2021 · 2021
Later among the works it cites.
A language-aware approach to code-switched morphological tagging
Şaziye Betül Özateş and Özlem Çetinoğlu. 2021 · 2021
Later among the works it cites.
Normalization and back-transliteration for code-switched data
Dwija Parikh and Thamar Solorio. 2021 · 2021
Later among the works it cites.
The effectiveness of intermediate-task training for code-switched natural language understanding
Archiki Prasad, Mohammad Ali Rehan, Shreya Pathak, and Preethi Jyothi. 2021 · 2021
Later among the works it cites.
Comparing grammatical theories of code-mixing
Adithya Pratapa and Monojit Choudhury. 2021 · 2021
Later among the works it cites.
“something something hota hai!” an explainable approach towards sentiment analysis on indian code-mixed data
Aman Priyanshu, Aleti Vardhan, Sudarshan Sivakumar, Supriti Vijay, and Nipuna Chhabra. 2021 · 2021
Later among the works it cites.
Dlrg@ dravidianlangtech-eacl2021: Transformer based approachfor offensive language identification on code-mixed tamil
Ratnavel Rajalakshmi, Yashwant Reddy, and Lokesh Kumar. 2021 · 2021
Later among the works it cites.
Dosa: Dravidian code-mixed offensive span identification dataset
Manikandan Ravikiran and Subbiah Annamalai. 2021 · 2021
Later among the works it cites.
Xtreme-r: Towards more challenging and nuanced multilingual evaluation
Sebastian Ruder, Noah Constant, Jan Botha, Aditya Siddhant, Orhan Firat, Jinlan Fu, Pengfei Liu, Junjie Hu, Dan Garrette, Graham Neubig, et al. 2021 · 2021
Later among the works it cites.
Bertologicomix: How does code-mixing interact with multilingual bert?
Sebastin Santy, Anirudh Srinivasan, and Monojit Choudhury. 2021 · 2021
Later among the works it cites.
Offensive language identification in dravidian code mixed social media text
Sunil Saumya, Abhinav Kumar, and Jyoti Prakash Singh. 2021 · 2021
Later among the works it cites.
Nlp-cuet@ dravidianlangtech-eacl2021: Offensive language detection from multilingual code-mixed text using transformers
Omar Sharif, Eftekhar Hossain, and Mohammed Moshiul Hoque. 2021 · 2021
Later among the works it cites.
Political discourse analysis: a case study of code mixing and code switching in political speeches
Dama Sravani, Lalitha Kameswari, and Radhika Mamidi. 2021 · 2021
Later among the works it cites.
Transliteration for low-resource code-switching texts: Building an automatic cyrillic-to-latin converter for tatar
Chihiro Taguchi, Yusuke Sakai, and Taro Watanabe. 2021 · 2021
Later among the works it cites.
From machine translation to code-switching: Generating high-quality code-switched text
Ishan Tarunesh, Syamantak Kumar, and Preethi Jyothi. 2021 · 2021
Later among the works it cites.
Hypers@ dravidianlangtech-eacl2021: Offensive language identification in dravidian code-mixed youtube comments and posts
Charangan Vasantharajan and Uthayasanker Thayasivam. 2021 · 2021
Later among the works it cites.
Towards emotion recognition in hindi-english code-mixed data: A transformer based approach
Anshul Wadhawan and Akshita Aggarwal. 2021 · 2021
Later among the works it cites.
Training hybrid models on noisy transliterated transcripts for code-switched speech recognition
Matthew Wiesner, Mousmita Sarma, Ashish Arora, Desh Raj, Dongji Gao, Ruizhe Huang, Supreet Preet, Moris Johnson, Zikra Iqbal, Nagendra Goel, Jan Trmal, Leibny Paola García Perera, and Sanjeev Khudanpur. 2021 · 2021
Later among the works it cites.
Can you traducir this? machine translation for code-switched input
Jitao Xu and François Yvon. 2021 · 2021
Later among the works it cites.
A network science approach to bilingual code-switching
Qihui Xu, Magdalena Markowska, Martin Chodorow, and Ping Li. 2021 · 2021
Later among the works it cites.
End-to-end spelling correction conditioned on acoustic feature for code-switching speech recognition
Shuai Zhang, Jiangyan Yi, Zhengkun Tian, Ye Bai, Jianhua Tao, Xuefei Liu, and Zhengqi Wen. 2021a · 2021
Later among the works it cites.
Cross-lingual aspect-based sentiment analysis with aspect term code-switching
Wenxuan Zhang, Ruidan He, Haiyun Peng, Lidong Bing, and Wai Lam. 2021b · 2021
Later among the works it cites.
Indorobusta: Towards robustness against diverse code-mixed indonesian local languages
Muhammad Farid Adilazuarda, Samuel Cahyawijaya, Genta Indra Winata, Pascale Fung, and Ayu Purwarianti. 2022 · 2022
Closest in time.
Basco: An annotated basque-spanish code-switching corpus for natural language understanding
Maia Aguirre, Laura García-Sardiña, Manex Serras, Ariane Méndez, and Jacobo López. 2022 · 2022
Closest in time.
One country, 700+ languages: Nlp challenges for underrepresented languages and dialects in indonesia
Alham Aji, Genta Indra Winata, Fajri Koto, Samuel Cahyawijaya, Ade Romadhony, Rahmad Mahendra, Kemal Kurniawan, David Moeljadi, Radityo Eko Prasojo, Timothy Baldwin, et al. 2022 · 2022
Closest in time.
Few-shot cross-lingual transfer for coarse-grained de-identification of code-mixed clinical texts
Saadullah Amin, Noon Pokaratsiri Goldstein, Morgan Wixted, Alejandro Garcia-Rudolph, Catalina Martínez-Costa, and Günter Neumann. 2022 · 2022
Closest in time.
An improved transformer transducer architecture for hindi-english code switched speech recognition
Ansen Antony, Sumanth Reddy Kota, Akhilesh Lade, Spoorthy V, and Shashidhar G. Koolagudi. 2022 · 2022
Closest in time.
mslam: Massively multilingual joint pre-training for speech and text
Ankur Bapna, Colin Cherry, Yu Zhang, Ye Jia, Melvin Johnson, Yong Cheng, Simran Khanuja, Jason Riesa, and Alexis Conneau. 2022 · 2022
Closest in time.
Iiitdwd@ tamilnlp-acl2022: Transformer-based approach to classify abusive content in dravidian code-mixed text
Shankar Biradar and Sunil Saumya. 2022 · 2022
Closest in time.
Cmnerone at semeval-2022 task 11: Code-mixed named entity recognition by leveraging multilingual data
Suman Dowlagar and Radhika Mamidi. 2022 · 2022
Closest in time.
Word-level language identification using subword embeddings for code-mixed bangla-english social media data
Aparna Dutta. 2022 · 2022
Closest in time.
Um6p-cs at semeval-2022 task 11: Enhancing multilingual and code-mixed complex named entity recognition via pseudo labels using multilingual transformer
Abdellah El Mekki, Abdelkader El Mahdaouy, Mohammed Akallouch, Ismail Berrada, and Ahmed Khoumsi. 2022 · 2022
Closest in time.
Leveraging sub label dependencies in code mixed indian languages for part-of-speech tagging using conditional random fields
Akash Kumar Gautam. 2022 · 2022
Closest in time.
Benchmarking evaluation metrics for code-switching automatic speech recognition
Injy Hamed, Amir Hussein, Oumnia Chellah, Shammur Chowdhury, Hamdy Mubarak, Sunayana Sitaram, Nizar Habash, and Ahmed Ali. 2022 · 2022
Closest in time.
Corpus creation for sentiment analysis in code-mixed tulu text
Asha Hegde, Mudoor Devadas Anusha, Sharal Coelho, Hosahalli Lakshmaiah Shashirekha, and Bharathi Raja Chakravarthi. 2022 · 2022
Closest in time.
Coswid, a code switching identification method suitable for under-resourced languages
Laurent Kevers. 2022 · 2022
Closest in time.
Talcs: An open-source mandarin-english code-switching corpus and a speech recognition baseline
Chengfei Li, Shuhao Deng, Yaoping Wang, Guangjing Wang, Yaguang Gong, Changbin Chen, and Jinfeng Bai. 2022 · 2022
Closest in time.
Ascend: A spontaneous chinese-english dataset for code-switching in multi-turn conversation
Holy Lovenia, Samuel Cahyawijaya, Genta Winata, Peng Xu, Yan Xu, Zihan Liu, Rita Frieske, Tiezheng Yu, Wenliang Dai, Elham J Barezi, et al. 2022 · 2022
Closest in time.
Semeval-2022 task 11: Multilingual complex named entity recognition (multiconer)
Shervin Malmasi, Anjie Fang, Besnik Fetahu, Sudipta Kar, and Oleg Rokhlenko. 2022b · 2022
Closest in time.
Normalization of code-switched text for speech synthesis
Sreeram Manghat, Sreeja Manghat, and Tanja Schultz. 2022 · 2022
Closest in time.
Borrowing or codeswitching? annotating for finer-grained distinctions in language mixing
Elena Álvarez Mellado and Constantine Lignos. 2022 · 2022
Closest in time.
Crosslingual generalization through multitask finetuning
Niklas Muennighoff, Thomas Wang, Lintang Sutawika, Adam Roberts, Stella Biderman, Teven Le Scao, M Saiful Bari, Sheng Shen, Zheng-Xin Yong, Hailey Schoelkopf, et al. 2022 · 2022
Closest in time.
Ksc2: An industrial-scale open-source kazakh speech corpus
Saida Mussakhojayeva, Yerbolat Khassanov, and Huseyin Atakan Varol. 2022b · 2022
Closest in time.
L3cube-hingcorpus and hingbert: A code mixed hindi-english dataset and bert language models
Ravindra Nayak and Raviraj Joshi. 2022 · 2022
Closest in time.
Speaker information can guide models to better inductive biases: A case study on predicting code-switching
Alissa Ostapenko, Shuly Wintner, Melinda Fricke, and Yulia Tsvetkov. 2022 · 2022
Closest in time.
Intonation in advice-giving in kenyan english and kiswahili
Billian Khalayi Otundo and Martine Grice. 2022 · 2022
Closest in time.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. 2022 · 2022
Closest in time.
Improving code-switching dependency parsing with semi-supervised auxiliary tasks
Şaziye Özateş, Arzucan Özgür, Tunga Güngör, and Özlem Çetinoğlu. 2022 · 2022
Closest in time.
Robust speech recognition via large-scale weak supervision
Alec Radford, Jong Wook Kim, Tao Xu, Greg Brockman, Christine McLeavey, and Ilya Sutskever. 2022 · 2022
Closest in time.
Mhe: Code-mixed corpora for similar language identification
Priya Rani, John Philip McCrae, and Theodorus Fransen. 2022 · 2022
Closest in time.
Zero-shot code-mixed offensive span identification through rationale extraction
Manikandan Ravikiran and Bharathi Raja Chakravarthi. 2022 · 2022
Closest in time.
Findings of the shared task on offensive span identification fromcode-mixed tamil-english comments
Manikandan Ravikiran, Bharathi Raja Chakravarthi, S Sangeetha, Ratnavel Rajalakshmi, Sajeetha Thavareesan, Rahul Ponnusamy, Shankar Mahadevan, et al. 2022 · 2022
Closest in time.
Bloom: A 176b-parameter open-access multilingual language model
Teven Le Scao, Angela Fan, Christopher Akiki, Ellie Pavlick, Suzana Ilić, Daniel Hesslow, Roman Castagné, Alexandra Sasha Luccioni, François Yvon, Matthias Gallé, et al. 2022 · 2022
Closest in time.
An improved deliberation network with text pre-training for code-switching automatic speech recognition
Zhijie Shen and Wu Guo. 2022 · 2022
Closest in time.
Language-specific characteristic assistance for code-switching speech recognition
Tongtong Song, Qiang Xu, Meng Ge, Longbiao Wang, Hao Shi, Yongjie Lv, Yuqin Lin, and Jianwu Dang. 2022 · 2022
Closest in time.
Identifying emotions in code mixed hindi-english tweets
Sanket Sonu, Rejwanul Haque, Mohammed Hasanuzzaman, Paul Stynes, and Pramod Pathak. 2022 · 2022
Closest in time.
Sentiment analysis on code-switched dravidian languages with kernel based extreme learning machines
Mithun Kumar SR, Lov Kumar, and Aruna Malapati. 2022 · 2022
Closest in time.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao, Abu Awal Md Shoeb, Abubakar Abid, Adam Fisch, Adam R Brown, Adam Santoro, Aditya Gupta, Adrià Garriga-Alonso, et al. 2022 · 2022
Closest in time.
Pandas@ abusive comment detection in tamil code-mixed data using custom embeddings with labse
Krithika Swaminathan, K Divyasri, GL Gayathri, Thenmozhi Durairaj, and B Bharathi. 2022 · 2022
Closest in time.
Universal dependencies treebank for tatar: Incorporating intra-word code-switching information
Chihiro Taguchi, Sei Iwata, and Taro Watanabe. 2022 · 2022
Closest in time.
Lae: Language-aware encoder for monolingual and multilingual asr
Jinchuan Tian, Jianwei Yu, Chunlei Zhang, Yuexian Zou, and Dong Yu. 2022 · 2022
Closest in time.
Nunc profana tractemus. detecting code-switching in a large corpus of 16th century letters
Martin Volk, Lukas Fischer, Patricia Scheurer, Bernard Silvan Schroffenegger, Raphael Schwitter, Phillip Ströbel, and Benjamin Suter. 2022 · 2022
Closest in time.
Cross-lingual few-shot learning on unseen languages
Genta Winata, Shijie Wu, Mayank Kulkarni, Thamar Solorio, and Daniel Preoţiuc-Pietro. 2022 · 2022
Closest in time.
Identifying tension in holocaust survivors’ interview: Code-switching/code-mixing as cues
Xinyuan Xia, Lu Xiao, Kun Yang, and Yueyue Wang. 2022 · 2022
Closest in time.
Improving recognition of out-of-vocabulary words in e2e code-switching asr by fusing speech generation methods
Lingxuan Ye, Gaofeng Cheng, Runyan Yang, Zehui Yang, Sanli Tian, Pengyuan Zhang, and Yonghong Yan. 2022 · 2022
Closest in time.
reducing multilingual context confusion for end-to-end code-switching automatic speech recognition
Shuai Zhang, Jiangyan Yi, Zhengkun Tian, Jianhua Tao, Yu Ting Yeung, and Liqun Deng. 2022 · 2022
Closest in time.
Zheng-Xin Yong, Ruochen Zhang, Jessica Zosa Forde, Skyler Wang, Samuel Cahyawijaya, Holy Lovenia, Genta Indra Winata, Lintang Sutawika, Jan Christian Blaise Cruz, Long Phan, Yin Lin Tan, and Alham Fikri Aji. 2023 · 2023
Closest in time.
Mixed-lingual text analysis for polyglot tts synthesis
Beat Pfister and Harald Romsdorfer. 2003 · 2040
Closest in time.
Building a mixed-lingual neural tts system with only monolingual data
Liumeng Xue, Wei Song, Guanghui Xu, Lei Xie, and Zhizheng Wu. 2019 · 2064
Closest in time.
Tweettaglish: A dataset for investigating tagalog-english code-switching
Megan Herrera, Ankit Aich, and Natalie Parde. 2022 · 2097
Closest in time.