Fetching the paper…
Reading the bibliography…
Singlish, a Creole language rooted in English, is a key focus in linguistic research within multilingual and multicultural contexts.
Seame: a mandarin-english code-switching speech corpus in south-east asia
Dau-Cheng Lyu, Tien Ping Tan, Engsiong Chng, and Haizhou Li. 2010 · 1989
Earlier work this paper cites.
The nie corpus of spoken singapore english (niecsse)
David Deterding and Ee Ling Low. 2001 · 2001
Earlier work this paper cites.
Singapore English
David Deterding. 2007 · 2007
Earlier work this paper cites.
The development of a singapore english call resource
Wenda Chen, Y Tan, E Chng, and Haizhou Li. 2010 · 2010
Earlier work this paper cites.
Tone in singlish: Substrate features from sinitic and malay
Lisa Lim. 2011 · 2011
Earlier work this paper cites.
The kaldi speech recognition toolkit
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, et al. 2011 · 2011
Earlier work this paper cites.
An overview of speaker identification: Accuracy and robustness issues
Roberto Togneri and Daniel Pullella. 2011 · 2011
Earlier work this paper cites.
The anatomy of singlish: globalisation, multiculturalism and the construction of the ‘local’in singapore
Robbie BH Goh. 2016 · 2016
Earlier work this paper cites.
Joint ctc-attention based end-to-end speech recognition using multi-task learning
Suyoun Kim, Takaaki Hori, and Shinji Watanabe. 2017 · 2017
Earlier work this paper cites.
Sigmoid-weighted linear units for neural network function approximation in reinforcement learning
Stefan Elfwing, Eiji Uchibe, and Kenji Doya. 2018 · 2018
Earlier work this paper cites.
ESPnet: End-to-end speech processing toolkit
Shinji Watanabe, Takaaki Hori, Shigeki Karita, Tomoki Hayashi, Jiro Nishitoba, Yuya Unno, Nelson Enrique Yalta Soplin, Jahn Heymann, Matthew Wiesner, Nanxin Chen, Adithya Renduchintala, and Tsubasa Ochiai. 2018 · 2018
Earlier work this paper cites.
Building the singapore english national speech corpus
Jia Xin Koh, Aqilah Mislan, Kevin Khoo, Brian Ang, Wilson Ang, Charmaine Ng, and YY Tan. 2019 · 2019
Earlier work this paper cites.
Spontaneous speech elicitation for large speech corpus in multilingual singapore
Ying-Ying Tan. 2019 · 2019
Earlier work this paper cites.
Voxceleb enrichment for age and gender recognition
Khaled Hechmi, Trung Ngo Trong, Ville Hautamäki, and Tomi Kinnunen. 2021 · 2021
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
An analysis of colloquial singapore english lah and its interpretation across speech acts
Junwen Lee. 2022 · 2022
Cited alongside, same era.
Particle stacking in singlish–new data from the national speech corpus
Ashley Boo, Junwen Lee, and Ying-Ying Tan. 2023 · 2023
Cited alongside, same era.
Qwen-audio: Advancing universal audio understanding via unified large-scale audio-language models
Yunfei Chu, Jin Xu, Xiaohuan Zhou, Qian Yang, Shiliang Zhang, Zhijie Yan, Chang Zhou, and Jingren Zhou. 2023 · 2023
Cited alongside, same era.
Cosmic: Data efficient instruction-tuning for speech in-context learning
Meralion-audiollm: Technical report
Yingxu He, Zhuohan Liu, Shuo Sun, Bin Wang, Wenyu Zhang, Xunlong Zou, Nancy F Chen, and Ai Ti Aw. 2024 · 2024
Later among the works it cites.
Wavllm: Towards robust and adaptive speech large language model
Shujie Hu, Long Zhou, Shujie Liu, Sanyuan Chen, Hongkun Hao, Jing Pan, Xunying Liu, Jinyu Li, Sunit Sivasankaran, Linquan Liu, et al. 2024 · 2024
Later among the works it cites.
Wavchat: A survey of spoken dialogue models
Shengpeng Ji, Yifu Chen, Minghui Fang, Jialong Zuo, Jingyu Lu, Hanting Wang, Ziyue Jiang, Long Zhou, Shujie Liu, Xize Cheng, et al. 2024 · 2024
Later among the works it cites.
Paralinguistics-aware speech-empowered large language models for natural conversation
Heeseung Kim, Soonshin Seo, Kyeongseok Jeong, Ohsung Kwon, Soyoon Kim, Jungwhan Kim, Jaehong Lee, Eunwoo Song, Myungwoo Oh, Jung-Woo Ha, et al. 2024 · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jing Pan, Jian Wu, Yashesh Gaur, Sunit Sivasankaran, Zhuo Chen, Shujie Liu, and Jinyu Li. 2023 · 2023
Cited alongside, same era.
Robust speech recognition via large-scale weak supervision
Alec Radford, Jong Wook Kim, Tao Xu, Greg Brockman, Christine McLeavey, and Ilya Sutskever. 2023 · 2023
Cited alongside, same era.
Audiopalm: A large language model that can speak and listen
Paul K Rubenstein, Chulayuth Asawaroengchai, Duc Dung Nguyen, Ankur Bapna, Zalán Borsos, Félix de Chaumont Quitry, Peter Chen, Dalia El Badawy, Wei Han, Eugene Kharitonov, et al. 2023 · 2023
Cited alongside, same era.
Gemini: a family of highly capable multimodal models
Gemini Team, Rohan Anil, Sebastian Borgeaud, Yonghui Wu, Jean-Baptiste Alayrac, Jiahui Yu, Radu Soricut, Johan Schalkwyk, Andrew M Dai, Anja Hauth, et al. 2023 · 2023
Cited alongside, same era.
Instructive dialogue summarization with query aggregations
Bin Wang, Zhengyuan Liu, and Nancy Chen. 2023a · 2023
Cited alongside, same era.
Yunfei Chu, Jin Xu, Qian Yang, Haojie Wei, Xipin Wei, Zhifang Guo, Yichong Leng, Yuanjun Lv, Jinzheng He, Junyang Lin, et al. 2024 · 2024
Cited alongside, same era.
Recent advances in speech language models: A survey
Wenqian Cui, Dianzhi Yu, Xiaoqi Jiao, Ziqiao Meng, Guangyan Zhang, Qichao Wang, Yiwen Guo, and Irwin King. 2024 · 2024
Cited alongside, same era.
Moshi: a speech-text foundation model for real-time dialogue
Alexandre Défossez, Laurent Mazaré, Manu Orsini, Amélie Royer, Patrick Pérez, Hervé Jégou, Edouard Grave, and Neil Zeghidour. 2024 · 2024
Cited alongside, same era.
Yadong Li, Haoze Sun, Mingan Lin, Tianpeng Li, Guosheng Dong, Tao Zhang, Bowen Ding, Wei Song, Zhenglin Cheng, Yuqi Huo, et al. 2024 · 2024
Later among the works it cites.
Developing instruction-following speech language model without speech instruction-tuning data
Ke-Han Lu, Zhehuai Chen, Szu-Wei Fu, Chao-Han Huck Yang, Jagadeesh Balam, Boris Ginsburg, Yu-Chiang Frank Wang, and Hung-yi Lee. 2024 · 2024
Later among the works it cites.
Spirit-lm: Interleaved spoken and written language model
Tu Anh Nguyen, Benjamin Muller, Bokai Yu, Marta R Costa-Jussa, Maha Elbayad, Sravya Popuri, Paul-Ambroise Duquenne, Robin Algayres, Ruslan Mavlyutov, Itai Gat, et al. 2024 · 2024
Later among the works it cites.
A survey on speech large language models
Jing Peng, Yucheng Wang, Yu Xi, Xv Li, and Kai Yu. 2024 · 2024
Later among the works it cites.
SALMONN: Towards generic hearing abilities for large language models
Changli Tang, Wenyi Yu, Guangzhi Sun, Xianzhao Chen, Tian Tan, Wei Li, Lu Lu, Zejun MA, and Chao Zhang. 2024 · 2024
Later among the works it cites.
Gemma: Open models based on gemini research and technology
Gemma Team, Thomas Mesnard, Cassidy Hardin, Robert Dadashi, Surya Bhupatiraju, Shreya Pathak, Laurent Sifre, Morgane Rivière, Mihir Sanjay Kale, Juliette Love, et al. 2024 · 2024
Later among the works it cites.
Audiobench: A universal benchmark for audio large language models
Bin Wang, Xunlong Zou, Geyu Lin, Shuo Sun, Zhuohan Liu, Wenyu Zhang, Zhengyuan Liu, AiTi Aw, and Nancy F Chen. 2024 · 2024
Later among the works it cites.
Building a taiwanese mandarin spoken language model: A first attempt
Chih-Kai Yang, Yu-Kuan Fu, Chen-An Li, Yi-Cheng Lin, Yu-Xiang Lin, Wei-Chih Chen, Ho Lam Chung, Chun-Yi Kuan, Wei-Ping Huang, Ke-Han Lu, et al. 2024 · 2024
Later among the works it cites.
Mowe-audio: Multitask audiollms with mixture of weak encoders
Wenyu Zhang, Shuo Sun, Bin Wang, Xunlong Zou, Zhuohan Liu, Yingxu He, Geyu Lin, Nancy F Chen, and Ai Ti Aw. 2025 · 2025
Closest in time.