Fetching the paper…
Reading the bibliography…
In this demo paper, we introduce SAPIEN, a platform for high-fidelity virtual agents driven by large language models that can hold open domain conversations with users in 13 different languages, and display emotions through facial expressions and voice.
M. E. Hoque, M. Courgeon, J.-C. Martin, B. Mutlu, and R. W. Picard, “Mach: My automated conversation coach,” in Proceedings of the 2013 ACM International Joint Conference on Pervasive and Ubiquitous Computing , ser. UbiComp ’13. New York, NY, USA: Association for Computing Machinery, 2013, p. 697–706. [Online]. Available: https://doi.org/10.1145/2493432.2493502
2013
Earlier work this paper cites.
M. E. Hoque and R. W. Picard, “Rich nonverbal sensing technology for automated social skills training,” Computer , vol. 47, no. 4, pp. 28–35, 2014
2014
Earlier work this paper cites.
M. Fung, Y. Jin, R. Zhao, and M. E. Hoque, “Roc speak: Semi-automated personalized feedback on nonverbal behavior from recorded videos,” in Proceedings of the 2015 ACM International Joint Conference on Pervasive and Ubiquitous Computing , ser. UbiComp ’15. New York, NY, USA: Association for Computing Machinery, 2015, p. 1167–1178. [Online]. Available: https://doi.org/10.1145/2750858.2804265
2015
Earlier work this paper cites.
M. R. Ali, D. Crasta, L. Jin, A. Baretto, J. Pachter, R. D. Rogge, and M. E. Hoque, “Lissa — live interactive social skill assistance,” in 2015 International Conference on Affective Computing and Intelligent Interaction (ACII) , 2015, pp. 173–179
2015
Earlier work this paper cites.
S. Z. Razavi, M. R. Ali, T. H. Smith, L. K. Schubert, and M. E. Hoque, “The lissa virtual human and asd teens: An overview of initial experiments,” in Intelligent Virtual Agents , D. Traum, W. Swartout, P. Khooshabeh, S. Kopp, S. Scherer, and A. Leuski, Eds. Cham: Springer International Publishing, 2016, pp. 460–463
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
W. Xiong, L. Wu, F. Alleva, J. Droppo, X. Huang, and A. Stolcke, “The microsoft 2017 conversational speech recognition system,” in 2018 IEEE international conference on acoustics, speech and signal processing (ICASSP) . IEEE, 2018, pp. 5934–5938
2018
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
M. R. Ali, S. Z. Razavi, R. Langevin, A. Al Mamun, B. Kane, R. Rawassizadeh, L. K. Schubert, and E. Hoque, “A virtual conversational agent for teens with autism spectrum disorder: Experimental results and design lessons,” in Proceedings of the 20th ACM International Conference on Intelligent Virtual Agents , ser. IVA ’20. New York, NY, USA: Association for Computing Machinery, 2020. [Online]. Available: https://doi.org/10.1145/3383652.3423900
2020
Earlier work this paper cites.
R. Luo, X. Tan, R. Wang, T. Qin, J. Li, S. Zhao, E. Chen, and T.-Y. Liu, “Lightspeech: Lightweight and fast text to speech with neural architecture search,” in ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2021, pp. 5699–5703
2021
Cited alongside, same era.
Y. Leng, X. Tan, L. Zhu, J. Xu, R. Luo, L. Liu, T. Qin, X. Li, E. Lin, and T.-Y. Liu, “Fastcorrect: Fast error correction with edit alignment for automatic speech recognition,” Advances in Neural Information Processing Systems , vol. 34, pp. 21 708–21 719, 2021
2021
Cited alongside, same era.
W. Hou, J. Wang, X. Tan, T. Qin, and T. Shinozaki, “Cross-domain speech recognition with unsupervised character-level distribution matching,” INTERSPEECH , 2021
2021
Cited alongside, same era.
M. R. Ali, T. Sen, B. Kane, S. Bose, T. M. Carroll, R. Epstein, L. Schubert, and E. Hoque, “Novel computational linguistic measures, dialogue system and the development of sophie: Standardized online patient for healthcare interaction education,” IEEE Trans. Affect. Comput. , vol. 14, no. 1, p. 223–235, jan 2023. [Online]. Available: https://doi.org/10.1109/TAFFC.2021.3054717
OpenAI, “Introducing chatgpt,” https://openai.com/blog/chatgpt, (Accessed on 06/22/2023)
2023
Closest in time.
“Anthropic — introducing claude,” https://www.anthropic.com/index/introducing-claude, (Accessed on 06/22/2023)
2023
Closest in time.
G. AI, “An important next step on our ai journey,” 2023. [Online]. Available: https://blog.google/technology/ai/bard-google-ai-search-updates/
2023
Closest in time.
R. Taori, I. Gulrajani, T. Zhang, Y. Dubois, X. Li, C. Guestrin, P. Liang, and T. B. Hashimoto, “Stanford alpaca: An instruction-following llama model,” https://github.com/tatsu-lab/stanford_alpaca, 2023
2023
Closest in time.
OpenAI, “Gpt-4 technical report,” 2023
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2021
Cited alongside, same era.
J. Li, “Recent advances in end-to-end automatic speech recognition,” APSIPA Transactions on Signal and Information Processing , April 2022. [Online]. Available: https://www.microsoft.com/en-us/research/publication/recent-advances-in-end-to-end-automatic-speech-recognition/
2022
Cited alongside, same era.
S.-g. Lee, H. Kim, C. Shin, X. Tan, C. Liu, Q. Meng, T. Qin, W. Chen, S. Yoon, and T.-Y. Liu, “Priorgrad: Improving conditional denoising diffusion models with data-driven adaptive prior,” ICLR , 2022
2022
Cited alongside, same era.
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray et al. , “Training language models to follow instructions with human feedback,” Advances in Neural Information Processing Systems , vol. 35, pp. 27 730–27 744, 2022
2022
Cited alongside, same era.
S. Z. Razavi, L. K. Schubert, K. van Orden, M. R. Ali, B. Kane, and E. Hoque, “Discourse behavior of older adults interacting with a dialogue agent competent in multiple topics,” ACM Trans. Interact. Intell. Syst. , vol. 12, no. 2, jul 2022. [Online]. Available: https://doi.org/10.1145/3484510
2022
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
W.-L. Chiang, Z. Li, Z. Lin, Y. Sheng, Z. Wu, H. Zhang, L. Zheng, S. Zhuang, Y. Zhuang, J. E. Gonzalez, I. Stoica, and E. P. Xing, “Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,” March 2023. [Online]. Available: https://lmsys.org/blog/2023-03-30-vicuna/
2023
Closest in time.