Fetching the paper…
Reading the bibliography…
In this paper, we present Safe Guard, an LLM-agent for the detection of hate speech in voice-based interactions in social VR (VRChat).
The psychology of profanity
G. T. W. Patrick · 1901
Earlier work this paper cites.
Gmm supervector based svm with spectral features for speech emotion recognition
H. Hu, M.-X. Xu, and W. Wu · 2007
Earlier work this paper cites.
Speech emotion recognition using both spectral and prosodic features
Y. Zhou, Y. Sun, J. Zhang, and Y. Yan · 2009
Earlier work this paper cites.
Detecting offensive user video blogs: An adaptive keyword spotting approach
M. S. Barakat, C. H. Ritz, and D. A. Stirling · 2012
Earlier work this paper cites.
Hateful symbols or hateful people? predictive features for hate speech detection on Twitter
Z. Waseem and D. Hovy · 2013
Earlier work this paper cites.
Mel frequency cepstral coefficients (mfcc) feature extraction enhancement in the application of speech recognition: A comparison study
S. Majeed, H. HUSAIN, S. Samad, and T. Idbeaa · 2015
Earlier work this paper cites.
Should we design for control, trust or involvement? a discourses survey about children’s online safety
H. Hartikainen, N. Iivari, and M. Kinnula · 2016
Earlier work this paper cites.
Characterizations of online harassment: Comparing policies across social media platforms
J. A. Pater, M. K. Kim, E. D. Mynatt, and C. Fiesler · 2016
Earlier work this paper cites.
Classification and its consequences for online harassment: Design insights from heartmob
L. Blackwell, J. Dimond, S. Schoenebeck, and C. Lampe · 2017
Earlier work this paper cites.
Automated hate speech detection and the problem of offensive language, 2017
T. Davidson, D. Warmsley, M. Macy, and I. Weber · 2017
Earlier work this paper cites.
Online harassment 2017, 2017
M. Duggan · 2017
Earlier work this paper cites.
Emotion recognition using prosodie and spectral features of speech and naïve bayes classifier
A. Khan and U. K. Roy · 2017
Earlier work this paper cites.
Large scale crowdsourcing and characterization of twitter abusive behavior, 2018
A.-M. Founta, C. Djouvas, D. Chatzakou, I. Leontiadis, J. Blackburn, G. Stringhini, A. Vakali, M. Sirivianos, and N. Kourtellis · 2018
Earlier work this paper cites.
All you need is ”love”: Evading hate speech detection
T. Gröndahl, L. Pajola, M. Juuti, M. Conti, and N. Asokan · 2018
Earlier work this paper cites.
An evaluation of preprocessing techniques for text classification
A. Kadhim · 2018
Earlier work this paper cites.
Harassment in social virtual reality: Challenges for platform governance
L. Blackwell, N. Ellison, N. Elliott-Deflo, and R. Schwartz · 2019
Earlier work this paper cites.
Moderation challenges in voice-based online communities on discord
J. A. Jiang, C. Kiene, S. Middler, J. R. Brubaker, and C. Fiesler · 2019
Earlier work this paper cites.
Overview of the hasoc track at fire 2019: Hate speech and offensive content identification in indo-european languages
T. Mandl, S. Modha, P. Majumder, D. Patel, M. Dave, C. Mandlia, and A. Patel · 2019
Earlier work this paper cites.
Mel frequency cepstral coefficient: A review
S. Ali, S. Tanweer, S. Khalid, and N. Rao · 2020
Cited alongside, same era.
Language agnostic hate speech detection
A. Arango · 2020
Cited alongside, same era.
It is complicated: Interacting with children in social virtual reality
D. Maloney, G. Freeman, and A. Robb · 2020
Cited alongside, same era.
How to motivate your dragon: Teaching goal-driven agents to speak and act in fantasy worlds
P. Ammanabrolu, J. Urbanek, M. Li, A. Szlam, T. Rocktäschel, and J. Weston · 2021
Cited alongside, same era.
Overview of the hasoc track at fire 2020: Hate speech and offensive language identification in tamil, malayalam, hindi, english and german
T. Mandl, S. Modha, A. Kumar M, and B. R. Chakravarthi · 2021
Cited alongside, same era.
Fine-tuning gpt-2 on annotated rpg quests for npc dialogue generation
Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing
P. Liu, W. Yuan, J. Fu, Z. Jiang, H. Hayashi, and G. Neubig · 2023
Later among the works it cites.
User reviews as a reporting mechanism for emergent issues within social vr communities
J. O’Hagan, F. Mathis, and M. McGill · 2023
Later among the works it cites.
Generative agents: Interactive simulacra of human behavior, 2023
J. S. Park, J. C. O’Brien, C. J. Cai, M. R. Morris, P. Liang, and M. S. Bernstein · 2023
Later among the works it cites.
Challenges of moderating social virtual reality
N. Sabri, B. Chen, A. Teoh, S. P. Dow, K. Vaccaro, and M. Elsherief · 2023
Later among the works it cites.
Towards leveraging ai-based moderation to address emergent harassment in social virtual reality
K. Schulenberg, L. Li, G. Freeman, S. Zamanifard, and N. J. McNeese · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. van Stegeren and J. Myśliwiec · 2021
Cited alongside, same era.
Disturbing the peace: Experiencing and mitigating emerging harassment in social virtual reality
G. Freeman, S. Zamanifard, D. Maloney, and D. Acena · 2022
Cited alongside, same era.
Human-ai collaboration via conditional delegation: A case study of content moderation
V. Lai, S. Carton, R. Bhatnagar, Q. V. Liao, Y. Zhang, and C. Tan · 2022
Cited alongside, same era.
Exploring the viability of conversational ai for non-playable characters: A comprehensive survey
A. Mehta, Y. Kunjadiya, A. Kulkarni, and M. Nagar · 2022
Cited alongside, same era.
Emotion based hate speech detection using multimodal learning
A. Rana and S. Jha · 2022
Cited alongside, same era.
Personalized quest and dialogue generation in role-playing games: A knowledge graph- and language model-based approach
T. Ashby, B. K. Webb, G. Knapp, J. Searle, and N. Fulda · 2023
Cited alongside, same era.
The effect of context-aware llm-based npc conversations on player engagement in role-playing video games
L. M. Csepregi · 2023
Cited alongside, same era.
Q. Zheng, S. Xu, L. Wang, Y. Tang, R. C. Salvi, G. Freeman, and Y. Huang · 2023
Later among the works it cites.
Deploying ml for voice safety
K. Bhat · 2024
Closest in time.
”pikachu would electrocute people who are misbehaving”: Expert, guardian and child perspectives on automated embodied moderators for safeguarding children in social virtual reality
C. Fiani, R. Bretin, S. A. Macdonald, M. Khamis, and M. Mcgill · 2024
Closest in time.
Exploring the perspectives of social vr-aware non-parent adults and parents on children’s use of social virtual reality
C. Fiani, P. Saeghe, M. McGill, and M. Khamis · 2024
Closest in time.
An investigation of large language models for real-world hate speech detection, 2024
K. Guo, A. Hu, J. Mu, Z. Shi, Z. Zhao, N. Vishwamitra, and H. Hu · 2024
Closest in time.
Llm-mod: Can large language models assist content moderation?
M. Kolla, S. Salunkhe, E. Chandrasekharan, and K. Saha · 2024
Closest in time.
Watch your language: Investigating content moderation with large language models
D. Kumar, Y. AbuHashem, and Z. Durumeric · 2024
Closest in time.
“hot” chatgpt: The promise of chatgpt in detecting and discriminating hateful, offensive, and toxic comments on social media
L. Li, L. Fan, S. Atreja, and L. Hemphill · 2024
Closest in time.
What is hate speech by united nation, 2024
U. Nation · 2024
Closest in time.
Enabling contextual soft moderation on social media through contrastive textual deviation
P. Paudel, M. Saeed, R. Auger, C. Wells, and G. Stringhini · 2024
Closest in time.
What are the cons of content moderation? Explore the Drawbacks
M. Rahaman · 2024
Closest in time.
Building llm-based ai agents in social virtual reality
H. Wan, J. Zhang, A. A. Suria, B. Yao, D. Wang, Y. Coady, and M. Prpa · 2024
Closest in time.