中文

Safe Guard:一种用于社交虚拟现实中基于语音的仇恨言论检测的 LLM 智能体

音频与语音处理 2024-09-25 v1 人工智能 声音

摘要

In this paper, we present Safe Guard, an LLM-agent for the detection of hate speech in voice-based interactions in social VR (VRChat). Our system leverages Open AI GPT and audio feature extraction for real-time voice interactions. We contribute a system design and evaluation of the system that demonstrates the capability of our approach in detecting hate speech, and reducing false positives compared to currently available approaches. Our results indicate the potential of LLM-based agents in creating safer virtual environments and set the groundwork for further advancements in LLM-driven moderation approaches.

关键词

引用

@article{arxiv.2409.15623,
  title  = {Safe Guard: an LLM-agent for Real-time Voice-based Hate Speech Detection in Social Virtual Reality},
  author = {Yiwen Xu and Qinyang Hou and Hongyu Wan and Mirjana Prpa},
  journal= {arXiv preprint arXiv:2409.15623},
  year   = {2024}
}