中文
相关论文

相关论文: AfriSenti: A Twitter Sentiment Analysis Benchmark …

200 篇论文

This paper describes the development of a multilingual, manually annotated dataset for three under-resourced Dravidian languages generated from social media comments. The dataset was annotated for sentiment analysis and offensive language…

Recent studies showed a huge interest in social networks sentiment analysis. Twitter, which is a microblogging service, can be a great source of information on how the users feel about a certain topic, or what their opinion is regarding a…

计算与语言 · 计算机科学 2020-10-01 Reda Khalaf , Mireille Makary

Sentiment Analysis, a popular subtask of Natural Language Processing, employs computational methods to extract sentiment, opinions, and other subjective aspects from linguistic data. Given its crucial role in understanding human sentiment,…

计算与语言 · 计算机科学 2025-02-07 Zhiqiang Shi , Ruchit Agrawal

It is important to be able to analyze the emotional state of people around the globe. There are 7100+ active languages spoken around the world and building emotion classification for each language is labor intensive. Particularly for…

计算与语言 · 计算机科学 2024-11-11 Shabnam Tafreshi , Shubham Vatsal , Mona Diab

Musicians frequently use social media to express their opinions, but they often convey different messages in their music compared to their posts online. Some utilize these platforms to abuse their colleagues, while others use it to show…

计算与语言 · 计算机科学 2024-11-12 Sunday Oluyele , Juwon Akingbade , Victor Akinode

Social media sites such as YouTube and Facebook have become an integral part of everyone's life and in the last few years, hate speech in the social media comment section has increased rapidly. Detection of hate speech on social media…

计算与语言 · 计算机科学 2020-12-18 Nauros Romim , Mosahed Ahmed , Hriteshwar Talukder , Md Saiful Islam

Large language models and vision-language models (which we jointly call LMs) have transformed NLP and CV, demonstrating remarkable potential across various fields. However, their capabilities in affective analysis (i.e. sentiment analysis…

计算与语言 · 计算机科学 2025-06-02 Zhiwei Liu , Lingfei Qian , Qianqian Xie , Jimin Huang , Kailai Yang , Sophia Ananiadou

Due to the breathtaking growth of social media or newspaper user comments, online product reviews comments, sentiment analysis (SA) has captured substantial interest from the researchers. With the fast increase of domain, SA work aims not…

计算与语言 · 计算机科学 2020-12-02 Mahfuz Ahmed Masum , Sheikh Junayed Ahmed , Ayesha Tasnim , Md Saiful Islam

We present the first self-supervised multilingual speech model trained exclusively on African speech. The model learned from nearly 60 000 hours of unlabeled speech segments in 21 languages and dialects spoken in sub-Saharan Africa. On the…

计算与语言 · 计算机科学 2024-04-23 Antoine Caubrière , Elodie Gauthier

Automatic speech recognition (ASR) for African languages remains constrained by limited labeled data and the lack of systematic guidance on model selection, data scaling, and decoding strategies. Large pre-trained systems such as Whisper,…

Although, the fair amount of works in sentiment analysis (SA) and opinion mining (OM) systems in the last decade and with respect to the performance of these systems, but it still not desired performance, especially for morphologically-Rich…

计算与语言 · 计算机科学 2015-06-08 Hossam S. Ibrahim , Sherif M. Abdou , Mervat Gheith

Hate speech is a growing problem on social media. It can seriously impact society, especially in countries like Ethiopia, where it can trigger conflicts among diverse ethnic and religious groups. While hate speech detection in resource rich…

计算与语言 · 计算机科学 2024-08-08 Samuel Minale Gashe , Seid Muhie Yimam , Yaregal Assabie

Arabic is one of the oldest languages still in use today. As a result, several Arabic-speaking regions have developed dialects that are unique to them. Dialect and emotion recognition have various uses in Arabic text analysis, such as…

计算与语言 · 计算机科学 2025-02-14 Nasser A Alsadhan

We introduce a new reading comprehension dataset, dubbed MultiWikiQA, which covers 306 languages and has 1,220,757 samples in total. We start with Wikipedia articles, which also provide the context for the dataset samples, and use an LLM to…

计算与语言 · 计算机科学 2026-03-05 Dan Saattrup Smart

Tunisians on social media tend to express themselves in their local dialect using Latin script (TUNIZI). This raises an additional challenge to the process of exploring and recognizing online opinions. To date, very little work has…

计算与语言 · 计算机科学 2020-10-15 Abir Messaoudi , Hatem Haddad , Moez Ben HajHmida , Chayma Fourati , Abderrazak Ben Hamida

Since state-of-the-art approaches to offensive language detection rely on supervised learning, it is crucial to quickly adapt them to the continuously evolving scenario of social media. While several approaches have been proposed to tackle…

计算与语言 · 计算机科学 2022-10-17 Elisa Leonardelli , Stefano Menini , Alessio Palmero Aprosio , Marco Guerini , Sara Tonelli

Arabic poetry, with its rich linguistic features and profound cultural significance, presents a unique challenge to the Natural Language Processing (NLP) field. The complexity of its structure and context necessitates advanced computational…

计算与语言 · 计算机科学 2024-03-20 Faisal Qarah

Large Language Models (LLMs) have shown remarkable capabilities, not only in generating human-like text, but also in acquiring knowledge. This highlights the need to go beyond the typical Natural Language Processing downstream benchmarks…

When building NLP models, there is a tendency to aim for broader coverage, often overlooking cultural and (socio)linguistic nuance. In this position paper, we make the case for care and attention to such nuances, particularly in dataset…

计算与语言 · 计算机科学 2022-03-21 A. Stevie Bergman , Mona T. Diab

Africa's rich linguistic diversity remains significantly underrepresented in speech technologies, creating barriers to digital inclusion. To alleviate this challenge, we systematically map the continent's speech space of datasets and…