English
Related papers

Related papers: CroSentiNews 2.0: A Sentence-Level News Sentiment …

200 papers

Emotion classification is often formulated as the task to categorize texts into a predefined set of emotion classes. So far, this task has been the recognition of the emotion of writers and readers, as well as that of entities mentioned in…

Computation and Language · Computer Science 2022-04-12 Enrica Troiano , Laura Oberländer , Maximilian Wegge , Roman Klinger

The Facebook network allows its users to record their reactions to text via a typology of emotions. This network, taken at scale, is therefore a prime data set of annotated sentiment data. This paper uses millions of such reactions, derived…

Machine Learning · Computer Science 2022-08-04 Vihanga Jayawickrama , Gihan Weeraprameshwara , Nisansa de Silva , Yudhanjaya Wijeratne

Sentiment analysis is an essential technique for investigating the emotional climate within developer teams, contributing to both team productivity and project success. Existing sentiment analysis tools in software engineering primarily…

Software Engineering · Computer Science 2025-07-11 Martin Obaidi , Marc Herrmann , Elisa Schmid , Raymond Ochsner , Kurt Schneider , Jil Klünder

Rapid increase in internet users along with growing power of online review sites and social media has given birth to sentiment analysis or opinion mining, which aims at determining what other people think and comment. Sentiments or Opinions…

Computation and Language · Computer Science 2016-07-12 Aurangzeb khan , Khairullah khan , Shakeel Ahmad , Fazal Masood Kundi , Irum Tareen , Muhammad Zubair Asghar

We introduce \textsc{PoliteRewrite} -- a dataset for polite language rewrite which is a novel sentence rewrite task. Compared with previous text style transfer tasks that can be mostly addressed by slight token- or phrase-level edits,…

Computation and Language · Computer Science 2022-12-21 Xun Wang , Tao Ge , Allen Mao , Yuki Li , Furu Wei , Si-Qing Chen

The evolution of the Internet has increased the amount of information that is expressed by people on different platforms. This information can be product reviews, discussions on forums, or social media platforms. Accessibility of these…

Computation and Language · Computer Science 2021-04-20 Gati L. Martin , Medard E. Mswahili , Young-Seob Jeong

Informational bias is widely present in news articles. It refers to providing one-sided, selective or suggestive information of specific aspects of certain entity to guide a specific interpretation, thereby biasing the reader's opinion.…

Computation and Language · Computer Science 2022-01-26 Shijia Guo , Kenny Q. Zhu

This paper fills a gap in aspect-based sentiment analysis and aims to present a new method for preparing and analysing texts concerning opinion and generating user-friendly descriptive reports in natural language. We present a comprehensive…

Computation and Language · Computer Science 2017-09-15 Łukasz Augustyniak , Krzysztof Rajda , Tomasz Kajdanowicz

Sentiment classification is widely used for product reviews and in online social media such as forums, Twitter, and blogs. However, the problem of classifying the sentiment of user comments on news sites has not been addressed yet. News…

Computation and Language · Computer Science 2015-06-12 Prakhar Biyani , Cornelia Caragea , Narayan Bhamidipati

We present PubMed 200k RCT, a new dataset based on PubMed for sequential sentence classification. The dataset consists of approximately 200,000 abstracts of randomized controlled trials, totaling 2.3 million sentences. Each sentence of each…

Computation and Language · Computer Science 2017-10-18 Franck Dernoncourt , Ji Young Lee

This paper focuses on sentiment mining and sentiment correlation analysis of web events. Although neural network models have contributed a lot to mining text information, little attention is paid to analysis of the inter-sentiment…

Computation and Language · Computer Science 2018-11-27 Xinzhi Wang , Shengcheng Yuan , Hui Zhang , Yi Liu

The growth of deep learning (DL) relies heavily on huge amounts of labelled data for tasks such as natural language processing and computer vision. Specifically, in image-to-text or image-to-image pipelines, opinion (sentiment) may be…

Computer Vision and Pattern Recognition · Computer Science 2024-09-17 Aleksei Krotov , Alison Tebo , Dylan K. Picart , Aaron Dean Algave

Sentiment Analysis of code-mixed text has diversified applications in opinion mining ranging from tagging user reviews to identifying social or political sentiments of a sub-population. In this paper, we present an ensemble architecture of…

Computation and Language · Computer Science 2020-07-23 Ayush Kumar , Harsh Agarwal , Keshav Bansal , Ashutosh Modi

ParlaSpeech is a collection of spoken parliamentary corpora currently spanning four Slavic languages - Croatian, Czech, Polish and Serbian - all together 6 thousand hours in size. The corpora were built in an automatic fashion from the…

Computation and Language · Computer Science 2026-04-16 Nikola Ljubešić , Peter Rupnik , Ivan Porupski , Taja Kuzman Pungeršek

This paper provides a detailed description of a new Twitter-based benchmark dataset for Arabic Sentiment Analysis (ASAD), which is launched in a competition3, sponsored by KAUST for awarding 10000 USD, 5000 USD and 2000 USD to the first,…

Computation and Language · Computer Science 2021-03-11 Basma Alharbi , Hind Alamro , Manal Alshehri , Zuhair Khayyat , Manal Kalkatawi , Inji Ibrahim Jaber , Xiangliang Zhang

Researchers commonly perform sentiment analysis on large collections of short texts like tweets, Reddit posts or newspaper headlines that are all focused on a specific topic, theme or event. Usually, general-purpose sentiment analysis…

Computation and Language · Computer Science 2024-07-11 James C. Young , Rudy Arthur , Hywel T. P. Williams

People are sharing their opinions, stories and reviews through online video sharing websites every day. Studying sentiment and subjectivity in these opinion videos is experiencing a growing attention from academia and industry. While…

Computation and Language · Computer Science 2016-11-18 Amir Zadeh , Rowan Zellers , Eli Pincus , Louis-Philippe Morency

We present a corpus of 100 documents, OBSINFOX, selected from 17 sources of French press considered unreliable by expert agencies, annotated using 11 labels by 8 annotators. By collecting more labels than usual, by more annotators than is…

There are many general purpose benchmark datasets for Semantic Textual Similarity but none of them are focused on technical concepts found in patents and scientific publications. This work aims to fill this gap by presenting a new human…

Computation and Language · Computer Science 2022-08-03 Grigor Aslanyan , Ian Wetherbee

Toxic comments in online platforms are an unavoidable social issue under the cloak of anonymity. Hate speech detection has been actively done for languages such as English, German, or Italian, where manually labeled corpus has been…

Computation and Language · Computer Science 2020-05-27 Jihyung Moon , Won Ik Cho , Junbum Lee