English
Related papers

Related papers: NutCracker at WNUT-2020 Task 2: Robustly Identifyi…

200 papers

X (formerly Twitter) has evolved into a contemporary agora, offering a platform for individuals to express opinions and viewpoints on current events. The majority of the topics discussed on Twitter are directly related to ongoing events,…

Recent studies on domain-specific BERT models show that effectiveness on downstream tasks can be improved when models are pretrained on in-domain data. Often, the pretraining data used in these models are selected based on their subject…

Computation and Language · Computer Science 2020-10-06 Xiang Dai , Sarvnaz Karimi , Ben Hachey , Cecile Paris

Since the classification of COVID-19 as a global pandemic, there have been many attempts to treat and contain the virus. Although there is no specific antiviral treatment recommended for COVID-19, there are several drugs that can…

Information Retrieval · Computer Science 2020-10-12 Ramya Tekumalla , Juan M. Banda

This paper introduces a study on tweet sentiment classification. Our task is to classify a tweet as either positive or negative. We approach the problem in two steps, namely embedding and classifying. Our baseline methods include several…

Computation and Language · Computer Science 2021-10-01 Tommaso Macrì , Freya Murphy , Yunfan Zou , Yves Zumbach

Toxic comment detection on social media has proven to be essential for content moderation. This paper compares a wide set of different models on a highly skewed multi-label hate speech dataset. We consider inference time and several metrics…

Computation and Language · Computer Science 2023-01-27 Corentin Duchene , Henri Jamet , Pierre Guillaume , Reda Dehak

Supervised learning algorithms are heavily reliant on annotated datasets to train machine learning models. However, the curation of the annotated datasets is laborious and time consuming due to the manual effort involved and has become a…

Computation and Language · Computer Science 2022-09-27 Ramya Tekumalla , Juan M. Banda

Background. After a year and half and over 4 million deaths, the COVID-19 pandemic continues to be widespread, and its related topics continue to dominate the global media. Although COVID-19 diagnoses have been well monitored, neither the…

Computers and Society · Computer Science 2021-08-19 Guangqing Chi , Junjun Yin , M. Luke Smith , Yosef Bodovski

Adverse drug reactions (ADRs) are one of the leading causes of mortality in health care. Current ADR surveillance systems are often associated with a substantial time lag before such events are officially published. On the other hand,…

Information Retrieval · Computer Science 2018-02-15 Shashank Gupta , Manish Gupta , Vasudeva Varma , Sachin Pawar , Nitin Ramrakhiyani , Girish K. Palshikar

COVID-19 pandemic has generated what public health officials called an infodemic of misinformation. As social distancing and stay-at-home orders came into effect, many turned to social media for socializing. This increase in social media…

Social and Information Networks · Computer Science 2021-06-15 Mir Mehedi A. Pritom , Rosana Montanez Rodriguez , Asad Ali Khan , Sebastian A. Nugroho , Esra'a Alrashydah , Beatrice N. Ruiz , Anthony Rios

Fake tweets are observed to be ever-increasing, demanding immediate countermeasures to combat their spread. During COVID-19, tweets with misinformation should be flagged and neutralized in their early stages to mitigate the damages. Most of…

Computation and Language · Computer Science 2021-04-27 Rachit Bansal , William Scott Paka , Nidhi , Shubhashis Sengupta , Tanmoy Chakraborty

The ongoing COVID-19 pandemic has caused immeasurable losses for people worldwide. To contain the spread of the virus and further alleviate the crisis, various health policies (e.g., stay-at-home orders) have been issued which spark heated…

Computation and Language · Computer Science 2023-01-26 Feng Xie , Zhong Zhang , Xuechen Zhao , Haiyang Wang , Jiaying Zou , Lei Tian , Bin Zhou , Yusong Tan

Neural Networks (NNs) are vulnerable to adversarial examples. Such inputs differ only slightly from their benign counterparts yet provoke misclassifications of the attacked NNs. The required perturbations to craft the examples are often…

Cryptography and Security · Computer Science 2020-09-30 Philip Sperl , Konstantin Böttinger

We introduce BERTweetFR, the first large-scale pre-trained language model for French tweets. Our model is initialized using the general-domain French language model CamemBERT which follows the base architecture of RoBERTa. Experiments show…

Computation and Language · Computer Science 2021-09-22 Yanzhu Guo , Virgile Rennard , Christos Xypolopoulos , Michalis Vazirgiannis

As a major social media platform, Twitter publishes a large number of user-generated text (tweets) on a daily basis. Mining such data can be used to address important social, public health, and emergency management issues that are…

Computation and Language · Computer Science 2021-12-07 Qing Han , Shubo Tian , Jinfeng Zhang

Subjectivity and difference of opinion are key social phenomena, and it is crucial to take these into account in the annotation and detection process of derogatory textual content. In this paper, we use four datasets provided by…

Computation and Language · Computer Science 2023-05-03 Sadat Shahriar , Thamar Solorio

The act of appearing kind or helpful via the use of but having a feeling of superiority condescending and patronizing language can have have serious mental health implications to those that experience it. Thus, detecting this condescending…

Computation and Language · Computer Science 2022-03-29 Xingmeng Zhao , Anthony Rios

This paper presents the different models submitted by the LT@Helsinki team for the SemEval 2020 Shared Task 12. Our team participated in sub-tasks A and C; titled offensive language identification and offense target identification,…

Computation and Language · Computer Science 2020-08-04 Marc Pàmies , Emily Öhman , Kaisla Kajava , Jörg Tiedemann

This paper presents models created for the Social Media Mining for Health 2023 shared task. Our team addressed the first task, classifying tweets that self-report Covid-19 diagnosis. Our approach involves a classification model that…

Computation and Language · Computer Science 2023-11-08 Sumam Francis , Marie-Francine Moens

In this paper, we present various systems submitted by our team problemConquero for SemEval-2020 Shared Task 12 Multilingual Offensive Language Identification in Social Media. We participated in all the three sub-tasks of OffensEval-2020,…

Computation and Language · Computer Science 2020-07-23 Karishma Laud , Jagriti Singh , Randeep Kumar Sahu , Ashutosh Modi

In elections around the world, the candidates may turn their campaigns toward negativity due to the prospect of failure and time pressure. In the digital age, social media platforms such as Twitter are rich sources of political discourse.…

Machine Learning · Computer Science 2023-11-02 Fatemeh Rajabi , Ali Mohades