中文
相关论文

相关论文: Text Augmentations with R-drop for Classification …

200 篇论文

Data augmentation techniques are widely used for enhancing the performance of machine learning models by tackling class imbalance issues and data sparsity. State-of-the-art generative language models have been shown to provide significant…

计算与语言 · 计算机科学 2023-01-10 Aleksandra Edwards , Asahi Ushio , Jose Camacho-Collados , Hélène de Ribaupierre , Alun Preece

With the rise of the Internet, there is a growing need to build intelligent systems that are capable of efficiently dealing with early risk detection (ERD) problems on social media, such as early depression detection, early rumor detection…

计算机与社会 · 计算机科学 2024-04-18 Sergio G. Burdisso , Marcelo Errecalde , Manuel Montes-y-Gómez

This paper discusses the approach used by the Accenture Team for CLEF2021 CheckThat! Lab, Task 1, to identify whether a claim made in social media would be interesting to a wide audience and should be fact-checked. Twitter training and test…

计算与语言 · 计算机科学 2021-07-14 Evan Williams , Paul Rodrigues , Sieu Tran

In this paper, we present our work participating in the BioCreative VII Track 3 - automatic extraction of medication names in tweets, where we implemented a multi-task learning model that is jointly trained on text classification and…

计算与语言 · 计算机科学 2021-11-30 Dongfang Xu , Shan Chen , Timothy Miller

We present semi-supervised models with data augmentation (SMDA), a semi-supervised text classification system to classify interactive affective responses. SMDA utilizes recent transformer-based models to encode each sentence and employs…

计算与语言 · 计算机科学 2020-04-24 Jiaao Chen , Yuwei Wu , Diyi Yang

Personality detection aims to detect one's personality traits underlying in social media posts. One challenge of this task is the scarcity of ground-truth personality traits which are collected from self-report questionnaires. Most existing…

计算与语言 · 计算机科学 2024-03-13 Linmei Hu , Hongyu He , Duokang Wang , Ziwang Zhao , Yingxia Shao , Liqiang Nie

Based on recent advances in natural language modeling and those in text generation capabilities, we propose a novel data augmentation method for text classification tasks. We use a powerful pre-trained neural network model to artificially…

The categorization of massive e-Commerce data is a crucial, well-studied task, which is prevalent in industrial settings. In this work, we aim to improve an existing product categorization model that is already in use by a major web…

机器学习 · 计算机科学 2023-05-31 Guy Horowitz , Stav Yanovsky Daye , Noa Avigdor-Elgrabli , Ariel Raviv

In recent years, social media platforms have hosted an explosion of hate speech and objectionable content. The urgent need for effective automatic hate speech detection models have drawn remarkable investment from companies and researchers.…

计算与语言 · 计算机科学 2020-10-27 Sayyed M. Zahiri , Ali Ahmadvand

Billions of people across the globe have been using social media platforms in their local languages to voice their opinions about the various topics related to the COVID-19 pandemic. Several organizations, including the World Health…

计算与语言 · 计算机科学 2022-10-13 Rabin Adhikari , Safal Thapaliya , Nirajan Basnet , Samip Poudel , Aman Shakya , Bishesh Khanal

Sentiment analysis has been an active area of research in the past two decades and recently, with the advent of social media, there has been an increasing demand for sentiment analysis on social media texts. Since the social media texts are…

计算与语言 · 计算机科学 2020-10-21 Sainik Kumar Mahata , Dipankar Das , Sivaji Bandyopadhyay

Background. After a year and half and over 4 million deaths, the COVID-19 pandemic continues to be widespread, and its related topics continue to dominate the global media. Although COVID-19 diagnoses have been well monitored, neither the…

计算机与社会 · 计算机科学 2021-08-19 Guangqing Chi , Junjun Yin , M. Luke Smith , Yosef Bodovski

Adverse drug reactions (ADRs) are one of the leading causes of mortality in health care. Current ADR surveillance systems are often associated with a substantial time lag before such events are officially published. On the other hand,…

信息检索 · 计算机科学 2018-02-15 Shashank Gupta , Manish Gupta , Vasudeva Varma , Sachin Pawar , Nitin Ramrakhiyani , Girish K. Palshikar

The reported work is our straightforward approach for the shared task Classification of tweets self-reporting age organized by the Social Media Mining for Health Applications (SMM4H) workshop. This literature describes the approach that was…

计算与语言 · 计算机科学 2023-01-16 Keshav Kapur , Rajitha Harikrishnan

Creating explanations for answers to science questions is a challenging task that requires multi-hop inference over a large set of fact sentences. This year, to refocus the Textgraphs Shared Task on the problem of gathering relevant…

计算与语言 · 计算机科学 2021-07-29 Vivek Kalyan , Sam Witteveen , Martin Andrews

The detection of suicide risk in social media is a critical task with potential life-saving implications. This paper presents a study on leveraging state-of-the-art natural language processing solutions for identifying suicide risk in…

计算与语言 · 计算机科学 2024-10-14 Jakub Pokrywka , Jeremi I. Kaczmarek , Edward J. Gorzelańczyk

Semi-supervised learning approaches have been investigated as a means to enhance the analysis of social media data in disaster management contexts. In this work, we present the first empirical evaluation of large language model (LLM) guided…

In recent years, social media has been widely explored as a potential source of communication and information in disasters and emergency situations. Several interesting works and case studies of disaster analytics exploring different…

计算与语言 · 计算机科学 2023-01-03 Wisal Mukhtiar , Waliiya Rizwan , Aneela Habib , Yasir Saleem Afridi , Laiq Hasan , Kashif Ahmad

In this paper we present a method to identify tweets that a user may find interesting enough to retweet. The method is based on a global, but personalized classifier, which is trained on data from several users, represented in terms of…

社会与信息网络 · 计算机科学 2017-09-20 Michail Vougioukas , Ion Androutsopoulos , Georgios Paliouras

Data augmentation has been widely used to improve deep neural networks in many research fields, such as computer vision. However, less work has been done in the context of text, partially due to its discrete nature and the complexity of…

计算与语言 · 计算机科学 2021-01-12 Ping Yu , Ruiyi Zhang , Yang Zhao , Yizhe Zhang , Chunyuan Li , Changyou Chen