中文
相关论文

相关论文: Text Augmentations with R-drop for Classification …

200 篇论文

Amid ongoing health crisis, there is a growing necessity to discern possible signs of Wellness Dimensions (WD) manifested in self-narrated text. As the distribution of WD on social media data is intrinsically imbalanced, we experiment the…

计算与语言 · 计算机科学 2023-06-08 Chandreen Liyanage , Muskan Garg , Vijay Mago , Sunghwan Sohn

Vaccination is important to minimize the risk and spread of various diseases. In recent years, vaccination has been a key step in countering the COVID-19 pandemic. However, many people are skeptical about the use of vaccines for various…

计算与语言 · 计算机科学 2023-12-19 Aniket Deroy , Subhankar Maity

In this paper, we describe our system for the AAAI 2021 shared task of COVID-19 Fake News Detection in English, where we achieved the 3rd position with the weighted F1 score of 0.9859 on the test set. Specifically, we proposed an ensemble…

计算与语言 · 计算机科学 2021-09-24 Xiangyang Li , Yu Xia , Xiang Long , Zheng Li , Sujian Li

With the COVID-19 pandemic, there is a growing urgency for medical community to keep up with the accelerating growth in the new coronavirus-related literature. As a result, the COVID-19 Open Research Dataset Challenge has released a corpus…

计算与语言 · 计算机科学 2020-06-04 Virapat Kieuvongngam , Bowen Tan , Yiming Niu

We present working notes for DS@GT team in the eRisk 2024 for Tasks 1 and 3. We propose a ranking system for Task 1 that predicts symptoms of depression based on the Beck Depression Inventory (BDI-II) questionnaire using binary classifiers…

计算与语言 · 计算机科学 2024-07-12 David Guecha , Aaryan Potdar , Anthony Miyaguchi

Over the course of the COVID-19 pandemic, large volumes of biomedical information concerning this new disease have been published on social media. Some of this information can pose a real danger to people's health, particularly when false…

计算与语言 · 计算机科学 2022-04-27 Isabelle Mohr , Amelie Wührl , Roman Klinger

This paper presents the approach that we employed to tackle the EMNLP WNUT-2020 Shared Task 2 : Identification of informative COVID-19 English Tweets. The task is to develop a system that automatically identifies whether an English Tweet…

计算与语言 · 计算机科学 2020-12-17 Anshul Wadhawan

In many cases of machine learning, research suggests that the development of training data might have a higher relevance than the choice and modelling of classifiers themselves. Thus, data augmentation methods have been developed to improve…

计算与语言 · 计算机科学 2022-07-25 Markus Bayer , Marc-André Kaufhold , Björn Buchhold , Marcel Keller , Jörg Dallmeyer , Christian Reuter

The option of sharing images, videos and audio files on social media opens up new possibilities for distinguishing between false information and fake news on the Internet. Due to the vast amount of data shared every second on social media,…

机器学习 · 计算机科学 2023-07-28 Raphael Frick , Inna Vogel

This paper presents our system employed for the Social Media Mining for Health 2023 Shared Task 4: Binary classification of English Reddit posts self-reporting a social anxiety disorder diagnosis. We systematically investigate and contrast…

计算与语言 · 计算机科学 2023-12-18 Sourabh Zanwar , Daniel Wiechmann , Yu Qiao , Elma Kerz

Nowadays, topic classification from tweets attracts considerable research attention. Different classification systems have been suggested thanks to these research efforts. Nevertheless, they face major challenges owing to low performance…

计算与语言 · 计算机科学 2024-07-04 Kheir Eddine Daouadi , Yaakoub Boualleg , Oussama Guehairia

The gold standard for COVID-19 is RT-PCR, testing facilities for which are limited and not always optimally distributed. Test results are delayed, which impacts treatment. Expert radiologists, one of whom is a co-author, are able to…

计算机视觉与模式识别 · 计算机科学 2021-02-17 Kartikeya Badola , Sameer Ambekar , Himanshu Pant , Sumit Soman , Anuradha Sural , Rajiv Narang , Suresh Chandra , Jayadeva

Over the last decade, there has been a vast increase in eating disorder diagnoses and eating disorder-attributed deaths, reaching their zenith during the Covid-19 pandemic. This immense growth derived in part from the stressors of the…

机器学习 · 计算机科学 2023-11-07 Jonathan Feldman

In online forums like Reddit, users share their experiences with medical conditions and treatments, including making claims, asking questions, and discussing the effects of treatments on their health. Building systems to understand this…

计算与语言 · 计算机科学 2023-04-28 Giridhar Kaushik Ramachandran , Haritha Gangavarapu , Kevin Lybarger , Ozlem Uzuner

This paper presents our submission to Task 1, Subjectivity Detection, of the CheckThat! Lab at CLEF 2025. We investigate the effectiveness of transfer-learning and stylistic data augmentation to improve classification of subjective and…

计算与语言 · 计算机科学 2025-07-09 Maximilian Heil , Dionne Bang

Mental illness affects a significant portion of the worldwide population. Online mental health forums can provide a supportive environment for those afflicted and also generate a large amount of data which can be mined to predict mental…

计算与语言 · 计算机科学 2019-07-12 Derek Howard , Marta Maslej , Justin Lee , Jacob Ritchie , Geoffrey Woollard , Leon French

Automatically associating social media posts with topics is an important prerequisite for effective search and recommendation on many social media platforms. However, topic classification of such posts is quite challenging because of (a) a…

计算与语言 · 计算机科学 2022-05-04 Vivek Kulkarni , Kenny Leung , Aria Haghighi

Bragging is a speech act employed with the goal of constructing a favorable self-image through positive statements about oneself. It is widespread in daily communication and especially popular in social media, where users aim to build a…

计算与语言 · 计算机科学 2022-03-14 Mali Jin , Daniel Preoţiuc-Pietro , A. Seza Doğruöz , Nikolaos Aletras

In fact-checking, structure and phrasing of claims critically influence a model's ability to predict verdicts accurately. Social media content in particular rarely serves as optimal input for verification systems, which necessitates…

计算与语言 · 计算机科学 2024-12-17 Amelie Wührl , Roman Klinger

Data augmentation techniques are widely used in text classification tasks to improve the performance of classifiers, especially in low-resource scenarios. Most previous methods conduct text augmentation without considering the different…

计算与语言 · 计算机科学 2022-09-07 Biyang Guo , Songqiao Han , Hailiang Huang