中文
相关论文

相关论文: Beyond the Crowd: LLM-Augmented Community Notes fo…

200 篇论文

Crowdsourcing provides a practical way to obtain large amounts of labeled data at a low cost. However, the annotation quality of annotators varies considerably, which imposes new challenges in learning a high-quality model from the…

机器学习 · 计算机科学 2021-06-15 Zhendong Chu , Jing Ma , Hongning Wang

The release note is a crucial document outlining changes in new software versions. Yet, many developers view the process of writing software release notes as a tedious and dreadful task. Consequently, numerous tools have been developed by…

软件工程 · 计算机科学 2025-05-26 Farbod Daneshyan , Runzhi He , Jianyu Wu , Minghui Zhou

Investigating the public experience of urgent care facilities is essential for promoting community healthcare development. Traditional survey methods often fall short due to limited scope, time, and spatial coverage. Crowdsourcing through…

计算与语言 · 计算机科学 2026-05-19 Xiaoran Xu , Zhaoqian Xue , Chi Zhang , Jhonatan Medri , Junjie Xiong , Jiayan Zhou , Jin Jin , Yongfeng Zhang , Siyuan Ma , Lingyao Li

Prior research on Twitter (now X) data has provided positive evidence of its utility in developing supplementary health surveillance systems. In this study, we present a new framework to surveil public health, focusing on mental health (MH)…

社会与信息网络 · 计算机科学 2025-05-26 Vijeta Deshpande , Minhwa Lee , Zonghai Yao , Zihao Zhang , Jason Brian Gibbons , Hong Yu

Recently, the misinformation problem has been addressed with a crowdsourcing-based approach: to assess the truthfulness of a statement, instead of relying on a few experts, a crowd of non-expert is exploited. We study whether crowdsourcing…

We present crowdsourcing as an additional modality to aid radiologists in the diagnosis of lung cancer from clinical chest computed tomography (CT) scans. More specifically, a complete workflow is introduced which can help maximize the…

计算机视觉与模式识别 · 计算机科学 2018-09-19 Saeed Boorboor , Saad Nadeem , Ji Hwan Park , Kevin Baker , Arie Kaufman

Health literacy is a critical determinant of patient outcomes, yet current screening tools are not always feasible and differ considerably in the number of items, question format, and dimensions of health literacy they capture, making…

Online Mental Health Communities (OMHCs) provide crucial peer and expert support, yet many posts remain unanswered due to missing support attributes that signal the need for help. We present a novel framework that identifies these gaps and…

计算与语言 · 计算机科学 2025-08-26 Bhagesh Gaur , Karan Gupta , Aseem Srivastava , Manish Gupta , Md Shad Akhtar

Clinical note generation aims to produce free-text summaries of a patient's condition and diagnostic process, with discharge instructions being a representative long-form example. While recent LLM-based methods pre-trained on general…

计算与语言 · 计算机科学 2025-08-12 Lo Pang-Yun Ting , Chengshuai Zhao , Yu-Hua Zeng , Yuan Jee Lim , Kun-Ta Chuang , Huan Liu

Fears about the destabilizing impact of misinformation online have motivated individuals and platforms to respond. Individuals have increasingly challenged others' online claims with fact-checks in pursuit of a healthier information…

计算机与社会 · 计算机科学 2025-01-27 Junsol Kim , Zhao Wang , Haohan Shi , Hsin-Keng Ling , James Evans

People enjoy sharing "notes" including their experiences within online communities. Therefore, recommending notes aligned with user interests has become a crucial task. Existing online methods only input notes into BERT-based models to…

信息检索 · 计算机科学 2024-03-26 Chao Zhang , Shiwei Wu , Haoxin Zhang , Tong Xu , Yan Gao , Yao Hu , Di Wu , Enhong Chen

Microtask crowdsourcing has enabled dataset advances in social science and machine learning, but existing crowdsourcing schemes are too expensive to scale up with the expanding volume of data. To scale and widen the applicability of…

An important unexplored aspect in previous work on user satisfaction estimation for Task-Oriented Dialogue (TOD) systems is their evaluation in terms of robustness for the identification of user dissatisfaction: current benchmarks for user…

计算与语言 · 计算机科学 2024-08-21 Amin Abolghasemi , Zhaochun Ren , Arian Askari , Mohammad Aliannejadi , Maarten de Rijke , Suzan Verberne

In this article, we propose a simulated crowd counting dataset CrowdX, which has a large scale, accurate labeling, parameterized realization, and high fidelity. The experimental results of using this dataset as data enhancement show that…

计算机视觉与模式识别 · 计算机科学 2022-01-25 Yi Hou , Chengyang Li , Yuheng Lu , Liping Zhu , Yuan Li , Huizhu Jia , Xiaodong Xie

Generative information extraction using large language models, particularly through few-shot learning, has become a popular method. Recent studies indicate that providing a detailed, human-readable guideline-similar to the annotation…

Large Language Models (LLMs) are increasingly used as scalable tools for pilot testing, predicting public opinion distributions before deploying costly surveys. To serve as effective pilot testing tools, the performance of these LLMs is…

社会与信息网络 · 计算机科学 2025-11-11 Xutao Mao , Ezra Xuanru Tao , Leyao Wang

Rapid integration of large language models (LLMs) in health care is sparking global discussion about their potential to revolutionize health care quality and accessibility. At a time when improving health care quality and access remains a…

计算机与社会 · 计算机科学 2025-04-01 Troy Zada , Natalie Tam , Francois Barnard , Marlize Van Sittert , Venkat Bhat , Sirisha Rambhatla

A clinical study is often necessary for exploring important research questions; however, this approach is sometimes time and money consuming. Another extreme approach, which is to collect and aggregate opinions from crowds, provides a…

人机交互 · 计算机科学 2022-05-17 Shoko Wakamiya , Toshiki Mera , Eiji Aramaki , Masaki Matsubara , Atsuyuki Morishima

Distant supervision is a popular method for performing relation extraction from text that is known to produce noisy labels. Most progress in relation extraction and classification has been made with crowdsourced corrections to…

计算与语言 · 计算机科学 2022-09-21 Anca Dumitrache , Lora Aroyo , Chris Welty

Data labeling is a necessary but often slow process that impedes the development of interactive systems for modern data analysis. Despite rising demand for manual data labeling, there is a surprising lack of work addressing its high and…

数据库 · 计算机科学 2015-09-22 Daniel Haas , Jiannan Wang , Eugene Wu , Michael J. Franklin