English
Related papers

Related papers: Overview and Insights from the SciVer Shared Task …

200 papers

In this paper, we introduce CIBER (Claim Investigation Based on Evidence Retrieval), an extension of the Retrieval-Augmented Generation (RAG) framework designed to identify corroborating and refuting documents as evidence for scientific…

Artificial Intelligence · Computer Science 2025-03-12 Siyuan Wang , James R. Foulds , Md Osman Gani , Shimei Pan

Recent studies investigated the challenge of assessing the strength of a given claim extracted from a dataset, particularly the claim's potential of being misleading and cherry-picked. We focus on claims that compare answers to an aggregate…

Databases · Computer Science 2024-08-28 Shunit Agmon , Amir Gilad , Brit Youngmann , Shahar Zoarets , Benny Kimelfeld

We describe the third edition of the CheckThat! Lab, which is part of the 2020 Cross-Language Evaluation Forum (CLEF). CheckThat! proposes four complementary tasks and a related task from previous lab editions, offered in English, Arabic,…

Computation and Language · Computer Science 2020-01-24 Alberto Barron-Cedeno , Tamer Elsayed , Preslav Nakov , Giovanni Da San Martino , Maram Hasanain , Reem Suwaileh , Fatima Haouari

This report summarizes the 6th International Verification of Neural Networks Competition (VNN-COMP 2025), held as a part of the 8th International Symposium on AI Verification (SAIV), that was collocated with the 37th International…

An important component of an automated fact-checking system is the claim check-worthiness detection system, which ranks sentences by prioritising them based on their need to be checked. Despite a body of research tackling the task, previous…

Computation and Language · Computer Science 2022-12-19 Amani S. Abumansour , Arkaitz Zubiaga

Scientific papers make claims about prior work backed by citations. Verifying those citations at scale (that each cited paper exists, says what the citation claims, and is itself reliable) is structurally beyond what human review can…

Digital Libraries · Computer Science 2026-05-26 Sergey V Samsonau

We introduce CLIMATE-FEVER, a new publicly available dataset for verification of climate change-related claims. By providing a dataset for the research community, we aim to facilitate and encourage work on improving algorithms for…

Computation and Language · Computer Science 2021-01-05 Thomas Diggelmann , Jordan Boyd-Graber , Jannis Bulian , Massimiliano Ciaramita , Markus Leippold

In the digital age, seeking health advice on the Internet has become a common practice. At the same time, determining the trustworthiness of online medical content is increasingly challenging. Fact-checking has emerged as an approach to…

Computation and Language · Computer Science 2024-03-26 Juraj Vladika , Phillip Schneider , Florian Matthes

Claim verification with large language models (LLMs) has recently attracted growing attention, due to their strong reasoning capabilities and transparent verification processes compared to traditional answer-only judgments. However,…

Computation and Language · Computer Science 2025-10-07 Qi He , Cheng Qian , Xiusi Chen , Bingxiang He , Yi R. Fung , Heng Ji

Evidence retrieval is a core part of automatic fact-checking. Prior work makes simplifying assumptions in retrieval that depart from real-world use cases: either no access to evidence, access to evidence curated by a human fact-checker, or…

Computation and Language · Computer Science 2024-06-18 Jifan Chen , Grace Kim , Aniruddh Sriram , Greg Durrett , Eunsol Choi

This paper presents an overview of the ImageArg shared task, the first multimodal Argument Mining shared task co-located with the 10th Workshop on Argument Mining at EMNLP 2023. The shared task comprises two classification subtasks - (1)…

Computation and Language · Computer Science 2023-10-25 Zhexiong Liu , Mohamed Elaraby , Yang Zhong , Diane Litman

Students have enthusiastically taken to online programming lessons and contests. Unfortunately, they tend to struggle due to lack of personalized feedback when they make mistakes. The overwhelming number of submissions precludes manual…

Software Engineering · Computer Science 2016-03-16 Shalini Kaleeswaran , Anirudh Santhiar , Aditya Kanade , Sumit Gulwani

Following the global COVID-19 pandemic, the number of scientific papers studying the virus has grown massively, leading to increased interest in automated literate review. We present a clinical text mining system that improves on previous…

Computation and Language · Computer Science 2020-12-09 Veysel Kocaman , David Talby

Large Language Models (LLMs) are at the forefront of NLP achievements but fall short in dealing with shortcut learning, factual inconsistency, and vulnerability to adversarial inputs.These shortcomings are especially critical in medical…

Computation and Language · Computer Science 2024-04-09 Mael Jullien , Marco Valentino , André Freitas

Fact-checking is the task of verifying the veracity of claims by assessing their assertions against credible evidence. The vast majority of fact-checking studies focus exclusively on political claims. Very little research explores…

Computation and Language · Computer Science 2020-10-21 Neema Kotonya , Francesca Toni

We introduce a FEVER-like dataset COVID-Fact of $4,086$ claims concerning the COVID-19 pandemic. The dataset contains claims, evidence for the claims, and contradictory claims refuted by the evidence. Unlike previous approaches, we…

Computation and Language · Computer Science 2021-06-08 Arkadiy Saakyan , Tuhin Chakrabarty , Smaranda Muresan

This paper presents the results of the shared task on Lay Summarisation of Biomedical Research Articles (BioLaySumm), hosted at the BioNLP Workshop at ACL 2023. The goal of this shared task is to develop abstractive summarisation models…

Computation and Language · Computer Science 2023-10-26 Tomas Goldsack , Zheheng Luo , Qianqian Xie , Carolina Scarton , Matthew Shardlow , Sophia Ananiadou , Chenghua Lin

Scientific contributions are a direct reflection of a research paper's value, illustrating its impact on existing theories or practices. Existing measurement methods assess contributions based on the authors' perceived or self-identified…

Digital Libraries · Computer Science 2024-10-18 Liyue Chen , Jielan Ding , Donghuan Song , Zihao Qu

Evaluating scientific arguments requires assessing the strict consistency between a claim and its underlying multimodal evidence. However, existing benchmarks lack the scale, domain diversity, and visual complexity needed to evaluate this…

Computation and Language · Computer Science 2026-04-21 Abolfazl Ansari , Delvin Ce Zhang , Zhuoyang Zou , Wenpeng Yin , Dongwon Lee

Despite rapid progress in claim verification, we lack a systematic understanding of what reasoning these benchmarks actually exercise. We generate structured reasoning traces for 24K claim-verification examples across 9 datasets using…

Computation and Language · Computer Science 2026-04-03 Delip Rao , Chris Callison-Burch
‹ Prev 1 8 9 10 Next ›