English
Related papers

Related papers: Team HUMANE at AVeriTeC 2025: HerO 2 for Efficient…

200 papers

To tackle the AVeriTeC shared task hosted by the FEVER-24, we introduce a system that only employs publicly available large language models (LLMs) for each step of automated fact-checking, dubbed the Herd of Open LLMs for verifying…

Computation and Language · Computer Science 2024-10-22 Yejun Yoon , Jaeyoon Jung , Seunghyun Yoon , Kunwoo Park

The Automatic Verification of Image-Text Claims (AVerImaTeC) shared task aims to advance system development for retrieving evidence and verifying real-world image-text claims. Participants were allowed to either employ external knowledge…

Computation and Language · Computer Science 2026-02-27 Rui Cao , Zhenyun Deng , Yulong Chen , Michael Schlichtkrull , Andreas Vlachos

The Automated Verification of Textual Claims (AVeriTeC) shared task asks participants to retrieve evidence and predict veracity for real-world claims checked by fact-checkers. Evidence can be found either via a search engine, or via a…

Separating disinformation from fact on the web has long challenged both the search and the reasoning powers of humans. We show that the reasoning power of large language models (LLMs) and the retrieval power of modern search engines can be…

Computation and Language · Computer Science 2024-11-11 Christopher Malon

Controllable generation through Stable Diffusion (SD) fine-tuning aims to improve fidelity, safety, and alignment with human guidance. Existing reinforcement learning from human feedback methods usually rely on predefined heuristic reward…

Even though many machine algorithms have been proposed for entity resolution, it remains very challenging to find a solution with quality guarantees. In this paper, we propose a novel HUman and Machine cOoperation (HUMO) framework for…

Databases · Computer Science 2018-04-03 Zhaoqiang Chen , Qun Chen , Fengfeng Fan , Yanyan Wang , Zhuo Wang , Youcef Nafa , Zhanhuai Li , Hailong Liu , Wei Pan

This paper describes our $3^{rd}$ place submission in the AVeriTeC shared task in which we attempted to address the challenge of fact-checking with evidence retrieved in the wild using a simple scheme of Retrieval-Augmented Generation (RAG)…

Computation and Language · Computer Science 2024-10-16 Herbert Ullrich , Tomáš Mlynář , Jan Drchal

This paper presents the HYBRINFOX method used to solve Task 2 of Subjectivity detection of the CLEF 2024 CheckThat! competition. The specificity of the method is to use a hybrid system, combining a RoBERTa model, fine-tuned for subjectivity…

Computation and Language · Computer Science 2024-07-08 Morgane Casanova , Julien Chanson , Benjamin Icard , Géraud Faye , Guillaume Gadek , Guillaume Gravier , Paul Égré

Even though many approaches have been proposed for entity resolution (ER), it remains very challenging to find one with quality guarantees. To this end, we proposea risk-aware HUman-Machine cOoperation framework for ER, denoted by r-HUMO.…

Human-Computer Interaction · Computer Science 2018-11-27 Boyi Hou , Qun Chen , Zhaoqiang Chen , Youcef Nafa , Zhanhuai Li

We propose a novel model for learned query optimization which provides query hints leading to better execution plans. The model addresses the three key challenges in learned hint-based query optimization: reliable hint recommendation…

Databases · Computer Science 2024-12-06 Sergey Zinchenko , Sergey Iazov

Humanity's Last Exam (HLE) has become a widely used benchmark for evaluating frontier large language models on challenging, multi-domain questions. However, community-led analyses have raised concerns that HLE contains a non-trivial number…

Existing datasets for automated fact-checking have substantial limitations, such as relying on artificial claims, lacking annotations for evidence and intermediate reasoning, or including evidence published after the claim. In this paper we…

Computation and Language · Computer Science 2023-11-09 Michael Schlichtkrull , Zhijiang Guo , Andreas Vlachos

When designing evidence-based policies and programs, decision-makers must distill key information from a vast and rapidly growing literature base. Identifying relevant literature from raw search results is time and resource intensive, and…

Computation and Language · Computer Science 2023-05-03 Kristen M. Edwards , Binyang Song , Jaron Porciello , Mark Engelbert , Carolyn Huang , Faez Ahmed

In this paper we present our system for the FEVER Challenge. The task of this challenge is to verify claims by extracting information from Wikipedia. Our system has two parts. In the first part it performs a search for candidate sentences…

Information Retrieval · Computer Science 2018-12-31 Jan Kowollik , Ahmet Aker

We present the results of the first Fact Extraction and VERification (FEVER) Shared Task. The task challenged participants to classify whether human-written factoid claims could be Supported or Refuted using evidence retrieved from…

Computation and Language · Computer Science 2018-12-03 James Thorne , Andreas Vlachos , Oana Cocarascu , Christos Christodoulopoulos , Arpit Mittal

Human activities are particularly complex and variable, and this makes challenging for deep learning models to reason about them. However, we note that such variability does have an underlying structure, composed of a hierarchy of patterns…

Computer Vision and Pattern Recognition · Computer Science 2025-05-20 Simone Alberto Peirone , Francesca Pistilli , Giuseppe Averta

The increased focus on misinformation has spurred development of data and systems for detecting the veracity of a claim as well as retrieving authoritative evidence. The Fact Extraction and VERification (FEVER) dataset provides such a…

Computation and Language · Computer Science 2020-04-28 Christopher Hidey , Tuhin Chakrabarty , Tariq Alhindi , Siddharth Varia , Kriste Krstovski , Mona Diab , Smaranda Muresan

The Fact Extraction and VERification (FEVER) shared task was launched to support the development of systems able to verify claims by extracting supporting or refuting facts from raw text. The shared task organizers provide a large-scale…

Information Retrieval · Computer Science 2019-05-10 Andreas Hanselowski , Hao Zhang , Zile Li , Daniil Sorokin , Benjamin Schiller , Claudia Schulz , Iryna Gurevych

An important emerging application of coding agents is agent optimization: the iterative improvement of a target agent through edit-execute-evaluate cycles. Despite its relevance, the community lacks a systematic understanding of coding…

Artificial Intelligence · Computer Science 2026-05-12 Varun Ursekar , Apaar Shanker , Veronica Chatrath , Yuan Xue , Sam Denton

Automatic fact verification has become an increasingly popular topic in recent years and among datasets the Fact Extraction and VERification (FEVER) dataset is one of the most popular. In this work we present BEVERS, a tuned baseline system…

Computation and Language · Computer Science 2023-03-31 Mitchell DeHaven , Stephen Scott
‹ Prev 1 2 3 10 Next ›