English

ICPR 2024 Competition on Multilingual Claim-Span Identification

Computation and Language 2024-12-02 v1

Abstract

A lot of claims are made in social media posts, which may contain misinformation or fake news. Hence, it is crucial to identify claims as a first step towards claim verification. Given the huge number of social media posts, the task of identifying claims needs to be automated. This competition deals with the task of 'Claim Span Identification' in which, given a text, parts / spans that correspond to claims are to be identified. This task is more challenging than the traditional binary classification of text into claim or not-claim, and requires state-of-the-art methods in Pattern Recognition, Natural Language Processing and Machine Learning. For this competition, we used a newly developed dataset called HECSI containing about 8K posts in English and about 8K posts in Hindi with claim-spans marked by human annotators. This paper gives an overview of the competition, and the solutions developed by the participating teams.

Keywords

Cite

@article{arxiv.2411.19579,
  title  = {ICPR 2024 Competition on Multilingual Claim-Span Identification},
  author = {Soham Poddar and Biswajit Paul and Moumita Basu and Saptarshi Ghosh},
  journal= {arXiv preprint arXiv:2411.19579},
  year   = {2024}
}

Comments

To appear at ICPR 2024

R2 v1 2026-06-28T20:16:36.730Z