English

Findings of the Fourth Shared Task on Multilingual Coreference Resolution: Can LLMs Dethrone Traditional Approaches?

Computation and Language 2025-11-07 v2

Abstract

The paper presents an overview of the fourth edition of the Shared Task on Multilingual Coreference Resolution, organized as part of the CODI-CRAC 2025 workshop. As in the previous editions, participants were challenged to develop systems that identify mentions and cluster them according to identity coreference. A key innovation of this year's task was the introduction of a dedicated Large Language Model (LLM) track, featuring a simplified plaintext format designed to be more suitable for LLMs than the original CoNLL-U representation. The task also expanded its coverage with three new datasets in two additional languages, using version 1.3 of CorefUD - a harmonized multilingual collection of 22 datasets in 17 languages. In total, nine systems participated, including four LLM-based approaches (two fine-tuned and two using few-shot adaptation). While traditional systems still kept the lead, LLMs showed clear potential, suggesting they may soon challenge established approaches in future editions.

Keywords

Cite

@article{arxiv.2509.17796,
  title  = {Findings of the Fourth Shared Task on Multilingual Coreference Resolution: Can LLMs Dethrone Traditional Approaches?},
  author = {Michal Novák and Miloslav Konopík and Anna Nedoluzhko and Martin Popel and Ondřej Pražák and Jakub Sido and Milan Straka and Zdeněk Žabokrtský and Daniel Zeman},
  journal= {arXiv preprint arXiv:2509.17796},
  year   = {2025}
}

Comments

Accepted to CODI-CRAC 2025