English

Is this sentence valid? An Arabic Dataset for Commonsense Validation

Computation and Language 2020-08-26 v1 Computers and Society

Abstract

The commonsense understanding and validation remains a challenging task in the field of natural language understanding. Therefore, several research papers have been published that studied the capability of proposed systems to evaluate the models ability to validate commonsense in text. In this paper, we present a benchmark Arabic dataset for commonsense understanding and validation as well as a baseline research and models trained using the same dataset. To the best of our knowledge, this dataset is considered as the first in the field of Arabic text commonsense validation. The dataset is distributed under the Creative Commons BY-SA 4.0 license and can be found on GitHub.

Keywords

Cite

@article{arxiv.2008.10873,
  title  = {Is this sentence valid? An Arabic Dataset for Commonsense Validation},
  author = {Saja Tawalbeh and Mohammad AL-Smadi},
  journal= {arXiv preprint arXiv:2008.10873},
  year   = {2020}
}

Comments

4 pages

R2 v1 2026-06-23T18:05:05.193Z