Is Japanese CCGBank empirically correct? A case study of passive and causative constructions
Computation and Language
2023-03-01 v1
Abstract
The Japanese CCGBank serves as training and evaluation data for developing Japanese CCG parsers. However, since it is automatically generated from the Kyoto Corpus, a dependency treebank, its linguistic validity still needs to be sufficiently verified. In this paper, we focus on the analysis of passive/causative constructions in the Japanese CCGBank and show that, together with the compositional semantics of ccg2lambda, a semantic parsing system, it yields empirically wrong predictions for the nested construction of passives and causatives.
Keywords
Cite
@article{arxiv.2302.14708,
title = {Is Japanese CCGBank empirically correct? A case study of passive and causative constructions},
author = {Daisuke Bekki and Hitomi Yanaka},
journal= {arXiv preprint arXiv:2302.14708},
year = {2023}
}
Comments
To appear in Proceedings of Treebanks and Linguistic Theories (TLT) 2023, the workshop in the Georgetown University Round Table on Linguistics 2023 (GURT2023)