English

Is Japanese CCGBank empirically correct? A case study of passive and causative constructions

Computation and Language 2023-03-01 v1

Abstract

The Japanese CCGBank serves as training and evaluation data for developing Japanese CCG parsers. However, since it is automatically generated from the Kyoto Corpus, a dependency treebank, its linguistic validity still needs to be sufficiently verified. In this paper, we focus on the analysis of passive/causative constructions in the Japanese CCGBank and show that, together with the compositional semantics of ccg2lambda, a semantic parsing system, it yields empirically wrong predictions for the nested construction of passives and causatives.

Keywords

Cite

@article{arxiv.2302.14708,
  title  = {Is Japanese CCGBank empirically correct? A case study of passive and causative constructions},
  author = {Daisuke Bekki and Hitomi Yanaka},
  journal= {arXiv preprint arXiv:2302.14708},
  year   = {2023}
}

Comments

To appear in Proceedings of Treebanks and Linguistic Theories (TLT) 2023, the workshop in the Georgetown University Round Table on Linguistics 2023 (GURT2023)