中文

基于机器学习的日语名词短语指代属性估计方法

计算与语言 2007-05-23 v1

摘要

日语中名词短语的指代属性在日英机器翻译中的文章生成以及日语名词短语的指代消解中都很有用处。由于日语没有冠词,这些指代属性通常被分类为普通名词短语、明确指代名词短语和不明确指代名词短语。在之前的工作中,通过开发使用线索词的规则来估计指代属性。如果多个规则之间存在冲突,将选择给定总分最高的类别作为目标类别。每个规则给出的分数是通过人工设定的,因此人力成本较高。在本工作中,我们使用机器学习方法自动调整这些分数,成功地减少了人力调整这些分数所需的成本。

关键词

引用

@article{arxiv.cs/0103011,
  title  = {A Machine-Learning Approach to Estimating the Referential Properties of Japanese Noun Phrases},
  author = {Masaki Murata and Kiyotaka Uchimoto and Qing Ma and Hitoshi Isahara},
  journal= {arXiv preprint arXiv:cs/0103011},
  year   = {2007}
}

备注

9 pages. Computation and Language. This paper is included in the book entitled by "Computational Linguistics and Intelligent Text Processing, Second International Conference, CICLing 2001, Mexico City, February 2001 Proceedings", Alexander Gelbukh (Ed.), Springer Publisher, ISSN 0302-9743, ISBN 3-540-41687-0