English

Diverse Linguistic Features for Assessing Reading Difficulty of Educational Filipino Texts

Computation and Language 2021-08-03 v1 Machine Learning

Abstract

In order to ensure quality and effective learning, fluency, and comprehension, the proper identification of the difficulty levels of reading materials should be observed. In this paper, we describe the development of automatic machine learning-based readability assessment models for educational Filipino texts using the most diverse set of linguistic features for the language. Results show that using a Random Forest model obtained a high performance of 62.7% in terms of accuracy, and 66.1% when using the optimal combination of feature sets consisting of traditional and syllable pattern-based predictors.

Keywords

Cite

@article{arxiv.2108.00241,
  title  = {Diverse Linguistic Features for Assessing Reading Difficulty of Educational Filipino Texts},
  author = {Joseph Marvin Imperial and Ethel Ong},
  journal= {arXiv preprint arXiv:2108.00241},
  year   = {2021}
}

Comments

Accepted at ICCE 2021