English

Lexical Bias In Essay Level Prediction

Computation and Language 2018-09-25 v1 Artificial Intelligence

Abstract

Automatically predicting the level of non-native English speakers given their written essays is an interesting machine learning problem. In this work I present the system "balikasg" that achieved the state-of-the-art performance in the CAp 2018 data science challenge among 14 systems. I detail the feature extraction, feature engineering and model selection steps and I evaluate how these decisions impact the system's performance. The paper concludes with remarks for future work.

Keywords

Cite

@article{arxiv.1809.08935,
  title  = {Lexical Bias In Essay Level Prediction},
  author = {Georgios Balikas},
  journal= {arXiv preprint arXiv:1809.08935},
  year   = {2018}
}

Comments

CAp 2018

R2 v1 2026-06-23T04:16:23.433Z