中文

ChatGPT在韩国数学教学中的鲁棒性研究

人工智能 2025-02-18 v1 历史与综述

摘要

ChatGPT作为一种人工智能模型,有望革新教育。然而,其在解决非英语问题方面的有效性仍不确定。本研究使用586道韩国数学题目评估ChatGPT的鲁棒性。ChatGPT实现了66.72%的准确率,正确回答了586道题中的391道。我们还评估了其根据十一个标准对数学题目进行评分的能力,并进行了主题分析。我们的发现表明,ChatGPT的评分与教育理论和考生视角相一致。尽管ChatGPT在题目分类方面表现良好,但其在非英语语境中仍存在困难,这凸显了需要改进的领域。未来研究应解决语言偏见问题并提高跨语言的准确性。领域特定的优化和多语言训练可以改善ChatGPT在个性化教育中的作用。

关键词

引用

@article{arxiv.2502.11915,
  title  = {On the robustness of ChatGPT in teaching Korean Mathematics},
  author = {Phuong-Nam Nguyen and Quang Nguyen-The and An Vu-Minh and Diep-Anh Nguyen and Xuan-Lam Pham},
  journal= {arXiv preprint arXiv:2502.11915},
  year   = {2025}
}

备注

21 pages, 12 figures, includes statistical analysis of ChatGPT's robustness in solving and rating multilingual mathematics questions. Focus on Korean CSAT Mathematics. Evaluates AI accuracy, rating effectiveness, and topic analysis