中文

利用大语言模型增强系统综述:使用 GPT-4 与 Kimi

计算与语言 2025-04-30 v1 应用统计

摘要

本研究深入探讨了两种大语言模型(LLMs)——GPT-4 与 Kimi——在系统综述中的应用。我们通过将 LLM 生成的编码与一项经同行评审的评估类系统综述中由人工生成的编码进行比较,评估了它们的性能。研究发现,LLM 在系统综述中的表现会随数据量和问题复杂度而波动。

关键词

引用

@article{arxiv.2504.20276,
  title  = {Enhancing Systematic Reviews with Large Language Models: Using GPT-4 and Kimi},
  author = {Dandan Chen Kaptur and Yue Huang and Xuejun Ryan Ji and Yanhui Guo and Bradley Kaptur},
  journal= {arXiv preprint arXiv:2504.20276},
  year   = {2025}
}

备注

13 pages, Paper presented at the National Council on Measurement in Education (NCME) Conference, Denver, Colorado, in April 2025