English

Are Large Language Models Capable of Generating Human-Level Narratives?

Computation and Language 2024-10-08 v2

Abstract

This paper investigates the capability of LLMs in storytelling, focusing on narrative development and plot progression. We introduce a novel computational framework to analyze narratives through three discourse-level aspects: i) story arcs, ii) turning points, and iii) affective dimensions, including arousal and valence. By leveraging expert and automatic annotations, we uncover significant discrepancies between the LLM- and human- written stories. While human-written stories are suspenseful, arousing, and diverse in narrative structures, LLM stories are homogeneously positive and lack tension. Next, we measure narrative reasoning skills as a precursor to generative capacities, concluding that most LLMs fall short of human abilities in discourse understanding. Finally, we show that explicit integration of aforementioned discourse features can enhance storytelling, as is demonstrated by over 40% improvement in neural storytelling in terms of diversity, suspense, and arousal.

Keywords

Cite

@article{arxiv.2407.13248,
  title  = {Are Large Language Models Capable of Generating Human-Level Narratives?},
  author = {Yufei Tian and Tenghao Huang and Miri Liu and Derek Jiang and Alexander Spangher and Muhao Chen and Jonathan May and Nanyun Peng},
  journal= {arXiv preprint arXiv:2407.13248},
  year   = {2024}
}

Comments

EMNLP 2024

R2 v1 2026-06-28T17:45:35.816Z