English

Improving Diversity of Commonsense Generation by Large Language Models via In-Context Learning

Computation and Language 2024-09-30 v2

Abstract

Generative Commonsense Reasoning (GCR) requires a model to reason about a situation using commonsense knowledge, while generating coherent sentences. Although the quality of the generated sentences is crucial, the diversity of the generation is equally important because it reflects the model's ability to use a range of commonsense knowledge facts. Large Language Models (LLMs) have shown proficiency in enhancing the generation quality across various tasks through in-context learning (ICL) using given examples without the need for any fine-tuning. However, the diversity aspect in LLM outputs has not been systematically studied before. To address this, we propose a simple method that diversifies the LLM generations, while preserving their quality. Experimental results on three benchmark GCR datasets show that our method achieves an ideal balance between the quality and diversity. Moreover, the sentences generated by our proposed method can be used as training data to improve diversity in existing commonsense generators.

Keywords

Cite

@article{arxiv.2404.16807,
  title  = {Improving Diversity of Commonsense Generation by Large Language Models via In-Context Learning},
  author = {Tianhui Zhang and Bei Peng and Danushka Bollegala},
  journal= {arXiv preprint arXiv:2404.16807},
  year   = {2024}
}

Comments

EMNLP 2024 Findings, Camera-ready version

R2 v1 2026-06-28T16:06:43.065Z