English

A Study on Question-Answer Dataset for LLM Safety Evaluation with a Focus on Illegal Activities

Computation and Language 2026-05-29 v1

Abstract

In this paper, we discuss question-answer dataset for LLM safety evaluation, with a focus on illegal activities. Specifically, on the basis of manual analysis of AnswerCarefully, we introduce several additional information, methods for creating question-answer examples, and a rubric for evaluating LLM-generated responses. The outcomes of this study are intended to be shared with the "JAI-Trust" project.

Keywords

Cite

@article{arxiv.2605.29340,
  title  = {A Study on Question-Answer Dataset for LLM Safety Evaluation with a Focus on Illegal Activities},
  author = {Kenji Imamura and Masao Ideuchi and Atsushi Fujita},
  journal= {arXiv preprint arXiv:2605.29340},
  year   = {2026}
}

Comments

10 pages, 1 figure

R2 v1 2026-07-22T07:38:39.724Z