English

Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning

Computation and Language 2025-12-30 v2

Abstract

Large language models (LLMs) show strong reasoning via chain-of-thought (CoT) prompting, but the process is opaque, which makes verification, debugging, and control difficult in high-stakes settings. We present Vis-CoT, a human-in-the-loop framework that converts linear CoT text into an interactive reasoning graph. Users can visualize the logical flow, identify flawed steps, and intervene by pruning incorrect paths and grafting new, user-defined premises. This shifts interaction from passive observation to active collaboration, steering models toward more accurate and trustworthy conclusions. Across GSM8K and StrategyQA, Vis-CoT improves final-answer accuracy by up to 24 percentage points over non-interactive baselines. A user study also shows large gains in perceived usability and trust. Vis-CoT points to a practical path for more reliable, understandable, and collaborative reasoning by combining LLMs with targeted human oversight.

Keywords

Cite

@article{arxiv.2509.01412,
  title  = {Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning},
  author = {Kaviraj Pather and Elena Hadjigeorgiou and Arben Krasniqi and Claire Schmit and Irina Rusu and Marc Pons and Kabir Khan},
  journal= {arXiv preprint arXiv:2509.01412},
  year   = {2025}
}

Comments

12 pages, 7 figures

R2 v1 2026-07-01T05:15:16.378Z