中文

Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?

人工智能 2025-08-28 v1 计算与语言

摘要

最近的工作表明,Chain-of-Thought(CoT)在软推理问题(如分析和常识推理)方面往往仅能获得有限的提升。CoT 也可能对模型实际推理不够忠实。我们探讨了 CoT 在软推理任务中的动力学与忠诚度,涵盖 instruction-tuned、reasoning 和 reasoning-distilled 模型。我们的发现揭示了这些模型依赖 CoT 的差异,表明 CoT 影响力与忠诚度并不总是一致的。

关键词

引用

@article{arxiv.2508.19827,
  title  = {Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?},
  author = {Samuel Lewis-Lim and Xingwei Tan and Zhixue Zhao and Nikolaos Aletras},
  journal= {arXiv preprint arXiv:2508.19827},
  year   = {2025}
}

备注

Accepted at EMNLP 2025 Main Conference