Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?
人工智能
2025-08-28 v1 计算与语言
摘要
最近的工作表明,Chain-of-Thought(CoT)在软推理问题(如分析和常识推理)方面往往仅能获得有限的提升。CoT 也可能对模型实际推理不够忠实。我们探讨了 CoT 在软推理任务中的动力学与忠诚度,涵盖 instruction-tuned、reasoning 和 reasoning-distilled 模型。我们的发现揭示了这些模型依赖 CoT 的差异,表明 CoT 影响力与忠诚度并不总是一致的。
引用
@article{arxiv.2508.19827,
title = {Analysing Chain of Thought Dynamics: Active Guidance or Unfaithful Post-hoc Rationalisation?},
author = {Samuel Lewis-Lim and Xingwei Tan and Zhixue Zhao and Nikolaos Aletras},
journal= {arXiv preprint arXiv:2508.19827},
year = {2025}
}
备注
Accepted at EMNLP 2025 Main Conference