English

Chain of Thought Still Thinks Fast: APriCoT Helps with Thinking Slow

Computation and Language 2025-08-12 v3 Artificial Intelligence

Abstract

Language models are known to absorb biases from their training data, leading to predictions driven by statistical regularities rather than semantic relevance. We investigate the impact of these biases on answer choice preferences in the Massive Multi-Task Language Understanding (MMLU) task. Our findings show that these biases are predictive of model preference and mirror human test-taking strategies even when chain of thought (CoT) reasoning is used. To address this issue, we introduce Counterfactual Prompting with Agnostically Primed CoT (APriCoT). We demonstrate that while Counterfactual Prompting with CoT alone is insufficient to mitigate bias, APriCoT effectively reduces the influence of base-rate probabilities while improving overall accuracy. Our results suggest that mitigating bias requires a slow thinking process which CoT alone may not provide as it tends to reinforce fast thinking model bias under some prompting methodologies. APriCoT is a step toward developing more robust and fair language models that can think slow.

Keywords

Cite

@article{arxiv.2408.08651,
  title  = {Chain of Thought Still Thinks Fast: APriCoT Helps with Thinking Slow},
  author = {Kyle Moore and Jesse Roberts and Thao Pham and Douglas Fisher},
  journal= {arXiv preprint arXiv:2408.08651},
  year   = {2025}
}

Comments

Final version. Published In Proceedings of the Annual Meeting of the Cognitive Science Society (Vol. 47) 2025