English

CDRL: A Reinforcement Learning Framework Inspired by Cerebellar Circuits and Dendritic Computational Strategies

Machine Learning 2026-02-18 v1 Artificial Intelligence Neural and Evolutionary Computing

Abstract

Reinforcement learning (RL) has achieved notable performance in high-dimensional sequential decision-making tasks, yet remains limited by low sample efficiency, sensitivity to noise, and weak generalization under partial observability. Most existing approaches address these issues primarily through optimization strategies, while the role of architectural priors in shaping representation learning and decision dynamics is less explored. Inspired by structural principles of the cerebellum, we propose a biologically grounded RL architecture that incorporate large expansion, sparse connectivity, sparse activation, and dendritic-level modulation. Experiments on noisy, high-dimensional RL benchmarks show that both the cerebellar architecture and dendritic modulation consistently improve sample efficiency, robustness, and generalization compared to conventional designs. Sensitivity analysis of architectural parameters suggests that cerebellum-inspired structures can offer optimized performance for RL with constrained model parameters. Overall, our work underscores the value of cerebellar structural priors as effective inductive biases for RL.

Keywords

Cite

@article{arxiv.2602.15367,
  title  = {CDRL: A Reinforcement Learning Framework Inspired by Cerebellar Circuits and Dendritic Computational Strategies},
  author = {Sibo Zhang and Rui Jing and Liangfu Lv and Jian Zhang and Yunliang Zang},
  journal= {arXiv preprint arXiv:2602.15367},
  year   = {2026}
}

Comments

14pages, 8 figures, 6 tabels

R2 v1 2026-07-01T10:39:33.048Z