English

An initial attempt of combining visual selective attention with deep reinforcement learning

Machine Learning 2020-06-19 v3 Artificial Intelligence Computer Vision and Pattern Recognition Machine Learning

Abstract

Visual attention serves as a means of feature selection mechanism in the perceptual system. Motivated by Broadbent's leaky filter model of selective attention, we evaluate how such mechanism could be implemented and affect the learning process of deep reinforcement learning. We visualize and analyze the feature maps of DQN on a toy problem Catch, and propose an approach to combine visual selective attention with deep reinforcement learning. We experiment with optical flow-based attention and A2C on Atari games. Experiment results show that visual selective attention could lead to improvements in terms of sample efficiency on tested games. An intriguing relation between attention and batch normalization is also discovered.

Keywords

Cite

@article{arxiv.1811.04407,
  title  = {An initial attempt of combining visual selective attention with deep reinforcement learning},
  author = {Liu Yuezhang and Ruohan Zhang and Dana H. Ballard},
  journal= {arXiv preprint arXiv:1811.04407},
  year   = {2020}
}

Comments

7 pages, 8 figures

R2 v1 2026-06-23T05:11:48.505Z