English

Explore Beyond the Boundary Using Entropic Information

Machine Learning 2026-07-31 v1 Artificial Intelligence

Abstract

In reinforcement learning, exploration with sparse and delayed rewards presents a significant challenge due to the limited feedback available for guiding the learning process. Addressing this issue requires extensive exploration in the state space to discover valuable reward signals. In this paper, we propose Entropic Information for Exploration (ENTINEX), a novel method that enhances exploration by incentivizing agents to explore beyond the boundaries of the state distribution. ENTINEX achieves this by assigning intrinsic rewards to these boundaries, leveraging entropic information to identify them effectively. Through extensive experimentation, we demonstrate that ENTINEX consistently improves exploration performance in environments characterized by sparse and delayed rewards. Our experimental results show that ENTINEX outperforms existing exploration methods, highlighting its effectiveness in both sparse and delayed reward scenarios.

Cite

@article{arxiv.2607.29419,
  title  = {Explore Beyond the Boundary Using Entropic Information},
  author = {Bumgeun Park and Donghwan Lee},
  journal= {arXiv preprint arXiv:2607.29419},
  year   = {2026}
}