English
Related papers

Related papers: Centric selection: a way to tune the exploration/e…

200 papers

The inherent trade-off in on-line learning is between exploration and exploitation. A good balance between these two (conflicting) goals can achieve a better long-term performance. Can we define an optimal balance? We propose to study this…

Information Theory · Computer Science 2024-11-06 Ram Zamir , Kenneth Rose

In this paper, we study the influence of the selective pressure on the performance of cellular genetic algorithms. Cellular genetic algorithms are genetic algorithms where the population is embedded on a toroidal grid. This structure makes…

Artificial Intelligence · Computer Science 2008-12-18 David Simoncini , Philippe Collard , Sébastien Verel , Manuel Clergue

The Exploration-Exploitation tradeoff arises in Reinforcement Learning when one cannot tell if a policy is optimal. Then, there is a constant need to explore new actions instead of exploiting past experience. In practice, it is common to…

Machine Learning · Computer Science 2019-09-10 Lior Shani , Yonathan Efroni , Shie Mannor

Finding a good compromise between the exploitation of known resources and the exploration of unknown, but potentially more profitable choices, is a general problem, which arises in many different scientific disciplines. We propose a…

Disordered Systems and Neural Networks · Physics 2016-10-28 Thomas Gueudré , Alexander Dobrinevski , Jean-Philippe Bouchaud

We study a minimal model for the growth of a phenotypically heterogeneous population of cells subject to a fluctuating environment in which they can replicate (by exploiting available resources) and modify their phenotype within a given…

Populations and Evolution · Quantitative Biology 2019-01-23 Andrea De Martino , Thomas Gueudré , Mattia Miotto

In this paper we introduce a new selection scheme in cellular genetic algorithms (cGAs). Anisotropic Selection (AS) promotes diversity and allows accurate control of the selective pressure. First we compare this new scheme with the…

Artificial Intelligence · Computer Science 2008-02-19 David Simoncini , Sébastien Verel , Philippe Collard , Manuel Clergue

The exploration-exploitation (EE) trade-off is a central challenge in reinforcement learning (RL) for large language models (LLMs). With Group Relative Policy Optimization (GRPO), training tends to be exploitation driven: entropy decreases…

Machine Learning · Computer Science 2026-01-21 Zhaochun Li , Chen Wang , Jionghao Bai , Shisheng Cui , Ge Lan , Zhou Zhao , Yue Wang

The tradeoff between accuracy and speed is considered fundamental to individual and collective decision-making. In this paper, we focus on collective estimation as an example of collective decision-making. The task is to estimate the…

Multiagent Systems · Computer Science 2022-01-19 Mohsen Raoufi , Heiko Hamann , Pawel Romanczuk

The balance of exploration versus exploitation (EvE) is a key issue on evolutionary computation. In this paper we will investigate how an adaptive controller aimed to perform Operator Selection can be used to dynamically manage the EvE…

Neural and Evolutionary Computing · Computer Science 2014-09-08 Giacomo di Tollo , Frédéric Lardeux , Jorge Maturana , Frédéric Saubion

We develop a probabilistic framework for analysing model-based reinforcement learning in the episodic setting. We then apply it to study finite-time horizon stochastic control problems with linear dynamics but unknown coefficients and…

Machine Learning · Computer Science 2021-12-22 Lukasz Szpruch , Tanut Treetanthiploet , Yufei Zhang

Evolutionary computation (EC) algorithms, renowned as powerful black-box optimizers, leverage a group of individuals to cooperatively search for the optimum. The exploration-exploitation tradeoff (EET) plays a crucial role in EC, which,…

Neural and Evolutionary Computing · Computer Science 2024-04-15 Zeyuan Ma , Jiacheng Chen , Hongshu Guo , Yining Ma , Yue-Jiao Gong

The lack of diversity in a genetic algorithm's population may lead to a bad performance of the genetic operators since there is not an equilibrium between exploration and exploitation. In those cases, genetic algorithms present a fast and…

Artificial Intelligence · Computer Science 2017-02-14 Andrés Herrera-Poyatos , Francisco Herrera

In this paper, we develop a dynamic exploration/ exploitation (exr/exp) strategy for contextual recommender systems (CRS). Specifically, our methods can adaptively balance the two aspects of exr/exp by automatically learning the optimal…

Information Retrieval · Computer Science 2014-04-16 Djallel Bouneffouf

We present a novel dual control strategy for uncertain linear systems based on targeted harmonic exploration and gain-scheduling with performance and excitation guarantees. In the proposed sequential approach, robust control is implemented…

Systems and Control · Electrical Eng. & Systems 2024-07-30 Janani Venkatasubramanian , Johannes Köhler , Julian Berberich , Frank Allgöwer

Known as two cornerstones of problem solving by search, exploitation and exploration are extensively discussed for implementation and application of evolutionary algorithms (EAs). However, only a few researches focus on evaluation and…

Neural and Evolutionary Computing · Computer Science 2020-01-30 Yu Chen , Jun He

The number of ribosomes in a cell is considered as limiting, and gene expression is thus largely determined by their cellular concentration. In this work we develop a toy model to study the trade-off between the ribosomal supply and the…

Subcellular Processes · Quantitative Biology 2020-02-19 Pascal S. Rogalla , Timothy J. Rudge , Luca Ciandrini

We develop a coherent framework for integrative simultaneous analysis of the exploration-exploitation and model order selection trade-offs. We improve over our preceding results on the same subject (Seldin et al., 2011) by combining…

Machine Learning · Computer Science 2015-03-19 Yevgeny Seldin , Nicolò Cesa-Bianchi , François Laviolette , Peter Auer , John Shawe-Taylor , Jan Peters

A key challenge in reinforcement learning (RL) is managing the exploration-exploitation trade-off without sacrificing sample efficiency. Policy gradient (PG) methods excel in exploitation through fine-grained, gradient-based optimization…

Machine Learning · Computer Science 2025-04-18 Zelal Su "Lain" Mustafaoglu , Keshav Pingali , Risto Miikkulainen

Cultures around the world show varying levels of conservatism. While maintaining traditional ideas prevents wrong ones from being embraced, it also slows or prevents adaptation to new times. Without exploration there can be no improvement,…

Populations and Evolution · Quantitative Biology 2023-04-17 Brian Mintz , Feng Fu

Genetic information and environmental factors determine the path of an individuals life and therefore, the evolution of its entire species. We have succeeded in proposing and studying a model that captures this idea. In our model, a…

Populations and Evolution · Quantitative Biology 2021-07-28 N. Abadi , G. Abramson
‹ Prev 1 2 3 10 Next ›