English
Related papers

Related papers: Multi-step dual control for exploration and exploi…

200 papers

Bayes-optimal behavior, while well-defined, is often difficult to achieve. Recent advances in the use of Monte-Carlo tree search (MCTS) have shown that it is possible to act near-optimally in Markov Decision Processes (MDPs) with very large…

Artificial Intelligence · Computer Science 2012-02-20 John Asmuth , Michael L. Littman

Satellite observation scheduling plays a significant role in improving the efficiency of Earth observation systems. To solve the large-scale multi-satellite observation scheduling problem, this paper proposes an ensemble of meta-heuristic…

Instrumentation and Methods for Astrophysics · Physics 2024-10-30 Guohua Wu , Qizhang Luo , Xiao Du , Xinwei Wang , Yinguo Chen , Ponnuthurai Nagaratnam Suganthan

This paper is a sequel of our previous work in which we introduced the MapDE algorithm to determine the existence of analytic invertible mappings of an input (source) differential polynomial system (DPS) to a specific target DPS, and…

Analysis of PDEs · Mathematics 2020-01-01 Zahra. Mohammadi , Gregory J. Reid , S. -L. Tracy Huang

This paper proposes a distributed model predicted control (DMPC) approach for consensus control of multi-agent systems (MASs) with linear agent dynamics and bounded control input constraints. Within the proposed DMPC framework, each agent…

Systems and Control · Electrical Eng. & Systems 2020-09-16 Yougang Bian , Changkun Du , Manjiang Hu , Haikuo Liu

In this paper, we propose a parallel multiobjective evolutionary algorithm called Parallel Criterion-based Partitioning MOEA (PCPMOEA), with an application to the Mutliobjective Knapsack Problem (MOKP). The suggested search strategy is…

Optimization and Control · Mathematics 2018-11-07 Kantour Nedjmeddine , Bouroubi Sadek , Chaabane Djamel

Model Predictive Control (MPC) has shown to be a successful method for many applications that require control. Especially in the presence of prediction uncertainty, various types of MPC offer robust or efficient control system behavior. For…

Systems and Control · Electrical Eng. & Systems 2021-06-17 Tim Brüdigam , Jie Zhan , Dirk Wollherr , Marion Leibold

Model predictive control is a well established control technology for trajectory tracking. Its use requires the availability of an accurate model of the plant, but obtaining such a model is often time consuming and costly. Data-Enabled…

Optimization and Control · Mathematics 2025-10-01 Margarita A. Guerrero , Braghadeesh Lakshminarayanan , Cristian R. Rojas

Autonomous exploration is an application of growing importance in robotics. A promising strategy is ergodic trajectory planning, whereby an agent spends in each area a fraction of time which is proportional to its probability information…

Optimization and Control · Mathematics 2021-07-06 Dimitris Gkouletsos , Andrea Iannelli , Mathias Hudoba de Badyn , John Lygeros

Exogenous MDPs (Exo-MDPs) capture sequential decision-making where uncertainty comes solely from exogenous inputs that evolve independently of the learner's actions. This structure is especially common in operations research applications…

Machine Learning · Computer Science 2026-01-29 Hao Liang , Jiayu Cheng , Sean R. Sinclair , Yali Du

In this paper, we consider a modified version of the control problem in a model free Markov decision process (MDP) setting with large state and action spaces. The control problem most commonly addressed in the contemporary literature is to…

Artificial Intelligence · Computer Science 2018-02-01 Ajin George Joseph , Shalabh Bhatnagar

In trajectory optimization, Model Predictive Path Integral (MPPI) control is a sampling-based Model Predictive Control (MPC) framework that generates optimal inputs by efficiently simulating numerous trajectories. In practice, however, MPPI…

Systems and Control · Electrical Eng. & Systems 2025-02-21 Fanxin Wang , Yikun Cheng , Chuyuan Tao

This thesis focuses on developing advanced control methods for two industrial systems in discrete-time aiming to enhance their performance in delivering the control objectives as well as considering the practical aspects. The first part…

Systems and Control · Computer Science 2015-07-31 Arash Khatamianfar

In dynamic and resource-constrained environments, such as multi-hop wireless mesh networks, traditional routing protocols often falter by relying on predetermined paths that prove ineffective in unpredictable link conditions. Shortest…

Networking and Internet Architecture · Computer Science 2026-02-02 Narjes Nourzad , Bhaskar Krishnamachari

In this article, we propose a sampling-based motion planning algorithm equipped with an information-theoretic convergence criterion for incremental informative motion planning. The proposed approach allows dense map representations and…

Robotics · Computer Science 2019-05-24 Maani Ghaffari Jadidi , Jaime Valls Miro , Gamini Dissanayake

This paper introduces COR-MCTS (Conservation of Resources - Monte Carlo Tree Search), a novel tactical decision-making approach for automated driving focusing on maneuver planning over extended horizons. Traditional decision-making…

Robotics · Computer Science 2025-04-23 Karim Essalmi , Fernando Garrido , Fawzi Nashashibi

A common way to simulate the transport and spread of pollutants in the atmosphere is via stochastic Lagrangian dispersion models. Mathematically, these models describe turbulent transport processes with stochastic differential equations…

In recent years, autonomous underwater vehicle (AUV) systems have demonstrated significant potential in complex marine exploration. However, effective AUV-based tracking remains challenging in realistic underwater environments characterized…

Networking and Internet Architecture · Computer Science 2026-02-10 Kai Tian , Chuan Lin , Guangjie Han , Chen An , Qian Zhu , Shengzhao Zhu , Zhenyu Wang

A distributed implementation of a Robust Integral of the Sign of the Error (RISE) controller is developed for multi-agent target tracking problems with exponential convergence guarantees. Previous RISE-based approaches for multi-agent…

Systems and Control · Electrical Eng. & Systems 2025-06-02 Cristian F. Nino , Omkar Sudhir Patil , Sage C. Edwards , Warren E. Dixon

Dynamic multiobjective optimization problems (DMOPs) feature time-varying objectives, which cause the Pareto optimal solution (POS) set to drift over time and make it difficult to maintain both convergence and diversity under limited…

Neural and Evolutionary Computing · Computer Science 2026-03-31 Jian Guan , Huolong Wu , Zhenzhong Wang , Gary G. Yen , Min Jiang

We address the challenge of effective exploration while maintaining good performance in policy gradient methods. As a solution, we propose diverse exploration (DE) via conjugate policies. DE learns and deploys a set of conjugate policies…

Machine Learning · Computer Science 2019-02-12 Andrew Cohen , Xingye Qiao , Lei Yu , Elliot Way , Xiangrong Tong