中文
相关论文

相关论文: Multi-step dual control for exploration and exploi…

200 篇论文

We present a sequential distributed model predictive control (MPC) scheme for cooperative control of multi-agent systems with dynamically decoupled heterogeneous nonlinear agents subject to individual constraints. In the scheme, we explore…

系统与控制 · 电气工程与系统科学 2024-06-13 Matthias Köhler , Matthias A. Müller , Frank Allgöwer

The emerging research paradigm coined as multitasking optimization aims to solve multiple optimization tasks concurrently by means of a single search process. For this purpose, the exploitation of complementarities among the tasks to be…

人工智能 · 计算机科学 2020-05-14 Eneko Osaba , Aritz D. Martinez , Akemi Galvez , Andres Iglesias , Javier Del Ser

Policy Search and Model Predictive Control~(MPC) are two different paradigms for robot control: policy search has the strength of automatically learning complex policies using experienced data, while MPC can offer optimal control…

机器人学 · 计算机科学 2021-12-17 Yunlong Song , Davide Scaramuzza

The integration of autonomous vehicles into urban and highway environments necessitates the development of robust and adaptable behavior planning systems. This study presents an innovative approach to address this challenge by utilizing a…

机器人学 · 计算机科学 2023-10-19 Qianfeng Wen , Zhongyi Gong , Lifeng Zhou , Zhongshun Zhang

In distributed model predictive control (DMPC), where a centralized optimization problem is solved in distributed fashion using dual decomposition, it is important to keep the number of iterations in the solution algorithm, i.e. the amount…

最优化与控制 · 数学 2013-07-11 Pontus Giselsson , Anders Rantzer

This paper investigates goal-oriented communication for remote estimation of multiple Markov sources in resource-constrained networks. An agent decides the updating times of the sources and transmits the packet to a remote destination over…

系统与控制 · 电气工程与系统科学 2024-06-04 Jiping Luo , Nikolaos Pappas

Maximum Causal Entropy (MCE) Inverse Optimal Control (IOC) has become an effective tool for modelling human behaviour in many control tasks. Its advantage over classic techniques for estimating human policies is the transferability of the…

系统与控制 · 计算机科学 2016-07-20 Felix Schmitt , Hans-Joachim Bieg , Dietrich Manstetten , Michael Herman , Rainer Stiefelhagen

In this paper, we propose a simple yet efficient strategy for improving the multi-objective steepest descent method proposed by Fliege and Svaiter (Math Methods Oper Res, 2000, 3: 479--494). The core idea behind this strategy involves…

最优化与控制 · 数学 2024-01-15 Wang Chen , Liping Tang , Xinmin Yang

Within the framework of probably approximately correct Markov decision processes (PAC-MDP), much theoretical work has focused on methods to attain near optimality after a relatively long period of learning and exploration. However,…

人工智能 · 计算机科学 2016-04-06 Kenji Kawaguchi

Collision-free flight in cluttered environments is a critical capability for autonomous quadrotors. Traditional methods often rely on detailed 3D map construction, trajectory generation, and tracking. However, this cascade pipeline can…

机器人学 · 计算机科学 2025-08-12 Linzuo Zhang , Yu Hu , Yang Deng , Feng Yu , Danping Zou

Monte Carlo Tree Search (MCTS), most famously used in game-play artificial intelligence (e.g., the game of Go), is a well-known strategy for constructing approximate solutions to sequential decision problems. Its primary innovation is the…

最优化与控制 · 数学 2017-04-21 Daniel R. Jiang , Lina Al-Kanj , Warren B. Powell

This paper presents the development and evaluation of an optimization-based autonomous trajectory planning algorithm for the asteroid reconnaissance phase of a deep-space exploration mission. The reconnaissance phase is a low-altitude flyby…

In this paper we investigate the convergence of the Policy Iteration Algorithm (PIA) for a class of general continuous-time entropy-regularized stochastic control problems. In particular, instead of employing sophisticated PDE estimates for…

最优化与控制 · 数学 2025-04-24 Jin Ma , Gaozhan Wang , Jianfeng Zhang

This paper proposes an iterative methodology to integrate large-scale behavioral activity-based models with dynamic traffic assignment models. The main novelty of the proposed approach is the decoupling of the two parts, allowing the…

计算机与社会 · 计算机科学 2024-04-12 Serio Agriesti , Claudio Roncoli , Bat-hen Nahmias-Biran

Bayesian Decision Trees (DTs) are generally considered a more advanced and accurate model than a regular Decision Tree (DT) because they can handle complex and uncertain data. Existing work on Bayesian DTs uses Markov Chain Monte Carlo…

机器学习 · 计算机科学 2023-05-31 Efthyvoulos Drousiotis , Alexander M. Phillips , Paul G. Spirakis , Simon Maskell

In the field of model predictive control, Data-enabled Predictive Control (DeePC) offers direct predictive control, bypassing traditional modeling. However, challenges emerge with increased computational demand due to recursive data…

系统与控制 · 电气工程与系统科学 2024-03-26 Jicheng Shi , Yingzhao Lian , Colin N. Jones

In the trajectory planning of automated driving, data-driven statistical artificial intelligence (AI) methods are increasingly established for predicting the emergent behavior of other road users. While these methods achieve exceptional…

机器人学 · 计算机科学 2025-04-28 Lars Ullrich , Zurab Mujirishvili , Knut Graichen

Classical Distributed Model Predictive Control (DiMPC) requires multiple iterations to achieve convergence, leading to high computational and communication burdens. This work focuses on the improvement of an iteration-free distributed MPC…

最优化与控制 · 数学 2026-04-03 Parth R. Brahmbhatt , Hari S. Ganesh , Styliani Avraamidou

To coordinate the economy, security and environment protection in the power system operation, a two-step many-objective optimal power flow (MaOPF) solution method is proposed. In step 1, it is the first time that knee point-driven…

最优化与控制 · 数学 2018-12-05 Yahui Li , Yang Li

This paper presents a distributed stochastic model predictive control (SMPC) approach for large-scale linear systems with private and common uncertainties in a plug-and-play framework. Using the so-called scenario approach, the centralized…

最优化与控制 · 数学 2019-01-09 V. Rostampour , T. Keviczky