中文
相关论文

相关论文: Robust Policy Selection and Harvest Risk Quantific…

200 篇论文

Learning to make decisions from observed data in dynamic environments remains a problem of fundamental importance in a number of fields, from artificial intelligence and robotics, to medicine and finance. This paper concerns the problem of…

机器学习 · 统计学 2018-06-04 Jack Umenberger , Thomas B. Schön

This paper studies a risk-sensitive decision-making problem under uncertainty. It considers a decision-making process that unfolds over a fixed number of stages, in which a decision-maker chooses among multiple alternatives, some of which…

最优化与控制 · 数学 2026-01-07 Chung-Han Hsieh , Yi-Shan Wong

Robust control is a core approach for controlling systems with performance guarantees that are robust to modeling error, and is widely used in real-world systems. However, current robust control approaches can only handle small system…

最优化与控制 · 数学 2021-06-08 Dimitar Ho , Hoang M. Le , John C. Doyle , Yisong Yue

We study data-driven learning of robust stochastic control for infinite-horizon systems with potentially continuous state and action spaces. In many managerial settings--supply chains, finance, manufacturing, services, and dynamic…

机器学习 · 统计学 2025-11-18 Shengbo Wang , Jason Meng , Nian Si , Jose Blanchet , Zhengyuan Zhou

Robust control of mechanical systems with multiple uncertainties remains a fundamental challenge, particularly when nonlinear dynamics and operating-condition variations are intricately intertwined. Although deep reinforcement learning…

机器学习 · 计算机科学 2026-03-11 Heisei Yonezawa , Ansei Yonezawa , Itsuro Kajiwara

This paper concerns discrete-time infinite-horizon stochastic control systems with Borel state and action spaces and universally measurable policies. We study optimization problems on strategic measures induced by the policies in these…

最优化与控制 · 数学 2023-12-22 Huizhen Yu

We study the singular stochastic optimal control problem with model uncertainty, where the necessary conditions determined by the corresponding maximum principle are trivial. Robust integral form and pointwise second order necessary…

最优化与控制 · 数学 2025-12-09 Guangdong Jing

Randomized optimization is an established tool for control design with modulated robustness. While for uncertain convex programs there exist randomized approaches with efficient sampling, this is not the case for non-convex problems.…

系统与控制 · 计算机科学 2015-06-08 Sergio Grammatico , Xiaojing Zhang , Kostas Margellos , Paul Goulart , John Lygeros

Capacity expansion models used for policy support have increasingly represented both the variability and uncertainty of weather-dependent generation (wind and solar). However, although also uncertain, as demonstrated by the performance of…

最优化与控制 · 数学 2025-04-14 Kamran Forghani , Xiaoming Kan , Lina Reichenberg , Fredrik Hedenus

This paper studies an $\alpha$-robust utility maximization problem where an investor faces an intractable claim -- an exogenous contingent claim with known marginal distribution but unspecified dependence structure with financial market…

投资组合管理 · 定量金融 2026-04-07 Xinyu Chen , Zuo Quan Xu

Mine planning is a complex task that involves many uncertainties. During early stage feasibility, available mineral resources can only be estimated based on limited sampling of ore grades from sparse drilling, leading to large uncertainty…

神经与进化计算 · 计算机科学 2023-05-30 Michael Stimson , William Reid , Aneta Neumann , Simon Ratcliffe , Frank Neumann

Uncertainty sampling is a prevalent active learning algorithm that queries sequentially the annotations of data samples which the current prediction model is uncertain about. However, the usage of uncertainty sampling has been largely…

机器学习 · 计算机科学 2026-04-08 Shang Liu , Xiaocheng Li

An important problem in sequential decision-making under uncertainty is to use limited data to compute a safe policy, i.e., a policy that is guaranteed to perform at least as well as a given baseline strategy. In this paper, we develop and…

机器学习 · 统计学 2016-07-14 Marek Petrik , Yinlam Chow , Mohammad Ghavamzadeh

We study control of constrained linear systems with only partial statistical information about the uncertainty affecting the system dynamics and the sensor measurements. Specifically, given a finite collection of disturbance realizations…

This paper presents a strictly convex chance-constrained stochastic control framework that accounts for uncertainty in control specifications such as reference trajectories and operational constraints. By jointly optimizing control inputs…

系统与控制 · 电气工程与系统科学 2026-01-27 Teruki Kato , Ryotaro Shima , Kenji Kashima

We consider a distributionally robust formulation of stochastic optimization problems arising in statistical learning, where robustness is with respect to uncertainty in the underlying data distribution. Our formulation builds on…

最优化与控制 · 数学 2021-06-09 Mert Gürbüzbalaban , Andrzej Ruszczyński , Landi Zhu

This paper investigates the dynamics and optimal harvesting of age-structured populations governed by McKendrick--von Foerster equations, contrasting two distinct harvesting mechanisms: rate-control and effort-control. For the rate-control…

最优化与控制 · 数学 2026-04-03 Jiguang Yu , Louis Shuo Wang , Ye Liang

Policy gradient methods are widely used in reinforcement learning. Yet, the nonconvexity of policy optimization poses significant challenges in understanding the global convergence of policy gradient methods. For a class of finite-horizon…

最优化与控制 · 数学 2026-03-10 Xin Chen , Yifan Hu , Minda Zhao

Balancing the trade-off between safety and efficiency is of significant importance for path planning under uncertainty. Many risk-aware path planners have been developed to explicitly limit the probability of collision to an acceptable…

机器人学 · 计算机科学 2022-10-26 Fei Meng , Liangliang Chen , Han Ma , Jiankun Wang , Max Q. -H. Meng

This paper studies optimal control problems of unknown linear systems subject to stochastic disturbances of uncertain distribution. Uncertainty about the stochastic disturbances is usually described via ambiguity sets of probability…

系统与控制 · 电气工程与系统科学 2023-06-30 Guanru Pan , Timm Faulwasser