中文
相关论文

相关论文: Robust Policy Selection and Harvest Risk Quantific…

200 篇论文

Consider a species whose population density solves the steady diffusive logistic equation in a heterogeneous environment modeled with the help of a spatially non constant coefficient standing for a resources distribution in a given box. We…

偏微分方程分析 · 数学 2018-07-25 Idriss Mazari , Grégoire Nadin , Yannick Privat

We consider risk-sensitive Markov decision processes (MDPs), where the MDP model is influenced by a parameter which takes values in a compact metric space. We identify sufficient conditions under which small perturbations in the model…

最优化与控制 · 数学 2022-09-28 Shiping Shao , Abhishek Gupta , William B. Haskell

We study the problem of model selection in batch policy optimization: given a fixed, partial-feedback dataset and $M$ model classes, learn a policy with performance that is competitive with the policy derived from the best model class. We…

机器学习 · 计算机科学 2021-12-24 Jonathan N. Lee , George Tucker , Ofir Nachum , Bo Dai

In stochastic control applications, typically only an ideal model (controlled transition kernel) is assumed and the control design is based on the given model, raising the problem of performance loss due to the mismatch between the assumed…

系统与控制 · 计算机科学 2020-02-04 Ali Devran Kara , Serdar Yüksel

We introduce a framework for quantifying propagation of uncertainty arising in a dynamic setting. Specifically, we define dynamic uncertainty sets designed explicitly for discrete stochastic processes over a finite time horizon. These…

风险管理 · 定量金融 2024-02-05 Marlon Moresco , Mélina Mailhot , Silvana M. Pesenti

We develop a neural-network framework for multi-period risk--reward stochastic control problems with constrained two-step feedback policies that may be discontinuous in the state. We allow a broad class of objectives built on a…

计算金融 · 定量金融 2026-03-09 Chang Chen , Duy-Minh Dang

This paper studies a robust utility maximization problem for intractable claims under distributional ambiguity, where the distribution of the claim cannot be inferred from market information and its dependence with tradable assets is…

最优化与控制 · 数学 2026-04-17 Guohui Guan , Zongxia Liang , Xingjian Ma

Numerically computing global policies to optimal control problems for complex dynamical systems is mostly intractable. In consequence, a number of approximation methods have been developed. However, none of the current methods can quantify…

机器人学 · 计算机科学 2021-03-05 Ashwin Khadke , Hartmut Geyer

Balancing safety and efficiency when planning in crowded scenarios with uncertain dynamics is challenging where it is imperative to accomplish the robot's mission without incurring any safety violations. Typically, chance constraints are…

机器人学 · 计算机科学 2023-02-22 Khaled A. Mustafa , Oscar de Groot , Xinwei Wang , Jens Kober , Javier Alonso-Mora

Model based predictions of future trajectories of a dynamical system often suffer from inaccuracies, forcing model based control algorithms to re-plan often, thus being computationally expensive, suboptimal and not reliable. In this work,…

机器学习 · 计算机科学 2018-12-11 Norman Di Palo , Harri Valpola

The ability to compute reward-optimal policies for given and known finite Markov decision processes (MDPs) underpins a variety of applications across planning, controller synthesis, and verification. However, we often want policies (1) to…

计算机科学中的逻辑 · 计算机科学 2025-11-18 Linus Heck , Filip Macák , Milan Češka , Sebastian Junges

In robust combinatorial optimization, we would like to find a solution that performs well under all realizations of an uncertainty set of possible parameter values. How we model this uncertainty set has a decisive influence on the…

最优化与控制 · 数学 2024-04-30 Marc Goerigk , Mohammad Khosravi

Most work in mechanism design assumes that buyers are risk neutral; some considers risk aversion arising due to a non-linear utility for money. Yet behavioral studies have established that real agents exhibit risk attitudes which cannot be…

计算机科学与博弈论 · 计算机科学 2018-03-13 Shuchi Chawla , Kira Goldner , J. Benjamin Miller , Emmanouil Pountourakis

We consider the maximization problem of monotone submodular functions under an uncertain knapsack constraint. Specifically, the problem is discussed in the situation that the knapsack capacity is not given explicitly and can be accessed…

数据结构与算法 · 计算机科学 2018-03-08 Yasushi Kawase , Hanna Sumita , Takuro Fukunaga

This paper is concerned with the maximum principle of stochastic optimal control problems, where the coefficients of the state equation and the cost functional are uncertain, and the system is generally under Markovian regime switching.…

最优化与控制 · 数学 2025-04-15 Tao Hao , Jiaqiang Wen , Jie Xiong

We consider a discrete time stochastic Markovian control problem under model uncertainty. Such uncertainty not only comes from the fact that the true probability law of the underlying stochastic process is unknown, but the parametric family…

最优化与控制 · 数学 2022-03-23 Erhan Bayraktar , Tao Chen

Model-based reinforcement learning (MBRL) has demonstrated superior sample efficiency compared to model-free reinforcement learning (MFRL). However, the presence of inaccurate models can introduce biases during policy learning, resulting in…

机器学习 · 计算机科学 2025-03-27 Yongshuai Liu , Xin Liu

This paper proposes an analytical framework for modelling resource contention in multi-robot systems, where the travel times and task durations are uncertain. It uses several approximation methods to quickly and accurately calculate the…

多智能体系统 · 计算机科学 2020-03-17 Andrew W. Palmer , Andrew J. Hill , Steven J. Scheding

We consider a robust approach to address uncertainty in model parameters in Markov Decision Processes (MDPs), which are widely used to model dynamic optimization in many applications. Most prior works consider the case where the uncertainty…

最优化与控制 · 数学 2021-09-02 Vineet Goyal , Julien Grand-Clément

As the penetration level of transmission-scale time-intermittent renewable generation resources increases, control of flexible resources will become important to mitigating the fluctuations due to these new renewable resources. Flexible…

最优化与控制 · 数学 2013-09-12 Krishnamurthy Dvijotham , Scott Backhaus , Misha Chertkov
‹ 上一页 1 8 9 10 下一页 ›