中文
相关论文

相关论文: Particle Approximation for Conditional Control wit…

200 篇论文

We consider a stochastic linear system and address the design of a finite horizon control policy that is optimal according to some average cost criterion and accounts also for probabilistic constraints on both the input and state variables.…

最优化与控制 · 数学 2016-10-21 Luca Deori , Simone Garatti , Maria Prandini

We consider the dynamics of particle systems where the particles are confined by impenetrable barriers to a bounded, possibly non-convex domain $\Omega$. When particles hit the boundary, we consider an instant change in velocity, which…

偏微分方程分析 · 数学 2018-12-24 M. Kimura , P. van Meurs , Z. X. Yang

In this paper, we investigate an optimal control problem with terminal stochastic linear complementarity constraints (SLCC), and its discrete approximation using the relaxation, the sample average approximation (SAA) and the implicit Euler…

最优化与控制 · 数学 2022-08-17 Jianfeng Luo , Xiaojun Chen

Safe reinforcement learning aims to learn the optimal policy while satisfying safety constraints, which is essential in real-world applications. However, current algorithms still struggle for efficient policy updates with hard constraint…

机器学习 · 计算机科学 2022-06-20 Linrui Zhang , Li Shen , Long Yang , Shixiang Chen , Bo Yuan , Xueqian Wang , Dacheng Tao

We consider stochastic model predictive control of a multi-agent systems with constraints on the probabilities of inter-agent collisions. We first study a sample-based approximation of the collision probabilities and use this approximation…

系统与控制 · 计算机科学 2011-08-17 Daniel Lyons , Jan-P. Calliess , Uwe D. Hanebeck

We describe a general strategy for sampling configurations from a given (Gibbs-Boltzmann or other) distribution. It is {\it not} based on the Metropolis concept of establishing a Markov process whose stationary state is the wanted…

统计力学 · 物理学 2007-05-23 P. Grassberger , W. Nadler

In this paper we consider time-optimal control problems for systems with backlash. Such systems are described by second order differential equations coupled with restrictions modeling the inelastic shocks. A main feature of such systems is…

最优化与控制 · 数学 2024-01-17 Maria do Rosário de Pinho , Maria Margarida Amorim Ferreira , Georgi Smirnov

We consider a framework for solving optimal liquidation problems in limit order books. In particular, order arrivals are modeled as a point process whose intensity depends on the liquidation price. We set up a stochastic control problem in…

交易与市场微观结构 · 定量金融 2012-01-30 Erhan Bayraktar , Michael Ludkovski

We prove existence and uniqueness for some nonlinear stochastic differential equation used in molecular dynamics, whose nonlinearity comes from a conditional expectation term. We also introduce an interacting particle system in order to…

概率论 · 数学 2010-01-16 Benjamin Jourdain , Tony Lelievre , Raphaël Roux

This paper investigates large-population stochastic control problems in which agents share their state information and cooperate to minimize a convex cost functional. The latter is decomposed into individual and coupling costs, with the…

最优化与控制 · 数学 2025-10-28 Elise Devey

In this paper, we consider the problem of optimizing the worst-case behavior of a partially observed system. All uncontrolled disturbances are modeled as finite-valued uncertain variables. Using the theory of cost distributions, we present…

最优化与控制 · 数学 2023-02-21 Aditya Dave , Nishanth Venkatesh , Andreas A. Malikopoulos

In this paper, we propose an original approach to stochastic control problems. We consider a weak formulation that is written as an optimization (minimization) problem on the space of probability measures. We then introduce a penalized…

最优化与控制 · 数学 2025-08-05 Thibaut Bourdais , Nadia Oudjane , Francesco Russo

We propose a method for setting limits that avoids excluding parameter values for which the sensitivity falls below a specified threshold. These "power-constrained" limits (PCL) address the issue that motivated the widely used CLs…

数据分析、统计与概率 · 物理学 2011-05-17 Glen Cowan , Kyle Cranmer , Eilam Gross , Ofer Vitells

We provide an overview on how to use the measurable selection techniques to derive the dynamic programming principle for a general stochastic optimal control/stopping problem. By considering its martingale problem formulation on the…

最优化与控制 · 数学 2024-10-03 Nicole El Karoui , Xiaolu Tan

We derive a framework to compute optimal controls for problems with states in the space of probability measures. Since many optimal control problems constrained by a system of ordinary differential equations (ODE) modelling interacting…

最优化与控制 · 数学 2020-09-23 Martin Burger , René Pinnau , Claudia Totzeck , Oliver Tse

In this paper, we propose a class of penalty methods with stochastic approximation for solving stochastic nonlinear programming problems. We assume that only noisy gradients or function values of the objective function are available via…

最优化与控制 · 数学 2016-05-20 Xiao Wang , Shiqian Ma , Ya-xiang Yuan

We prove the existence of a solution to an equation governing the number density within a compact domain of a discrete particle system for a prescribed class of particle interactions taking into account the effects of the diffusion and…

概率论 · 数学 2007-05-23 Clive G. Wells

We investigate the numerical approximation of an elliptic optimal control problem which involves a nonconvex local regularization of the $L^q$-quasinorm penalization (with $q\in(0,1)$) in the cost function. Our approach is based on the…

最优化与控制 · 数学 2022-09-26 Pedro Merino , Alexander Nenjer

In this paper, we consider the stochastic optimal control problem for the interacting particle system. We obtain the stochastic maximum principle of the optimal control system by introducing a generalized backward stochastic differential…

概率论 · 数学 2025-05-14 Andrey A. Dorogovtsev , Yuecai Han , Kateryna Hlyniana , Yuhang Li

Despite its popularity in the reinforcement learning community, a provably convergent policy gradient method for continuous space-time control problems with nonlinear state dynamics has been elusive. This paper proposes proximal gradient…

最优化与控制 · 数学 2022-12-27 Christoph Reisinger , Wolfgang Stockinger , Yufei Zhang