English
Related papers

Related papers: Perov's Contraction Principle and Dynamic Programm…

200 papers

We employ constraints to control the parameter space of deep neural networks throughout training. The use of customized, appropriately designed constraints can reduce the vanishing/exploding gradients problem, improve smoothness of…

Machine Learning · Computer Science 2021-06-22 Benedict Leimkuhler , Tiffany Vlaar , Timothée Pouchon , Amos Storkey

In multi-objective optimization, a single decision vector must balance the trade-offs between many objectives. Solutions achieving an optimal trade-off are said to be Pareto optimal: these are decision vectors for which improving any one…

Optimization and Control · Mathematics 2023-08-07 Abhishek Roy , Geelon So , Yi-An Ma

We consider how to use the Bellman residual of the dynamic programming operator to compute suboptimality bounds for solutions to stochastic shortest path problems. Such bounds have been previously established only in the special case that…

Artificial Intelligence · Computer Science 2012-02-20 Eric A. Hansen

The most relevant problems in discounted reinforcement learning involve estimating the mean of a function under the stationary distribution of a Markov reward process, such as the expected return in policy evaluation, or the policy gradient…

Machine Learning · Computer Science 2023-04-17 Alberto Maria Metelli , Mirco Mutti , Marcello Restelli

The problem of minimizing convex functionals of probability distributions is solved under the assumption that the density of every distribution is bounded from above and below. A system of sufficient and necessary first-order optimality…

Information Theory · Computer Science 2018-12-05 Michael Fauss , Abdelhak M. Zoubir

We provide an overview on how to use the measurable selection techniques to derive the dynamic programming principle for a general stochastic optimal control/stopping problem. By considering its martingale problem formulation on the…

Optimization and Control · Mathematics 2024-10-03 Nicole El Karoui , Xiaolu Tan

In this paper, we prove that the Banach contraction principle proved by S. G. Matthews in 1994 on 0--complete partial metric spaces can be extended to cyclical mappings. However, the generalized contraction principle proved by D. Ili\'{c},…

General Topology · Mathematics 2011-12-30 Thabet Abdeljawad , Jehad O. Alzabut , Aiman Mukheimer , Younes Zaidan

In performative prediction, a predictive model impacts the distribution that generates future data, a phenomenon that is being ignored in classical supervised learning. In this closed-loop setting, the natural measure of performance named…

Machine Learning · Computer Science 2022-10-24 Yulai Zhao

A robust-to-dynamics optimization (RDO) problem is an optimization problem specified by two pieces of input: (i) a mathematical program (an objective function $f:\mathbb{R}^n\rightarrow\mathbb{R}$ and a feasible set…

Optimization and Control · Mathematics 2023-11-27 Amir Ali Ahmadi , Oktay Gunluk

We provide a first-order necessary and sufficient condition for optimality of lower semicontinuous functions on Banach spaces using the concept of subdifferential. From the sufficient condition we derive that any subdifferential operator is…

Optimization and Control · Mathematics 2014-01-23 Florence Jules , Marc Lassonde

We describe an abstract control-theoretic framework in which the validity of the dynamic programming principle can be established in continuous time by a verification of a small number of structural properties. As an application we treat…

Optimization and Control · Mathematics 2014-03-18 Gordan Zitkovic

We consider the linear programming approach for constrained and unconstrained Markov decision processes (MDPs) under the long-run average cost criterion, where the class of MDPs in our study have Borel state spaces and discrete countable…

Optimization and Control · Mathematics 2021-04-20 Huizhen Yu

We provide a control-theoretic perspective on optimal tensor algorithms for minimizing a convex function in a finite-dimensional Euclidean space. Given a function $\Phi: \mathbb{R}^d \rightarrow \mathbb{R}$ that is convex and twice…

Optimization and Control · Mathematics 2026-01-21 Tianyi Lin , Michael. I. Jordan

In the present paper, the Polyak's principle, concerning convexity of the images of small balls through C1,1 mappings, is employed in the study of vector optimization problems. This leads to extend to such a context achievements of local…

Optimization and Control · Mathematics 2013-06-24 Amos Uderzo

We suggest that the tools of contraction analysis for deterministic systems can be applied towards studying the convergence behavior of stochastic dynamical systems in the Wasserstein metric. In particular, we consider the case of Ito…

Optimization and Control · Mathematics 2019-03-01 Jake Bouvrie , Jean-Jacques Slotine

This paper studies a finite-fuel two-dimensional degenerate singular stochastic control problem under regime switching that is motivated by the optimal irreversible extraction problem of an exhaustible commodity. A company extracts a…

Optimization and Control · Mathematics 2017-12-29 Giorgio Ferrari , Shuzhen Yang

We analyze an optimal stopping problem with a constraint on the expected cost. When the reward function and cost function are Lipschitz continuous in state variable, we show that the value of such an optimal stopping problem is a continuous…

Optimization and Control · Mathematics 2017-08-08 Erhan Bayraktar , Song Yao

In this paper we present a dynamic programing approach to stochastic optimal control problems with dynamic, time-consistent risk constraints. Constrained stochastic optimal control problems, which naturally arise when one has to consider…

Optimization and Control · Mathematics 2015-11-24 Yin-Lam Chow , Marco Pavone

In this work we provide explicit conditions on the existence of optimal feedback controls for stochastic processes with regime-switching. We use the compactification method which needs less regularity conditions on the coefficients of the…

Optimization and Control · Mathematics 2020-01-14 Jinghai Shao

At each iteration of a Block Coordinate Descent method one minimizes an approximation of the objective function with respect to a generally small set of variables subject to constraints in which these variables are involved. The…

Optimization and Control · Mathematics 2023-04-28 E. G. Birgin , J. M. Martínez
‹ Prev 1 4 5 6 7 8 10 Next ›