English
Related papers

Related papers: On the correctness of monadic backward induction

200 papers

Inspired by Shapiro et al.~\cite{shapiro2023episodic}, we consider a stochastic optimal control (SOC) and Markov decision process (MDP) where the risks arising from epistemic and aleatoric uncertainties are assessed using Bayesian composite…

Optimization and Control · Mathematics 2025-09-01 Wentao Ma , Zhiping Chen , Huifu Xu

We describe methods for proving bounds on infinite-time averages in differential dynamical systems. The methods rely on the construction of nonnegative polynomials with certain properties, similarly to the way nonlinear stability can be…

Dynamical Systems · Mathematics 2021-06-25 David Goluskin

We study a novel general class of multidimensional type-I backward stochastic Volterra integral equations. Toward this goal, we introduce an infinite dimensional system of standard backward SDEs and establish its well-posedness, and we show…

Probability · Mathematics 2020-08-05 Camilo Hernández , Dylan Possamaï

In this paper, we develop an optimization-based framework for solving coupled forward-backward stochastic differential equations. We introduce an integral-form objective function and prove its equivalence to the error between consecutive…

Optimization and Control · Mathematics 2025-07-22 Yutian Wang , Yuan-Hua Ni , Xun Li

In this article, by using several new crucial {\it a priori} estimates which are still absent in the literature, we provide a comprehensive resolution of the first order generic mean field type control problems and also establish the…

Optimization and Control · Mathematics 2023-09-18 Alain Bensoussan , Tak Kwong Wong , Sheung Chi Phillip Yam , Hongwei Yuan

We introduce a generalized forward-backward splitting method with penalty term for solving monotone inclusion problems involving the sum of a finite number of maximally monotone operators and the normal cone to the nonempty set of zeros of…

Optimization and Control · Mathematics 2018-07-31 Nimit Nimana , Narin Petrot

We adopt an optimal-control framework for addressing the undiscounted infinite-horizon discrete-time restless $N$-armed bandit problem. Unlike most studies that rely on constructing policies based on the relaxed single-armed Markov Decision…

Optimization and Control · Mathematics 2024-03-19 Chen YAN

This work addresses the general problem of control synthesis for continuous-space, discrete-time stochastic systems with probabilistic guarantees via finite abstractions. While established methods exist, they often trade off accuracy for…

Systems and Control · Electrical Eng. & Systems 2025-07-04 Ibon Gracia , Morteza Lahijanian

Model-free reinforcement learning is known to be memory and computation efficient and more amendable to large scale problems. In this paper, two model-free algorithms are introduced for learning infinite-horizon average-reward Markov…

Machine Learning · Computer Science 2020-02-26 Chen-Yu Wei , Mehdi Jafarnia-Jahromi , Haipeng Luo , Hiteshi Sharma , Rahul Jain

Koopman operator theory provides a global linear representation of nonlinear dynamics and underpins many data-driven methods. In practice, however, finite-dimensional feature spaces induced by a user-chosen dictionary are rarely invariant,…

Finite dimensional solutions to a class of stochastic partial differential equations are obtained extending the differential constraints method for deterministic PDE to the stochastic framework. A geometrical reformulation of the stochastic…

Probability · Mathematics 2017-12-25 Francesco C. De Vecchi

The reinforcement learning algorithm SARSA combined with linear function approximation has been shown to converge for infinite horizon discounted Markov decision problems (MDPs). In this paper, we investigate the convergence of the…

Machine Learning · Computer Science 2023-06-08 Lina Palmborg

The solutions to many sequential decision-making problems are characterized by dynamic programming and Bellman's principle of optimality. However, due to the inherent complexity of solving Bellman's equation exactly, there has been…

Systems and Control · Electrical Eng. & Systems 2026-03-24 Bowen Li , Edwin K. P. Chong , Ali Pezeshki

We investigate Stochastic Mirror Descent (SMD) with matrix parameters and vector-valued predictions, a framework relevant to multi-class classification and matrix completion problems. Focusing on the overparameterized regime, where the…

Machine Learning · Statistics 2026-03-02 Danil Akhtiamov , Reza Ghane , Omead Pooladzandi , Babak Hassibi

Although neural networks have been applied to several systems in recent years, they still cannot be used in safety-critical systems due to the lack of efficient techniques to certify their robustness. A number of techniques based on convex…

Machine Learning · Computer Science 2021-09-28 Ziye Ma , Somayeh Sojoudi

We study the model-based undiscounted reinforcement learning for partially observable Markov decision processes (POMDPs). The oracle we consider is the optimal policy of the POMDP with a known environment in terms of the average reward over…

Machine Learning · Computer Science 2022-07-19 Yi Xiong , Ningyuan Chen , Xuefeng Gao , Xiang Zhou

Low rank matrix recovery problems appear widely in statistics, combinatorics, and imaging. One celebrated method for solving these problems is to formulate and solve a semidefinite program (SDP). It is often known that the exact solution to…

Optimization and Control · Mathematics 2021-07-26 Lijun Ding , Madeleine Udell

This paper investigates the exact controllability problem for multi-dimensional stochastic first-order symmetric hyperbolic systems with control inputs acting in two distinct ways: an internal control applied to the diffusion term and a…

Optimization and Control · Mathematics 2026-01-27 Zengyu Li , Qi Lü , Yu Wang , Haitian Yang

As a primary contribution, we present a convergence theorem for stochastic iterations, and in particular, Q-learning iterates, under a general, possibly non-Markovian, stochastic environment. Our conditions for convergence involve an…

Optimization and Control · Mathematics 2024-03-05 Ali Devran Kara , Serdar Yuksel

In this work we are interested in general linear inverse problems where the corresponding forward problem is solved iteratively using fixed point methods. Then one-shot methods, which iterate at the same time on the forward problem solution…

Numerical Analysis · Mathematics 2024-05-15 Marcella Bonazzoli , Houssem Haddar , Tuan Anh Vu
‹ Prev 1 3 4 5 6 7 10 Next ›