English
Related papers

Related papers: Finite-Step Bounds for Iterated Correlation Matric…

200 papers

Policy Iteration (PI) is a classical family of algorithms to compute an optimal policy for any given Markov Decision Problem (MDP). The basic idea in PI is to begin with some initial policy and to repeatedly update the policy to one from an…

This paper proposes a mechanism to fine-tune convex approximations of probabilistic reachable sets (PRS) of uncertain dynamic systems. We consider the case of unbounded uncertainties, for which it may be impossible to find a bounded…

Robotics · Computer Science 2024-02-06 Pengcheng Wu , Sonia Martinez , Jun Chen

A common problem to all applications of linear finite dynamical systems is analyzing the dynamics without enumerating every possible state transition. Of particular interest is the long term dynamical behaviour. In this paper, we study the…

Dynamical Systems · Mathematics 2019-04-01 Björn Lindenberg

In simulation-based inferences for partially observed Markov process models (POMP), the by-product of the Monte Carlo filtering is an approximation of the log likelihood function. Recently, iterated filtering [14, 13] has originally been…

Methodology · Statistics 2018-02-26 Dao Nguyen

In this paper we consider the problem of finding $\epsilon$-approximate stationary points of convex functions that are $p$-times differentiable with $\nu$-H\"{o}lder continuous $p$th derivatives. We present tensor methods with and without…

Optimization and Control · Mathematics 2021-06-07 Geovani Nunes Grapiglia , Yurii Nesterov

A Fixed-Parameter Tractable (\FPT) $\rho$-approximation algorithm for a minimization (resp. maximization) parameterized problem $P$ is an FPT algorithm that, given an instance $(x, k)\in P$ computes a solution of cost at most $k \cdot…

Data Structures and Algorithms · Computer Science 2013-08-19 Rajesh Chitnis , MohammadTaghi Hajiaghayi , Guy Kortsarz

In many branches of engineering, Banach contraction mapping theorem is employed to establish the convergence of certain deterministic algorithms. Randomized versions of these algorithms have been developed that have proved useful in…

Probability · Mathematics 2023-09-25 Abhishek Gupta , Rahul Jain , Peter Glynn

Partially observable Markov decision processes (POMDPs) are a fundamental model for sequential decision-making under uncertainty. However, many verification and synthesis problems for POMDPs are undecidable or intractable. Most prominently,…

Artificial Intelligence · Computer Science 2026-04-23 Nathanaël Fijalkow , Arka Ghosh , Roman Kniazev , Guillermo A. Pérez , Pierre Vandenhove

During recent years the interest of optimization and machine learning communities in high-probability convergence of stochastic optimization methods has been growing. One of the main reasons for this is that high-probability complexity…

This paper develops an optimal Chernoff type bound for the probabilities of large deviations of sums $\sum_{k=1}^n f (X_k)$ where $f$ is a real-valued function and $(X_k)_{k \in \mathbb{Z}_{\ge 0}}$ is a finite state Markov chain with an…

Probability · Mathematics 2019-12-24 Vrettos Moulos , Venkat Anantharam

We study a family of correlated one-dimensional random walks with a finite memory range M.These walks are extensions of the Taylor's walk as investigated by Goldstein, which has a memory range equal to one. At each step, with a probability…

adap-org · Physics 2009-10-31 Roger Bidaux , Nino Boccara

In this work, we describe a generic approach to show convergence with high probability for stochastic convex optimization. In previous works, either the convergence is only in expectation or the bound depends on the diameter of the domain.…

Optimization and Control · Mathematics 2022-10-04 Alina Ene , Huy L. Nguyen

Quantum states evolving at equidistant steps into a set of mutually orthogonal states of finite or infinite cardinality p exhibit an interesting physical effect. The analysis of the amplitudes of the state at half the step time with the…

Quantum Physics · Physics 2009-09-29 Hans-Rudolf Thomann

We consider the problem of learning control policies that optimize a reward function while satisfying constraints due to considerations of safety, fairness, or other costs. We propose a new algorithm, Projection-Based Constrained Policy…

Machine Learning · Computer Science 2020-10-08 Tsung-Yen Yang , Justinian Rosca , Karthik Narasimhan , Peter J. Ramadge

Partially Observable Markov Decision Processes (POMDPs) are fundamental to decision-making under uncertainty. We introduce a novel scalable approach to accelerate upper bound estimation in Point-Based Value Iteration (PBVI) algorithms, the…

Optimization and Control · Mathematics 2025-03-13 Siqiong Zhou , Ashif S. Iquebal , Esma S. Gel

We study both the positively and negatively step-reinforced random walks with parameter $p$. For a step distribution $\mu$ with finite second moment, the positively step-reinforced random walk with $p\in [1/2,1)$ and the negatively…

Probability · Mathematics 2025-04-04 Zhishui Hu

A long-standing open question is which graph class is the most general one permitting constant-time constant-factor approximations for dominating sets. The approximation ratio has been bounded by increasingly general parameters such as…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-08-26 Christoph Lenzen , Sophie Wenning

Nonconvex and nonsmooth optimization problems are frequently encountered in much of statistics, business, science and engineering, but they are not yet widely recognized as a technology in the sense of scalability. A reason for this…

Optimization and Control · Mathematics 2018-01-19 Bo Jiang , Tianyi Lin , Shiqian Ma , Shuzhong Zhang

We resolve several fundamental questions in the area of distributed functional monitoring, initiated by Cormode, Muthukrishnan, and Yi (SODA, 2008). In this model there are $k$ sites each tracking their input and communicating with a…

Data Structures and Algorithms · Computer Science 2013-06-13 David P. Woodruff , Qin Zhang

We develop an algorithm for computing bounded reachability probability for hybrid systems, i.e., the probability that the system reaches an unsafe region within a finite number of discrete transitions. In particular, we focus on hybrid…

Logic in Computer Science · Computer Science 2015-05-13 Fedor Shmarov , Paolo Zuliani