English
Related papers

Related papers: General limit value in Dynamic Programming

200 papers

We consider a finite number of $N$ statistically equal agents, each moving on a finite set of states according to a continuous-time Markov Decision Process (MDP). Transition intensities of the agents and generated rewards depend not only on…

Probability · Mathematics 2025-09-23 Nicole Bäuerle , Sebastian Höfer

We study the uniform verification problem for infinite state processes, which consists of proving that the parallel composition of an arbitrary number of processes satisfies a temporal property. Our practical motivation is to build a…

Logic in Computer Science · Computer Science 2014-01-10 Alejandro Sánchez , César Sánchez

In this paper we consider the rate of convergence of solutions of a scalar ordinary differential equation which is a perturbed version of an autonomous equation with a globally stable equilibrium. Under weak assumptions on the nonlinear…

Classical Analysis and ODEs · Mathematics 2016-07-12 John A. D. Appleby , Denis D. Patterson

Algorithms that solve zero-sum games, multi-objective agent objectives, or, more generally, variational inequality (VI) problems are notoriously unstable on general problems. Owing to the increasing need for solving such problems in machine…

Machine Learning · Statistics 2022-07-15 Tatjana Chavdarova , Ya-Ping Hsieh , Michael I. Jordan

This work takes up the challenges of utility maximization problem when the market is indivisible and the transaction costs are included. First there is a so-called solvency region given by the minimum margin requirement in the problem…

Portfolio Management · Quantitative Finance 2010-03-16 Qingshuo Song , G. Yin , Chao Zhu

The Central Limit Theorem states that, in the limit of a large number of terms, an appropriately scaled sum of independent random variables yields another random variable whose probability distribution tends to a stable distribution. The…

Data Analysis, Statistics and Probability · Physics 2024-04-08 Damián H. Zanette , Inés Samengo

We consider the continuous version of the Vicsek model with noise, proposed as a model for collective behavior of individuals with a fixed speed. We rigorously derive the kinetic mean-field partial differential equation satisfied when the…

Probability · Mathematics 2011-12-06 François Bolley , José A. Cañizo , José A. Carrillo

We address an optimal stopping problem over the set of Bermudan-type strategies $\Theta$ (which we understand in a more general sense than the stopping strategies for Bermudan options in finance) and with non-linear operators (non-linear…

Optimization and Control · Mathematics 2023-01-27 Miryana Grigorova , Marie-Claire Quenez , Peng Yuan

Previous work has shown that reasoning with real-time temporal logics is often simpler when restricted to models with bounded variability---where no more than v events may occur every V time units, for given v, V. When reasoning about…

Logic in Computer Science · Computer Science 2015-02-24 Carlo A. Furia , Paola Spoletini

In this paper, we study a stochastic recursive optimal control problem in which the value functional is defined by the solution of a backward stochastic differential equation (BSDE) under $\tilde{G}$-expectation. Under standard assumptions,…

Optimization and Control · Mathematics 2021-06-08 Mingshang Hu , Shaolin Ji , Xiaojuan Li

We present the first finite time global convergence analysis of policy gradient in the context of infinite horizon average reward Markov decision processes (MDPs). Specifically, we focus on ergodic tabular MDPs with finite state and action…

Machine Learning · Computer Science 2024-03-12 Navdeep Kumar , Yashaswini Murthy , Itai Shufaro , Kfir Y. Levy , R. Srikant , Shie Mannor

For a given distribution, learning algorithm, and performance metric, the rate of convergence (or data-scaling law) is the asymptotic behavior of the algorithm's test performance as a function of number of train samples. Many learning…

Machine Learning · Computer Science 2021-11-10 Preetum Nakkiran

This paper addresses the problem of utility maximization under uncertain parameters. In contrast with the classical approach, where the parameters of the model evolve freely within a given range, we constrain them via a penalty function. We…

Optimization and Control · Mathematics 2022-03-08 Ivan Guo , Nicolas Langrené , Grégoire Loeper , Wei Ning

The problem of constrained Markov decision process is considered. An agent aims to maximize the expected accumulated discounted reward subject to multiple constraints on its costs (the number of constraints is relatively small). A new dual…

Optimization and Control · Mathematics 2022-10-21 Egor Gladin , Maksim Lavrik-Karmazin , Karina Zainullina , Varvara Rudenko , Alexander Gasnikov , Martin Takáč

We introduce a new approach for the numerical pricing of American options. The main idea is to choose a finite number of suitable excessive functions (randomly) and to find the smallest majorant of the gain function in the span of these…

Computational Finance · Quantitative Finance 2013-10-17 Sören Christensen

We consider the two categories of termination problems of quantum programs with nondeterminism: 1) Is an input of a program terminating with probability one under all schedulers? If not, how can a scheduler be synthesized to evidence the…

Quantum Physics · Physics 2024-02-27 Jianling Fu , Hui Jiang , Ming Xu , Yuxin Deng , Zhi-Bin Li

We study the optimal dynamic pricing of an expiring ticket or voucher, sold by a time-sensitive seller to strategic buyers who arrive stochastically with private values. The expiring nature creates a conflict: the seller's urgency to sell…

Computer Science and Game Theory · Computer Science 2025-09-30 Suyeon Choi , Changhyun Kwon , Seungki Min

We give a simple combinatorial algorithm to deterministically approximately count the number of satisfying assignments of general constraint satisfaction problems (CSPs). Suppose that the CSP has domain size $q=O(1)$, each constraint…

Data Structures and Algorithms · Computer Science 2023-03-10 Kun He , Chunyang Wang , Yitong Yin

We provide a dynamic programming principle for stochastic optimal control problems with expectation constraints. A weak formulation, using test functions and a probabilistic relaxation of the constraint, avoids restrictions related to a…

Optimization and Control · Mathematics 2012-12-21 Bruno Bouchard , Marcel Nutz

We study global optimization of non-convex functions through optimal control theory. Our main result establishes that (quasi-)optimal trajectories of a discounted control problem converge globally and practically asymptotically to the set…

Optimization and Control · Mathematics 2025-11-17 Yuyang Huang , Dante Kalise , Hicham Kouhkouh
‹ Prev 1 8 9 10 Next ›