English
Related papers

Related papers: Two-grid Penalty Approximation Scheme for Doubly R…

200 papers

In this work, we develop analysis and algorithms for a class of (stochastic) bilevel optimization problems whose lower-level (LL) problem is strongly convex and linearly constrained. Most existing approaches for solving such problems rely…

Optimization and Control · Mathematics 2025-04-08 Prashant Khanduri , Ioannis Tsaknakis , Yihua Zhang , Sijia Liu , Mingyi Hong

We consider a non-Markovian optimal stopping problem on finite horizon. We prove that the value process can be represented by means of a backward stochastic differential equation (BSDE), defined on an enlarged probability space, containing…

Probability · Mathematics 2015-02-20 Marco Fuhrman , Huyên Pham , Federica Zeni

We introduce a new numerical method to approximate the solution of a finite horizon deterministic optimal control problem. We exploit two Hamilton-Jacobi-Bellman PDE, arising by considering the dynamics in forward and backward time. This…

Optimization and Control · Mathematics 2023-04-21 Marianne Akian , Stéphane Gaubert , Shanqing Liu

In this paper, we introduce proximal gradient temporal difference learning, which provides a principled way of designing and analyzing true stochastic gradient temporal difference learning algorithms. We show how gradient TD (GTD)…

Machine Learning · Computer Science 2020-06-09 Bo Liu , Ian Gemp , Mohammad Ghavamzadeh , Ji Liu , Sridhar Mahadevan , Marek Petrik

While many methods exist to discretize nonlinear time-dependent partial differential equations (PDEs), the rigorous estimation and adaptive control of their discretization errors remains challenging. In this paper, we present a methodology…

Numerical Analysis · Mathematics 2017-06-15 Xunxun Wu , Kristoffer van der Zee , Gorkem Simsek , Harald Van Brummelen

Computable estimates for the error of finite element discretisations of parabolic problems in the $L^\infty(0,T; L^2)$ norm are developed, which exhibit constant effectivities (the ratio of the estimated error to the true error) with…

Numerical Analysis · Mathematics 2018-03-09 Oliver J. Sutton

Port-Hamiltonian systems provide a highly-structured framework for modeling of physical systems. By definition, they encode a balance equation relating energy changes to supplied and dissipated energy. Capturing this energy balance in…

Numerical Analysis · Mathematics 2026-05-15 Aashutosh Sharma , Andreas Bartel , Manuel Schaller

In this paper we propose a randomized primal-dual proximal block coordinate updating framework for a general multi-block convex optimization model with coupled objective function and linear constraints. Assuming mere convexity, we establish…

Optimization and Control · Mathematics 2017-01-25 Xiang Gao , Yangyang Xu , Shuzhong Zhang

The optimal stopping problem is one of the core problems in financial markets, with broad applications such as pricing American and Bermudan options. The deep BSDE method [Han, Jentzen and E, PNAS, 115(34):8505-8510, 2018] has shown great…

Probability · Mathematics 2023-08-28 Chengfan Gao , Siping Gao , Ruimeng Hu , Zimu Zhu

We detail in this article the necessity of a change of paradigm for the delay-robust control of systems composed of two linear first order hyperbolic equations. One must go back to the classical trade-off between convergence rate and…

Optimization and Control · Mathematics 2017-09-14 Jean Auriol , Jakob Ulf , Philippe Martin , Florent Meglio

We consider a Markov chain approximation scheme for utility maximization problems in continuous time, which uses, in turn, a piecewise constant policy approximation, Euler-Maruyama time stepping, and a Gauss-Hermite approximation of the…

Optimization and Control · Mathematics 2020-01-07 Athena Picarelli , Christoph Reisinger

Parallel-in-time methods for partial differential equations (PDEs) have been the subject of intense development over recent decades, particularly for diffusion-dominated problems. It has been widely reported in the literature, however, that…

Numerical Analysis · Mathematics 2023-03-22 H. De Sterck , R. D. Falgout , O. A. Krzysik , J. B. Schroder

We show how a posteriori goal oriented error estimation can be used to efficiently solve the subproblems occurring in a Model Predictive Control (MPC) algorithm. In MPC, only an initial part of a computed solution is implemented as a…

Optimization and Control · Mathematics 2022-03-02 Lars Grüne , Manuel Schaller , Anton Schiela

This paper aims to investigate the numerical approximation of semilinear non-autonomous stochastic partial differential equations (SPDEs) driven by multiplicative or additive noise. Such equations are more realistic than autonomous SPDEs…

Numerical Analysis · Mathematics 2020-11-18 Jean Daniel Mukam , Antoine Tambue

We formulate a notion of doubly reflected BSDEs with a default time and two completely separated RCLL barriers. We demonstrate the existence and uniqueness of the solution. Within the defaultable setup, we introduce a type of generalized…

Probability · Mathematics 2025-07-09 Badr Elmansouri , Mohamed El Otmani

Bilevel programming has recently received attention in the literature due to its wide range of applications, including reinforcement learning and hyper-parameter optimization. However, it is widely assumed that the underlying bilevel…

Machine Learning · Computer Science 2024-10-11 Parvin Nazari , Ahmad Mousavi , Davoud Ataee Tarzanagh , George Michailidis

We study the problem of infinite-horizon average-reward reinforcement learning with linear Markov decision processes (MDPs). The associated Bellman operator of the problem not being a contraction makes the algorithm design challenging.…

Machine Learning · Statistics 2025-03-12 Kihyuk Hong , Woojin Chae , Yufan Zhang , Dabeen Lee , Ambuj Tewari

In this paper we present an algorithm for adaptive sparse grid approximations of quantities of interest computed from discretized partial differential equations. We use adjoint-based a posteriori error estimates of the physical…

Numerical Analysis · Computer Science 2015-06-22 John D. Jakeman , Timothy Wildey

Reinforcement learning is widely used in applications where one needs to perform sequential decisions while interacting with the environment. The problem becomes more challenging when the decision requirement includes satisfying some safety…

Machine Learning · Computer Science 2022-07-15 Qinbo Bai , Amrit Singh Bedi , Mridul Agarwal , Alec Koppel , Vaneet Aggarwal

Time-parallel algorithms seek greater concurrency by decomposing the temporal domain of a Partial Differential Equation (PDE), providing possibilities for accelerating the computation of its solution. While parallelisation in time has…

Numerical Analysis · Mathematics 2021-04-20 Federico Danieli , Scott MacLachlan