English
Related papers

Related papers: The Target Discounted-Sum Problem

200 papers

In recent years there has been intense interest in the vanishing discount problem for Hamilton-Jacobi equations. In the case of the scalar equation, B. Ziliotto has recently given an example of the Hamilton-Jacobi equation having non-convex…

Analysis of PDEs · Mathematics 2022-02-08 Hitoshi Ishii

This paper studies semiparametric contextual bandits, a generalization of the linear stochastic bandit problem where the reward for an action is modeled as a linear function of known action features confounded by an non-linear…

Machine Learning · Statistics 2018-07-17 Akshay Krishnamurthy , Zhiwei Steven Wu , Vasilis Syrgkanis

The Promise Constraint Satisfaction Problem (PCSP) is a generalization of the Constraint Satisfaction Problem (CSP) that includes approximation variants of satisfiability and graph coloring problems. Barto [LICS '19] has shown that a…

Computational Complexity · Computer Science 2025-06-09 Kristina Asimi , Libor Barto

While discounted payoff games and classic games that reduce to them, like parity and mean-payoff games, are symmetric, their solutions are not. We have taken a fresh view on the properties that optimal solutions need to have, and devised a…

Data Structures and Algorithms · Computer Science 2026-03-11 Daniele Dell'Erba , Arthur Dumas , Sven Schewe

While discounted payoff games and classic games that reduce to them, like parity and mean-payoff games, are symmetric, their solutions are not. We have taken a fresh view on the constraints that optimal solutions need to satisfy, and…

Data Structures and Algorithms · Computer Science 2023-10-03 Daniele Dell'Erba , Arthur Dumas , Sven Schewe

In this article we propose a Weighted Stochastic Mesh (WSM) Algorithm for approximating the value of a discrete and continuous time optimal stopping problem. We prove that in the discrete case the WSM algorithm leads to semi-tractability of…

Computational Finance · Quantitative Finance 2019-06-25 D. Belomestny , M. Kaledin , J. Schoenmakers

Let $(\xi_1, \eta_1)$, $(\xi_2, \eta_2),\ldots$ be independent identically distributed $\mathbb{R}^2$-valued random vectors. Assuming that $\xi_1$ has zero mean and finite variance and imposing three distinct groups of assumptions on the…

Probability · Mathematics 2022-08-03 Alexander Iksanov , Alexander Marynych , Anatolii Nikitin

The stochastic composition optimization proposed recently by Wang et al. [2014] minimizes the objective with the compositional expectation form: $\min_x~(\mathbb{E}_iF_i \circ \mathbb{E}_j G_j)(x).$ It summarizes many important applications…

Optimization and Control · Mathematics 2017-05-23 Xiangru Lian , Mengdi Wang , Ji Liu

In this paper we consider a fragment of the first-order theory of the real numbers that includes systems of equations of continuous functions in bounded domains, and for which all functions are computable in the sense that it is possible to…

Computational Complexity · Computer Science 2016-08-15 Peter Franek , Stefan Ratschan , Piotr Zgliczynski

Let $f\in \mathbb{Q}(x)$ be a non-constant rational function. We consider "Waring's Problem for $f(x)$," i.e., whether every element of $\bbq$ can be written as a bounded sum of elements of $\{f(a)\mid a\in \mathbb{Q}\}$. For rational…

Number Theory · Mathematics 2018-01-23 Bo-Hae Im , Michael Larsen

We consider optimization problems in which the objective requires an inner loop with many steps or is the limit of a sequence of increasingly costly approximations. Meta-learning, training recurrent neural networks, and optimization of the…

Machine Learning · Computer Science 2019-05-20 Alex Beatson , Ryan P. Adams

We consider a budget-constrained bandit problem where each arm pull incurs a random cost, and yields a random reward in return. The objective is to maximize the total expected reward under a budget constraint on the total cost. The model is…

Machine Learning · Computer Science 2020-03-03 Semih Cayci , Atilla Eryilmaz , R. Srikant

We study the links between the values of stochastic games with varying stage duration $h$, the corresponding Shapley operators $\bf{T}$ and ${\bf{T}}\_h$and the solution of $\dot f\_t = ({\bf{T}} - Id )f\_t$. Considering general non…

Optimization and Control · Mathematics 2016-01-11 Sylvain Sorin , Guillaume Vigeral

Given two weighted automata, we consider the problem of whether one is big-O of the other, i.e., if the weight of every finite word in the first is not greater than some constant multiple of the weight in the second. We show that the…

Formal Languages and Automata Theory · Computer Science 2023-06-22 Dmitry Chistikov , Stefan Kiefer , Andrzej S. Murawski , David Purser

In this paper we establish a new summation method by expanding $\prod_{k}(1-\frac{z}{a_{k}})^{-1}$ with two approaches: the Taylor expansion and the infinite partial fraction decomposition. Here we focus on the case when $a_{k}$ is…

Classical Analysis and ODEs · Mathematics 2021-02-09 Xiaowei Wang

We consider the two categories of termination problems of quantum programs with nondeterminism: 1) Is an input of a program terminating with probability one under all schedulers? If not, how can a scheduler be synthesized to evidence the…

Quantum Physics · Physics 2024-02-27 Jianling Fu , Hui Jiang , Ming Xu , Yuxin Deng , Zhi-Bin Li

We consider an adversarial variant of the classic $K$-armed linear contextual bandit problem where the sequence of loss functions associated with each arm are allowed to change without restriction over time. Under the assumption that the…

Machine Learning · Computer Science 2022-05-25 Gergely Neu , Julia Olkhovskaya

We take a unifying approach to single selection optimal stopping problems with random arrival order and independent sampling of items. In the problem we consider, a decision maker (DM) initially gets to sample each of $N$ items…

Computer Science and Game Theory · Computer Science 2021-08-11 José Correa , Andrés Cristi , Boris Epstein , José Soto

We design alignment-free techniques for comparing a sequence or word, called a target, against a set of words, called a reference. A target-specific factor of a target $T$ against a reference $R$ is a factor $w$ of a word in $T$ which is…

Data Structures and Algorithms · Computer Science 2023-04-07 Marie-Pierre Béal , Maxime Crochemore

We describe a novel optimization method for finite sums (such as empirical risk minimization problems) building on the recently introduced SAGA method. Our method achieves an accelerated convergence rate on strongly convex smooth problems.…

Machine Learning · Statistics 2016-10-31 Aaron Defazio
‹ Prev 1 8 9 10 Next ›