English
Related papers

Related papers: Uniform inference for value functions

200 papers

This paper concerns the use of a particular class of determinantal point processes (DPP), a class of repulsive spatial point processes, for Monte Carlo integration. Let $d\ge 1$, $I\subseteq \overline d=\{1,\dots,d\}$ with $\iota=|I|$.…

Computation · Statistics 2021-10-19 Jean-François Coeurjolly , Adrien Mazoyer , Pierre-Olivier Amblard

In this article, we study the notion of gH-Hadamard derivative for interval-valued functions (IVFs) and its applications to interval optimization problems (IOPs). It is shown that the existence of gH-Hadamard derivative implies the…

Optimization and Control · Mathematics 2022-10-07 Ram Surat Chauhan , Debdas Ghosh , Qamrul Hasan Ansari

In this paper we propose a new methodology to represent the results of the robust ordinal regression approach by means of a family of representative value functions for which, taken two alternatives $a$ and $b$, the following two conditions…

Optimization and Control · Mathematics 2021-07-19 Sally Giuseppe Arcidiacono , Salvatore Corrente , Salvatore Greco

We consider the oscillatory integrals with parameter-dependent phases. We decompose the integrals into a leading term and a remainder term. Instead of the pointwise estimate, we use some $L^p$-estimate for the remainder term and get various…

Classical Analysis and ODEs · Mathematics 2024-02-14 Zihua Guo

Constructing confidence intervals for the value of an (unknown) optimal treatment policy is a fundamental problem in causal inference. Insight into the optimal policy value can guide the development of reward-maximizing, individualized…

Econometrics · Economics 2026-04-01 Justin Whitehouse , Qizhao Chen , Morgane Austern , Vasilis Syrgkanis

We consider centralized and distributed mirror descent algorithms over a finite-dimensional Hilbert space, and prove that the problem variables converge to an optimizer of a possibly nonsmooth function when the step sizes are square…

Optimization and Control · Mathematics 2018-05-07 Thinh T. Doan , Subhonmesh Bose , D. Hoa Nguyen , Carolyn L. Beck

Estimating epistemic uncertainty in value functions is a crucial challenge for many aspects of reinforcement learning (RL), including efficient exploration, safe decision-making, and offline RL. While deep ensembles provide a robust method…

We develop a new proximal-gradient method for minimizing the sum of a differentiable, possibly nonconvex, function plus a convex, possibly non differentiable, function. The key features of the proposed method are the definition of a…

Numerical Analysis · Mathematics 2016-05-13 Silvia Bonettini , Ignace Loris , Federica Porta , Marco Prato

Several machine learning applications involve the optimization of higher-order derivatives (e.g., gradients of gradients) during training, which can be expensive in respect to memory and computation even with automatic differentiation. As a…

Machine Learning · Computer Science 2020-11-26 Tianyu Pang , Kun Xu , Chongxuan Li , Yang Song , Stefano Ermon , Jun Zhu

Objective functions that optimize deep neural networks play a vital role in creating an enhanced feature representation of the input data. Although cross-entropy-based loss formulations have been extensively used in a variety of supervised…

Computer Vision and Pattern Recognition · Computer Science 2023-12-19 Deen Dayal Mohan , Bhavin Jawade , Srirangaraj Setlur , Venu Govindaraj

We can, and should, do statistical inference on simulation models by adjusting the parameters in the simulation so that the values of {\em randomly chosen} functions of the simulation output match the values of those same functions…

Methodology · Statistics 2021-11-18 Cosma Rohilla Shalizi

General unsupervised learning is a long-standing conceptual problem in machine learning. Supervised learning is successful because it can be solved by the minimization of the training error cost function. Unsupervised learning is not as…

Machine Learning · Computer Science 2015-12-04 Ilya Sutskever , Rafal Jozefowicz , Karol Gregor , Danilo Rezende , Tim Lillicrap , Oriol Vinyals

We study the problem of directly optimizing arbitrary non-differentiable task evaluation metrics such as misclassification rate and recall. Our method, named MetricOpt, operates in a black-box setting where the computational details of the…

Machine Learning · Computer Science 2021-04-22 Chen Huang , Shuangfei Zhai , Pengsheng Guo , Josh Susskind

We give a technical overview of our exact-real implementation of various representations of the space of continuous unary real functions over the unit domain and a family of associated (partial) operations, including integration, range…

Logic in Computer Science · Computer Science 2019-10-14 Michal Konečný , Eike Neumann

We present a general convergent class of reinforcement learning algorithms that is founded on two distinct principles: (1) mapping value estimates to a different space using arbitrary functions from a broad class, and (2) linearly…

Machine Learning · Computer Science 2022-03-18 Mehdi Fatemi , Arash Tavakoli

This paper proposes a unique optimization approach for estimating the minimax rational approximation and its application for evaluating matrix functions. Our method enables the extension to generalized rational approximations and has the…

Numerical Analysis · Mathematics 2025-04-03 Nir Sharon , Vinesha Peiris , Nadia Sukhorukova , Julien Ugon

This paper concerns with iterative schemes for the perfect reconstruction of functions belonging to multiresolution spaces on bounded manifolds from nonuniform sampling. The schemes have optimal complexity in the sense that the…

Numerical Analysis · Mathematics 2007-05-23 Massimo Fornasier , Laura Gori

For multi-valued functions---such as when the conditional distribution on targets given the inputs is multi-modal---standard regression approaches are not always desirable because they provide the conditional mean. Modal regression…

Machine Learning · Statistics 2020-10-30 Yangchen Pan , Ehsan Imani , Martha White , Amir-massoud Farahmand

Partially observable Markov decision processes (POMDPs) have recently become popular among many AI researchers because they serve as a natural model for planning under uncertainty. Value iteration is a well-known algorithm for finding…

Artificial Intelligence · Computer Science 2011-06-02 N. L. Zhang , W. Zhang

We introduce an empirical functional $\Psi$ that is an optimal uniform mean estimator: Let $F\subset L_2(\mu)$ be a class of mean zero functions, $u$ is a real valued function, and $X_1,\dots,X_N$ are independent, distributed according to…

Probability · Mathematics 2026-03-06 Daniel Bartl , Shahar Mendelson