中文
相关论文

相关论文: Sharper lower bounds on the performance of the emp…

200 篇论文

Greedy-GQ with linear function approximation, originally proposed in \cite{maei2010toward}, is a value-based off-policy algorithm for optimal control in reinforcement learning, and it has a non-linear two timescale structure with the…

机器学习 · 计算机科学 2024-05-03 Yue Wang , Yi Zhou , Shaofeng Zou

Quantum Information scrambling (QI-scrambling) is a pivotal area of inquiry within the study of quantum many-body systems. This research derives mathematical upper and lower bounds for the scrambling rate by applying the Maligranda…

量子物理 · 物理学 2025-07-25 Ahmed Zahia , M. Y. Abd-Rabbou , Atta ur Rahman , Cong Feng Qiao

We study the high-dimensional asymptotics of empirical risk minimization (ERM) in over-parametrized two-layer neural networks with quadratic activations trained on synthetic data. We derive sharp asymptotics for both training and test…

机器学习 · 统计学 2026-02-03 Vittorio Erba , Emanuele Troiani , Lenka Zdeborová , Florent Krzakala

We provide lower error bounds for randomized algorithms that approximate integrals of functions depending on an unrestricted or even infinite number of variables. More precisely, we consider the infinite-dimensional integration problem on…

数值分析 · 数学 2021-02-09 Michael Gnewuch

Empirical risk minimization is the main tool for prediction problems, but its extension to relational data remains unsolved. We solve this problem using recent ideas from graph sampling theory to (i) define an empirical risk for relational…

机器学习 · 统计学 2019-02-25 Victor Veitch , Morgane Austern , Wenda Zhou , David M. Blei , Peter Orbanz

Ordinal regression is aimed at predicting an ordinal class label. In this paper, we consider its semi-supervised formulation, in which we have unlabeled data along with ordinal-labeled data to train an ordinal regressor. There are several…

机器学习 · 计算机科学 2021-06-11 Taira Tsuchiya , Nontawat Charoenphakdee , Issei Sato , Masashi Sugiyama

Adversarial training has been widely studied in recent years due to its role in improving model robustness against adversarial attacks. This paper focuses on comparing different distributed adversarial training algorithms--including…

机器学习 · 计算机科学 2025-09-16 Ying Cao , Kun Yuan , Ali H. Sayed

We establish optimal Statistical Query (SQ) lower bounds for robustly learning certain families of discrete high-dimensional distributions. In particular, we show that no efficient SQ algorithm with access to an $\epsilon$-corrupted binary…

数据结构与算法 · 计算机科学 2022-06-10 Ilias Diakonikolas , Daniel M. Kane , Yuxin Sun

We consider the problem of estimating small ball probabilities $\mathbb P\{f(G) \leqslant \delta \mathbb Ef(G)\}$ for sub-additive,positively homogeneous functions $f$ with respect to the Gaussian measure. We establish estimates that depend…

泛函分析 · 数学 2021-07-29 Grigoris Paouris , Konstantin Tikhomirov , Petros Valettas

This guide provides a reference for high-probability regret bounds in empirical risk minimization (ERM). The presentation is modular: we begin with intuition and general proof strategies, then state broadly applicable guarantees under…

机器学习 · 统计学 2026-03-04 Lars van der Laan

The solution to empirical risk minimization with $f$-divergence regularization (ERM-$f$DR) is presented under mild conditions on $f$. Under such conditions, the optimal measure is shown to be unique. Examples of the solution for particular…

机器学习 · 统计学 2024-10-25 Francisco Daunas , Iñaki Esnaola , Samir M. Perlaza , H. Vincent Poor

It has been experimentally observed in recent years that multi-layer artificial neural networks have a surprising ability to generalize, even when trained with far more parameters than observations. Is there a theoretical basis for this?…

机器学习 · 统计学 2018-09-19 Andrew R. Barron , Jason M. Klusowski

Having a perfect model to compute the optimal policy is often infeasible in reinforcement learning. It is important in high-stakes domains to quantify and manage risk induced by model uncertainties. Entropic risk measure is an exponential…

机器学习 · 计算机科学 2020-06-23 Reazul Hasan Russel , Bahram Behzadian , Marek Petrik

Uniform deviation bounds limit the difference between a model's expected loss and its loss on an empirical sample uniformly for all models in a learning problem. As such, they are a critical component to empirical risk minimization. In this…

机器学习 · 统计学 2017-02-28 Olivier Bachem , Mario Lucic , S. Hamed Hassani , Andreas Krause

We consider the classical statistical learning/regression problem, when the value of a real random variable Y is to be predicted based on the observation of another random variable X. Given a class of functions F and a sample of independent…

统计理论 · 数学 2016-08-03 Gabor Lugosi , Shahar Mendelson

We describe an algorithm for sampling a low-rank random matrix $Q$ that best approximates a fixed target matrix $P\in\mathbb{C}^{n\times m}$ in the following sense: $Q$ is unbiased, i.e., $\mathbb{E}[Q] = P$; $\mathsf{rank}(Q)\leq r$; and…

数据结构与算法 · 计算机科学 2026-03-18 Leighton Pate Barnes , Stephen Cameron , Benjamin Howard

In recent years, there is a growing need to train machine learning models on a huge volume of data. Designing efficient distributed optimization algorithms for empirical risk minimization (ERM) has therefore become an active and challenging…

最优化与控制 · 数学 2019-11-19 Ching-pei Lee , Kai-Wei Chang

We study the problem of empirical minimization for variance-type functionals over functional classes. Sharp non-asymptotic bounds for the excess variance are derived under mild conditions. In particular, it is shown that under some…

数值分析 · 数学 2021-08-03 D. Belomestny , L. Iosipoi , Q. Paris , N. Zhivotovskiy

Gibbs-ERM learning is a natural idealized model of learning with stochastic optimization algorithms (such as Stochastic Gradient Langevin Dynamics and ---to some extent--- Stochastic Gradient Descent), while it also arises in other…

机器学习 · 计算机科学 2019-02-06 Ilja Kuzborskij , Nicolò Cesa-Bianchi , Csaba Szepesvári

The multi-armed bandit (MAB) problem is a ubiquitous decision-making problem that exemplifies exploration-exploitation tradeoff. Standard formulations exclude risk in decision making. Risknotably complicates the basic reward-maximising…

机器学习 · 计算机科学 2021-05-17 Ming Liang Ang , Eloise Y. Y. Lim , Joel Q. L. Chang