中文
相关论文

相关论文: A Large Deviation Inequality for $\beta$-mixing Ti…

200 篇论文

Quantile Regression (QR) can be used to estimate aleatoric uncertainty in deep neural networks and can generate prediction intervals. Quantifying uncertainty is particularly important in critical applications such as clinical diagnosis,…

机器学习 · 计算机科学 2023-09-15 Haleh Akrami , Omar Zamzam , Anand Joshi , Sergul Aydore , Richard Leahy

Parallel tempering, or replica exchange, is a popular method for simulating complex systems. The idea is to run parallel simulations at different temperatures, and at a given swap rate exchange configurations between the parallel…

概率论 · 数学 2016-04-20 J. D. Doll , Paul Dupuis , Pierre Nyquist

The linear regression model is widely used in empirical work in Economics, Statistics, and many other disciplines. Researchers often include many covariates in their linear model specification in an attempt to control for confounders. We…

统计理论 · 数学 2017-12-12 Matias D. Cattaneo , Michael Jansson , Whitney K. Newey

We obtain error rates for large deviations of sums of i.i.d. random variables in, a particular case, of the domain of a non-symmetric infinite mean $\alpha=1$-stable law. The focus of this work is on the method of proof via analytic…

概率论 · 数学 2025-06-17 Jonny Imbierski , Dalia Terhesiu

Semiparametric mixture models are parametric models with latent variables. They are defined kernel, $p_\theta(x | z)$, where z is the unknown latent variable, and $\theta$ is the parameter of interest. We assume that the latent variables…

统计理论 · 数学 2024-12-03 Stefan Franssen , Jeanne Nguyen , Aad van der Vaart

The likelihood model of high dimensional data $X_n$ can often be expressed as $p(X_n|Z_n,\theta)$, where $\theta\mathrel{\mathop:}=(\theta_k)_{k\in[K]}$ is a collection of hidden features shared across objects, indexed by $n$, and $Z_n$ is…

机器学习 · 计算机科学 2019-05-14 Aonan Zhang , John Paisley

One of the distinguishing characteristics of modern deep learning systems is that they typically employ neural network architectures that utilize enormous numbers of parameters, often in the millions and sometimes even in the billions.…

机器学习 · 统计学 2021-11-15 Ben Adlam , Jake Levinson , Jeffrey Pennington

Let L be a positive line bundle over a projective complex manifold X. Consider the space of holomorphic sections of the tensor power of order p of L. The determinant of a basis of this space, together with some given probability measure on…

复变函数 · 数学 2016-03-14 Tien-Cuong Dinh , Viet-Anh Nguyen

Large deviation results are given for a class of perturbed nonhomogeneous Markov chains on finite state space which formally includes some stochastic optimization algorithms. Specifically, let {P_n} be a sequence of transition matrices on a…

概率论 · 数学 2007-05-23 Zach Dietz , Sunder Sethuraman

This paper studies an intriguing phenomenon related to the good generalization performance of estimators obtained by using large learning rates within gradient descent algorithms. First observed in the deep learning literature, we show that…

机器学习 · 统计学 2022-06-06 Gaspard Beugnot , Julien Mairal , Alessandro Rudi

Depth measures are powerful tools for defining level sets in emerging, non--standard, and complex random objects such as high-dimensional multivariate data, functional data, and random graphs. Despite their favorable theoretical properties,…

Let $X=\{X_n: n\in\mathbb{N}\}$ be a long memory linear process with innovations in the domain of attraction of an $\alpha$-stable law $(0<\alpha<2)$. Assume that the linear process $X$ has a bounded probability density function $f(x)$.…

统计理论 · 数学 2022-10-10 Hui Liu , Fangjun Xu

In real word applications, data generating process for training a machine learning model often differs from what the model encounters in the test stage. Understanding how and whether machine learning models generalize under such…

机器学习 · 统计学 2022-02-08 Abdulkadir Canatar , Blake Bordelon , Cengiz Pehlevan

Foundation models, particularly Large Language Models (LLMs), have revolutionized text and video processing, yet time series data presents distinct challenges for such approaches due to domain-specific features such as missing values,…

机器学习 · 计算机科学 2025-02-12 Defu Cao , Wen Ye , Yizhou Zhang , Yan Liu

In order to fully utilize "big data", it is often required to use "big models". Such models tend to grow with the complexity and size of the training data, and do not make strong parametric assumptions upfront on the nature of the…

机器学习 · 统计学 2015-04-17 Vikas Sindhwani , Haim Avron

The aim of this paper is to investigate the large deviations for a class of slow-fast mean-field diffusions, which extends some existing results to the case where the laws of fast process are also involved in the slow component. Due to the…

概率论 · 数学 2026-04-28 Wei Hong , Wei Liu , Shiyuan Yang

Generative models, like large language models, are becoming increasingly relevant in our daily lives, yet a theoretical framework to assess their generalization behavior and uncertainty does not exist. Particularly, the problem of…

机器学习 · 计算机科学 2024-07-11 Sebastian G. Gruber , Florian Buettner

This paper presents new results on Functional Analysis of Variance for fixed effect models with correlated Hilbert-valued Gaussian error components. The geometry of the Reproducing Kernel Hilbert Space (RKHS) of the error term is considered…

统计理论 · 数学 2015-09-04 M. D. Ruiz-Medina

Many scientific problems require identifying a small set of covariates that are associated with a target response and estimating their effects. Often, these effects are nonlinear and include interactions, so linear and additive methods can…

统计计算 · 统计学 2022-12-02 Raj Agrawal , Tamara Broderick

Finite mixtures of regression models provide a flexible modeling framework for many phenomena. Using moment-based estimation of the regression parameters, we develop unbiased estimators with a minimum of assumptions on the mixture…

统计理论 · 数学 2019-05-17 Claus Thorn Ekstrøm , Christian Bressen Pipper