English
Related papers

Related papers: Estimating thresholding levels for random fields v…

200 papers

For a class of stochastic models with Gaussian and rough mean-reverting volatility that embeds the genuine rough Stein-Stein model, we study the weak approximation rate when using a Euler type scheme with integrated kernels. Our first…

Probability · Mathematics 2026-02-23 Aurélien Alfonsi , Ahmed Kebaier

Common cross-validation (CV) methods like k-fold cross-validation or Monte-Carlo cross-validation estimate the predictive performance of a learner by repeatedly training it on a large portion of the given data and testing on the remaining…

Machine Learning · Computer Science 2021-11-30 Felix Mohr , Jan N. van Rijn

We study regression of $1$-Lipschitz functions under a log-concave measure $\mu$ on $\mathbb{R}^d$. We focus on the high-dimensional regime where the sample size $n$ is subexponential in $d$, in which distribution-free estimators are…

Probability · Mathematics 2025-09-15 Pierre Bizeul , Boaz Klartag

The formalism to compute the geometrical and topological one-point statistics of mildly non-Gaussian 2D and 3D cosmological fields is developed. Leveraging the isotropy of the target statistics, the Gram-Charlier expansion is reformulated…

Cosmology and Nongalactic Astrophysics · Physics 2015-05-30 Christophe Gay , Christophe Pichon , Dmitri Pogosyan

We study harmonic map regression, a nonparametric estimator for manifold-valued responses, that penalizes the empirical Fr\'echet risk by the Dirichlet energy. By connecting penalized regression to the theory of harmonic maps, the estimator…

Statistics Theory · Mathematics 2026-04-13 Xiaoyu Chen

Piecewise Linear-Quadratic (PLQ) penalties are widely used to develop models in statistical inference, signal processing, and machine learning. Common examples of PLQ penalties include least squares, Huber, Vapnik, 1-norm, and their…

Machine Learning · Statistics 2021-01-01 Peng Zheng , Aleksandr Y. Aravkin , Karthikeyan Natesan Ramamurthy

We analyse the convergence of sampling algorithms for functions in reproducing kernel Hilbert spaces (RKHS). To this end, we discuss approximation properties of kernel regression under minimalistic assumptions on both the kernel and the…

Machine Learning · Statistics 2025-04-21 Armin Iske

We consider a regression framework where the design points are deterministic and the errors possibly non-i.i.d. and heavy-tailed (with a moment of order $p$ in $[1,2]$). Given a class of candidate regression functions, we propose a…

Statistics Theory · Mathematics 2025-06-03 Yannick Baraud , Guillaume Maillard

A fundamental drawback of kernel-based statistical models is their limited scalability to large data sets, which requires resorting to approximations. In this work, we focus on the popular Gaussian kernel and on techniques to linearize…

Machine Learning · Statistics 2022-04-13 Jonas Wacker , Maurizio Filippone

We propose a new approach, along with refinements, based on $L_1$ penalties and aimed at jointly estimating several related regression models. Its main interest is that it can be rewritten as a weighted lasso on a simple transformation of…

Methodology · Statistics 2014-11-07 Edouard Ollier , Vivian Viallon

We will present a new method, which enables us to find threshold functions for many properties in random intersection graphs. This method will be used to establish sharp threshold functions in random intersection graphs for k-connectivity,…

Combinatorics · Mathematics 2013-01-04 Katarzyna Rybarczyk

We describe algorithms for finding the regression of t, a sequence of values, to the closest sequence s by mean squared error, so that s is always increasing (isotonicity) and so the values of two consecutive points do not increase by too…

Data Structures and Algorithms · Computer Science 2009-12-31 Pankaj K. Agarwal , Jeff M. Phillips , Bardia Sadri

Probabilistic Component Latent Analysis (PLCA) is a statistical modeling method for feature extraction from non-negative data. It has been fruitfully applied to various research fields of information retrieval. However, the EM-solved…

Methodology · Statistics 2017-03-16 D. Cazau , G. Nuel

We tackle estimating sparse coefficients in a linear regression when the covariates are sampled from an $L$-subexponential random vector. This vector belongs to a class of distributions that exhibit heavier tails than Gaussian random…

Statistics Theory · Mathematics 2024-02-07 Takeyuki Sasai

Latent class analysis (LCA) is a useful tool to investigate the heterogeneity of a disease population with time-to-event data. We propose a new method based on non-parametric maximum likelihood estimator (NPMLE), which facilitates…

Methodology · Statistics 2022-02-03 Teng Fei , John Hanfelt , Limin Peng

Local Fr\'echet regression is a nonparametric regression method for metric space valued responses and Euclidean predictors, which can be utilized to obtain estimates of smooth trajectories taking values in general metric spaces from noisy…

Methodology · Statistics 2021-07-07 Yaqing Chen , Hans-Georg Müller

Recent advances in large-margin classification of data residing in general metric spaces (rather than Hilbert spaces) enable classification under various natural metrics, such as string edit and earthmover distance. A general framework…

Machine Learning · Computer Science 2014-07-14 Lee-Ad Gottlieb , Aryeh Kontorovich , Robert Krauthgamer

This article is dedicated to the estimation of the regression function when the explanatory variable is a weakly dependent process whose correlation coefficient exhibits exponential decay and has a known bounded density function. The…

Statistics Theory · Mathematics 2025-07-17 Karine Bertin , Lisandro Fermin , Miguel Padrino

This paper studies robust regression for data on Riemannian manifolds. Geodesic regression is the generalization of linear regression to a setting with a manifold-valued dependent variable and one or more real-valued independent variables.…

Machine Learning · Statistics 2022-01-26 Ha-Young Shin , Hee-Seok Oh

The predictive quality of machine learning models is typically measured in terms of their (approximate) expected prediction accuracy or the so-called Area Under the Curve (AUC). Minimizing the reciprocals of these measures are the goals of…

Machine Learning · Statistics 2019-03-04 Hiva Ghanbari , Minhan Li , Katya Scheinberg