English
Related papers

Related papers: SMML estimators for linear regression and tessella…

200 papers

Estimation of signal-to-noise ratios and residual variances in high-dimensional linear models has various important applications including, e.g. heritability estimation in bioinformatics. One commonly used estimator, usually referred to as…

Statistics Theory · Mathematics 2023-06-09 Xiaohan Hu , Xiaodong Li

Manifold learning has been proven to be an effective method for capturing the implicitly intrinsic structure of non-Euclidean data, in which one of the primary challenges is how to maintain the distortion-free (isometry) of the data…

Machine Learning · Computer Science 2024-09-24 Zihao Chen , Wenyong Wang , Yu Xiang

Consider a regression model with infinitely many parameters and time series errors. We are interested in choosing weights for averaging across generalized least squares (GLS) estimators obtained from a set of approximating models. However,…

Statistics Theory · Mathematics 2016-10-05 Tzu-Chang F. Cheng , Ching-Kang Ing , Shu-Hui Yu

We consider the Lie group PSL(2) (the group of orientation preserving isometries of the hyperbolic plane) and a left-invariant Riemannian metric on this group with two equal eigenvalues that correspond to space-like eigenvectors (with…

Differential Geometry · Mathematics 2018-05-15 A. V. Podobryaev , Yu. L. Sachkov

A key challenge in environmental health research is unmeasured spatial confounding, driven by unobserved spatially structured variables that influence both treatment and outcome. A common approach is to fit a spatial regression that models…

Methodology · Statistics 2025-12-23 Sophie M. Woodward , Francesca Dominici , Jose R. Zubizarreta

Machine learning models and libraries can train datasets of different sizes and perform prediction and classification operations, but machine learning models and libraries cause slow and long training times on large datasets. This article…

Machine Learning · Computer Science 2025-09-17 Halil Hüseyin Çalışkan , Talha Koruk

Nonparametric methods have been very popular in the last couple of decades in time series and regression, but no such development has taken place for spatial models. A rather obvious reason for this is the curse of dimensionality. For…

Statistics Theory · Mathematics 2007-06-13 Jiti Gao , Zudi Lu , Dag Tjøstheim

A connection between the General Linear Model (GLM) in combination with classical statistical inference and the machine learning (MLE)-based inference is described in this paper. Firstly, the estimation of the GLM parameters is expressed as…

Machine Learning · Statistics 2022-02-10 Juan Manuel Gorriz , SIPBA group , John Suckling

The immense computational cost of traditional numerical weather and climate models has sparked the development of machine learning (ML) based emulators. Because ML methods benefit from long records of training data, it is common to use…

Machine Learning · Computer Science 2023-09-25 Timothy A. Smith , Stephen G. Penny , Jason A. Platt , Tse-Chun Chen

This paper deals with subspace estimation in the small sample size regime, where the number of samples is comparable in magnitude with the observation dimension. The traditional estimators, mostly based on the sample correlation matrix, are…

Methodology · Statistics 2015-06-19 Pascal Vallet , Xavier Mestre , Philippe Loubaton

Recently, the Multilinear Compressive Learning (MCL) framework was proposed to efficiently optimize the sensing and learning steps when working with multidimensional signals, i.e. tensors. In Compressive Learning in general, and in MCL in…

Computer Vision and Pattern Recognition · Computer Science 2020-09-23 Dat Thanh Tran , Moncef Gabbouj , Alexandros Iosifidis

Minimum Description Length (MDL) is an important principle for induction and prediction, with strong relations to optimal Bayesian learning. This paper deals with learning non-i.i.d. processes by means of two-part MDL, where the underlying…

Information Theory · Computer Science 2007-07-13 Jan Poland , Marcus Hutter

We propose to estimate a metamodel and the sensitivity indices of a complex model m in the Gaussian regression framework. Our approach combines methods for sensitivity analysis of complex models and statistical tools for sparse…

Statistics Theory · Mathematics 2019-11-19 Sylvie Huet , Marie-Luce Taupin

Dimensionality reduction is a fundamental task that aims to simplify complex data by reducing its feature dimensionality while preserving essential patterns, with core applications in data analysis and visualisation. To preserve the…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Thomas Dagès , Simon Weber , Ya-Wei Eileen Lin , Ronen Talmon , Daniel Cremers , Michael Lindenbaum , Alfred M. Bruckstein , Ron Kimmel

We study the Teichm\"uller metric on the Teichm\"uller space of a surface of finite type, in regions where the injectivity radius of the surface is small. The main result is that in such regions the Teichm\"uller metric is approximated up…

Geometric Topology · Mathematics 2016-09-06 Yair Minsky

In this paper, we introduce a novel high-dimensional Factor-Adjusted sparse Partially Linear regression Model (FAPLM), to integrate the linear effects of high-dimensional latent factors with the nonparametric effects of low-dimensional…

Methodology · Statistics 2025-01-14 Yanmei Shi , Meiling Hao , Yanlin Tang , Xu Guo

Nonlinear dimensionality reduction methods provide a valuable means to visualize and interpret high-dimensional data. However, many popular methods can fail dramatically, even on simple two-dimensional manifolds, due to problems such as…

Machine Learning · Statistics 2020-07-08 Daniel Ting , Michael I. Jordan

In this article we study the asymptotic behaviour of the least square estimator in a linear regression model based on random observation instances. We provide mild assumptions on the moments and dependence structure on the randomly spaced…

Statistics Theory · Mathematics 2021-10-07 Karine Bertin , Soledad Torres , Lauri Viitasaari

We demonstrate an equivalence between reproducing kernel Hilbert space (RKHS) embeddings of conditional distributions and vector-valued regressors. This connection introduces a natural regularized loss function which the RKHS embeddings…

Machine Learning · Computer Science 2012-07-25 Steffen Grünewälder , Guy Lever , Luca Baldassarre , Sam Patterson , Arthur Gretton , Massimilano Pontil

This study explores the estimation of parameters in a matrix-valued linear regression model, where the $T$ responses $(Y_t)_{t=1}^T \in \mathbb{R}^{n \times p}$ and predictors $(X_t)_{t=1}^T \in \mathbb{R}^{m \times q}$ satisfy the…

Statistics Theory · Mathematics 2025-12-08 Nayel Bettache