English
Related papers

Related papers: An analysis of penalized interaction models

200 papers

We consider the equivalent problems of estimating the residual variance, the proportion of explained variance $\eta$ and the signal strength in a high-dimensional linear regression model with Gaussian random design. Our aim is to understand…

Methodology · Statistics 2017-03-17 Nicolas Verzelen , Elisabeth Gassiat

We study high-dimensional sparse estimation under three natural constraints: communication constraints, local privacy constraints, and linear measurements (compressive sensing). Without sparsity assumptions, it has been established that…

Data Structures and Algorithms · Computer Science 2022-03-15 Jayadev Acharya , Clément L. Canonne , Ziteng Sun , Himanshu Tyagi

We establish the rate of convergence of distributions of sums of independent identically distributed random variables to the Gaussian distribution in terms of truncated pseudomoments by implementing the idea of Yu. Studnyev for getting…

Probability · Mathematics 2015-08-13 Yuliya Mishura , Yevheniya Munchak , Petro Slyusarchuk

The eigenvalue density for members of the Gaussian orthogonal and unitary ensembles follows the Wigner semi-circle law. If the Gaussian entries are all shifted by a constant amount c/Sqrt(2N), where N is the size of the matrix, in the large…

Mathematical Physics · Physics 2009-04-21 Kevin E. Bassler , Peter J. Forrester , Norman E. Frankel

Analyzing multi-layered graphical models provides insight into understanding the conditional relationships among nodes within layers after adjusting for and quantifying the effects of nodes from other layers. We obtain the penalized maximum…

Methodology · Statistics 2016-01-06 Jiahe Lin , Sumanta Basu , Moulinath Banerjee , George Michailidis

We perform a comprehensive analysis of a collective decision-making model inspired by honeybee behavior. This model integrates individual exploration for option discovery and social interactions for information sharing, while also…

Disordered Systems and Neural Networks · Physics 2024-12-19 David March-Pons , Ezequiel E. Ferrero , M. Carmen Miguel

Mixed-effect models are very popular for analyzing data with a hierarchical structure, e.g. repeated observations within subjects in a longitudinal design, patients nested within centers in a multicenter design. However, recently, due to…

Methodology · Statistics 2019-05-09 Abhik Ghosh , Magne Thoresen

Sparse high dimensional graphical model selection is a popular topic in contemporary machine learning. To this end, various useful approaches have been proposed in the context of $\ell_1$-penalized estimation in the Gaussian framework.…

Computation · Statistics 2022-02-04 Sang-Yun Oh , Onkar Dalal , Kshitij Khare , Bala Rajaratnam

We study full Bayesian procedures for sparse linear regression when errors have a symmetric but otherwise unknown distribution. The unknown error distribution is endowed with a symmetrized Dirichlet process mixture of Gaussians. For the…

Statistics Theory · Mathematics 2019-03-26 Minwoo Chae , Lizhen Lin , David B. Dunson

Learning from implicit feedback is challenging because of the difficult nature of the one-class problem: we can observe only positive examples. Most conventional methods use a pairwise ranking approach and negative samplers to cope with the…

Machine Learning · Computer Science 2021-05-12 Riku Togashi , Masahiro Kato , Mayu Otani , Tetsuya Sakai , Shin'ichi Satoh

In causal inference, and specifically in the \textit{Causes of Effects} problem, one is interested in how to use statistical evidence to understand causation in an individual case, and so how to assess the so-called {\em probability of…

Methodology · Statistics 2018-10-23 Fabio Corradi , Monica Musio

The problem of constructing confidence sets in the high-dimensional linear model with $n$ response variables and $p$ parameters, possibly $p\ge n$, is considered. Full honest adaptive inference is possible if the rate of sparse estimation…

Statistics Theory · Mathematics 2013-12-19 Richard Nickl , Sara van de Geer

We propose a unified framework for likelihood-based regression modeling when the response variable has finite support. Our work is motivated by the fact that, in practice, observed data are discrete and bounded. The proposed methods assume…

Methodology · Statistics 2022-09-13 Karl Oskar Ekvall , Matteo Bottai

We derive uniform convergence rates for the maximum likelihood estimator and minimax lower bounds for parameter estimation in two-component location-scale Gaussian mixture models with unequal variances. We assume the mixing proportions of…

Statistics Theory · Mathematics 2020-06-02 Tudor Manole , Nhat Ho

Univariate and multivariate general linear regression models, subject to linear inequality constraints, arise in many scientific applications. The linear inequality restrictions on model parameters are often available from phenomenological…

Methodology · Statistics 2021-12-07 Solmaz Seifollahi , Kaniav Kamary , Hossein Bevrani

As data-driven methods are deployed in real-world settings, the processes that generate the observed data will often react to the decisions of the learner. For example, a data source may have some incentive for the algorithm to provide a…

Machine Learning · Computer Science 2023-04-26 Roy Dong , Heling Zhang , Lillian J. Ratliff

This paper deals with empirical processes of the type \[C_n(B)=\sqrt{n}\{\mu_n(B)-P(X_{n+1}\in B\mid X_1,...,X_n)\},\] where $(X_n)$ is a sequence of random variables and $\mu_n=(1/n)\sum_{i=1}^n\delta_{X_i}$ the empirical measure.…

Statistics Theory · Mathematics 2010-01-14 Patrizia Berti , Irene Crimaldi , Luca Pratelli , Pietro Rigo

We consider an optimization problem with strongly convex objective and linear inequalities constraints. To be able to deal with a large number of constraints we provide a penalty reformulation of the problem. As penalty functions we use a…

Optimization and Control · Mathematics 2020-04-29 Angelia Nedich , Tatiana Tatarenko

A common approach to statistical learning with big-data is to randomly split it among $m$ machines and learn the parameter of interest by averaging the $m$ individual estimates. In this paper, focusing on empirical risk minimization, or…

Machine Learning · Statistics 2016-06-14 Jonathan Rosenblatt , Boaz Nadler

In this paper we study the joint distributional convergence of the largest eigenvalues of the sample covariance matrix of a $p$-dimensional time series with iid entries when $p$ converges to infinity together with the sample size $n$. We…

Probability · Mathematics 2016-08-26 Johannes Heiny , Thomas Mikosch