English
Related papers

Related papers: On Difference Between Two Types of $\gamma$-diverg…

200 papers

Dimension reduction plays a pivotal role in analysing high-dimensional data. However, observations with missing values present serious difficulties in directly applying standard dimension reduction techniques. As a large number of dimension…

Machine Learning · Statistics 2021-09-28 Yurong Ling , Zijing Liu , Jing-Hao Xue

The problem of identifying the most discriminating features when performing supervised learning has been extensively investigated. In particular, several methods for variable selection in model-based classification have been proposed.…

Applications · Statistics 2020-12-16 Andrea Cappozzo , Francesca Greselin , Thomas Brendan Murphy

Mixtures of regression are a powerful class of models for regression learning with respect to a highly uncertain and heterogeneous response variable of interest. In addition to being a rich predictive model for the response given some…

Statistics Theory · Mathematics 2024-01-30 Dat Do , Linh Do , XuanLong Nguyen

Modern algorithms for binary classification rely on an intermediate regression problem for computational tractability. In this paper, we establish a geometric distinction between classification and regression that allows risk in these two…

Machine Learning · Statistics 2022-05-19 Suhas Vijaykumar , Claire Lazar Reich

We study a discrete-to-continuous Gamma-limit of a family of high-contrast double porosity type functionals defined on a scaled integer lattice. Under periodicity and p-growth conditions we prove the homogenization result and describe the…

Functional Analysis · Mathematics 2014-06-09 Andrea Braides , Valeria Chiado Piat , Andrey Piatnitski

A systematic, comparative investigation into the effects of low-quality data reveals a stark spectrum of robustness across modern probabilistic models. We find that autoregressive language models, from token prediction to…

Artificial Intelligence · Computer Science 2025-12-16 Liu Peng , Yaochu Jin

Sources of bias in empirical studies can be separated in those coming from the modelling domain (e.g. multicollinearity) and those coming from outliers. We propose a two-step approach to counter both issues. First, by decontaminating data…

General Economics · Economics 2019-02-14 Mathias Kloss , Thomas Kirschstein , Steffen Liebscher , Martin Petrick

Functional logistic regression is a popular model to capture a linear relationship between binary response and functional predictor variables. However, many methods used for parameter estimation in functional logistic regression are…

Methodology · Statistics 2025-10-15 Berkay Akturk , Ufuk Beyaztas , Han Lin Shang

This paper considers the two-dataset problem, where data are collected from two potentially different populations sharing common aspects. This problem arises when data are collected by two different types of researchers or from two…

Methodology · Statistics 2022-09-27 Steven N. MacEachern , Koji Miyawaki

A large body of work in the statistics and computer science communities dating back to Huber (Huber, 1960) has led to statistically and computationally efficient outlier-robust estimators. Two particular outlier models have received…

Statistics Theory · Mathematics 2024-11-26 Yeshwanth Cherapanamjeri , Daniel Lee

A popular approach for comparing gene expression levels between (replicated) conditions of RNA sequencing data relies on counting reads that map to features of interest. Within such count-based methods, many flexible and advanced…

Quantitative Methods · Quantitative Biology 2014-03-17 Xiaobei Zhou , Helen Lindsay , Mark D. Robinson

Gaussian process regression (GPR) model is well-known to be susceptible to outliers. Robust process regression models based on t-process or other heavy-tailed processes have been developed to address the problem. However, due to the nature…

Methodology · Statistics 2017-07-10 Wang Zhanfeng , Noh Maengseok , Lee Youngjo , Shi Jianqing

In this paper a robust version of the classical Wald test statistics for linear hypothesis in the logistic regression model is introduced and its properties are explored. We study the problem under the assumption of random covariates…

Statistics Theory · Mathematics 2019-05-09 Ayandrendanath Basu , Abhik Ghosh , Abhijit Mandal , Nirian Martin , Leandro Pardo

We consider a variant of regression problem, where the correspondence between input and output data is not available. Such shuffled data is commonly observed in many real world problems. Taking flow cytometry as an example, the measuring…

Machine Learning · Computer Science 2021-02-12 Yujia Xie , Yixiu Mao , Simiao Zuo , Hongteng Xu , Xiaojing Ye , Tuo Zhao , Hongyuan Zha

Robust tests of general composite hypothesis under non-identically distributed observations is always a challenge. Ghosh and Basu (2018, Statistica Sinica, 28, 1133--1155) have proposed a new class of test statistics for such problems based…

Statistics Theory · Mathematics 2019-01-08 Abhik Ghosh , Ayanendranath Basu

Standard high-dimensional factor models assume that the comovements in a large set of variables could be modeled using a small number of latent factors that affect all variables. In many relevant applications in economics and finance,…

Econometrics · Economics 2022-02-08 Antoine Djogbenou , Razvan Sufana

Data analysis based on information from several sources is common in economic and biomedical studies. This setting is often referred to as the data fusion problem, which differs from traditional missing data problems since no complete data…

Methodology · Statistics 2022-04-07 Wei Li , Shanshan Luo , Wangli Xu

In diagnostic test accuracy meta-analysis (DTA-MA), standard inference methods using bivariate random-effects models for jointly synthesizing sensitivity and specificity can be sensitive to outlying studies and may yield misleading…

Methodology · Statistics 2026-05-01 Kotaro Sasaki , Hisashi Noma , Theodoros Evrenoglou

We study the problem of testing the covariance matrix of a high-dimensional Gaussian in a robust setting, where the input distribution has been corrupted in Huber's contamination model. Specifically, we are given i.i.d. samples from a…

Machine Learning · Computer Science 2021-01-01 Ilias Diakonikolas , Daniel M. Kane

A rigorous connection between large deviations theory and Gamma-convergence is established. Applications include representations formulas for rate functions, a contraction principle for measurable maps, a large deviations principle for…

Probability · Mathematics 2018-02-02 Mauro Mariani
‹ Prev 1 3 4 5 6 7 10 Next ›