English
Related papers

Related papers: An information criterion for model selection with …

200 papers

This study presents an efficient approach for incomplete data classification, where the entries of samples are missing or masked due to privacy preservation. To deal with these incomplete data, a new kernel function with asymmetric…

Machine Learning · Computer Science 2016-11-22 Bo-Wei Chen

Species distribution models (SDMs) are widely used to predict species' geographic distributions, serving as critical tools for ecological research and conservation planning. Typically, SDMs relate species occurrences to environmental…

Machine Learning · Computer Science 2025-08-12 Hager Radi Abdelwahed , Mélisande Teng , Robin Zbinden , Laura Pollock , Hugo Larochelle , Devis Tuia , David Rolnick

The predictability of a time series is determined by the sensitivity to initial conditions of its data generating process. In this paper our goal is to characterize this sensitivity from a finite sample by assuming few hypotheses on the…

Chaotic Dynamics · Physics 2012-12-13 Quentin Giai Gianetto , Jean-Marc Le Caillec , Erwan Marrec

In statistical classification and machine learning, classification error is an important performance measure, which is minimized by the Bayes decision rule. In practice, the unknown true distribution is usually replaced with a model…

Machine Learning · Computer Science 2025-01-28 Zijian Yang , Vahe Eminyan , Ralf Schlüter , Hermann Ney

Although G\"odel's incompleteness theorem made mathematician recognize that no axiomatic system could completely prove its correctness and that there is an eternal hole between our knowledge and the world, physicists so far continue to work…

Statistical Mechanics · Physics 2007-05-23 Qiuping A. Wang

The diffusion model has shown remarkable performance in modeling data distributions and synthesizing data. However, the vanilla diffusion model requires complete or fully observed data for training. Incomplete data is a common issue in…

Machine Learning · Computer Science 2023-07-04 Yidong Ouyang , Liyan Xie , Chongxuan Li , Guang Cheng

Recently, a so-called E-MS algorithm was developed for model selection in the presence of missing data. Specifically, it performs the Expectation step (E step) and Model Selection step (MS step) alternately to find the minimum point of the…

Methodology · Statistics 2021-06-22 Ping-Feng Xu , Lai-Xu Shang , Man-Lai Tang , Na Shan , Guoliang Tian

In traditional logistic regression models, the link function is often assumed to be linear and continuous in predictors. Here, we consider a threshold model that all continuous features are discretized into ordinal levels, which further…

Methodology · Statistics 2022-02-18 Yinan Lin , Wen Zhou , Zhi Geng , Gexin Xiao , Jianxin Yin

We consider the task of identifying and estimating a parameter of interest in settings where data is missing not at random (MNAR). In general, such parameters are not identified without strong assumptions on the missing data model. In this…

Methodology · Statistics 2024-02-29 Zixiao Wang , AmirEmad Ghassami , Ilya Shpitser

In this paper, we present a method of maximum a posteriori estimation of parameters in dynamic factor models with incomplete data. We extend maximum likelihood expectation maximization iterations by Ba\'nbura & Modugno (2014) to penalized…

Methodology · Statistics 2022-10-14 Erik Spånberg

Data with missing values is ubiquitous in many applications. Recent years have witnessed increasing attention on prediction with only incomplete data consisting of observed features and a mask that indicates the missing pattern. Existing…

Machine Learning · Computer Science 2023-05-22 Yichen Zhu , Jian Yuan , Bo Jiang , Tao Lin , Haiming Jin , Xinbing Wang , Chenghu Zhou

Data imputation, the process of filling in missing feature elements for incomplete data sets, plays a crucial role in data-driven learning. A fundamental belief is that data imputation is helpful for learning performance, and it follows…

Machine Learning · Computer Science 2025-09-30 Ruikai Yang , Fan He , Mingzhen He , Kaijie Wang , Xiaolin Huang

Nonignorable missing outcomes are common in real world datasets and often require strong parametric assumptions to achieve identification. These assumptions can be implausible or untestable, and so we may forgo them in favour of partially…

Methodology · Statistics 2023-10-19 Daniel Daly-Grafstein , Paul Gustafson

We study model selection and model averaging in generalized additive partial linear models (GAPLMs). Polynomial spline is used to approximate nonparametric functions. The corresponding estimators of the linear parameters are shown to be…

Statistics Theory · Mathematics 2011-03-09 Xinyu Zhang , Hua Liang

In quantum multi-parameter estimation, the precision of estimating unknown parameters is bounded by the Cramer-Rao bound (CRB), defined via the inverse of the Fisher information matrix (FIM). However, in certain scenarios such as…

Quantum Physics · Physics 2025-11-18 Min Namkung , Changhyoup Lee , Hyang-Tag Lim

The Fisher Information matrix is a widely used measure for applications ranging from statistical inference, information geometry, experiment design, to the study of criticality in biological systems. Yet there is no commonly accepted…

Computation · Statistics 2016-02-17 Omri Har Shemesh , Rick Quax , Borja Miñano , Alfons G. Hoekstra , Peter M. A. Sloot

The maximal information coefficient (MIC), which measures the amount of dependence between two variables, is able to detect both linear and non-linear associations. However, computational cost grows rapidly as a function of the dataset…

Information Theory · Computer Science 2015-08-18 Ali Mousavi , Richard G. Baraniuk

Information theory is a powerful framework to capture aspects of dynamical systems with multiple degrees of freedom. Mathematically, the dynamics can be represented as a continuous curve $\mathcal{C}$ on a suitable hyperplane in flat space…

Information Theory · Computer Science 2026-04-28 Mattia Carrino , Stefan Hohenegger

We prove lower bounds on the error of any estimator for the mean of a real probability distribution under the knowledge that the distribution belongs to a given set. We apply these lower bounds both to parametric and nonparametric…

Statistics Theory · Mathematics 2024-03-05 Rémy Degenne , Timothée Mathieu

In compositional data, detecting which part of the whole delineates heterogeneity is important. The aim is to propose a procedure to quantify this term in the multivariate regression context without abandoning the data's natural…

Methodology · Statistics 2023-02-21 P Solano
‹ Prev 1 8 9 10 Next ›