中文
相关论文

相关论文: Using Exact Tests from Algebraic Statistics in Spa…

200 篇论文

Conformal prediction is a distribution-free framework for uncertainty quantification that replaces point predictions with sets, offering marginal coverage guarantees (i.e., ensuring that the prediction sets contain the true label with a…

We present new families of goodness-of-fit tests of uniformity on a full-dimensional set $W\subset\R^d$ based on statistics related to edge lengths of random geometric graphs. Asymptotic normality of these statistics is proven under the…

统计理论 · 数学 2020-07-20 Bruno Ebner , Franz Nestmann , Matthias Schulte

Disentangled representation learning aims to uncover latent variables underlying the observed data, and generally speaking, rather strong assumptions are needed to ensure identifiability. Some approaches rely on sufficient changes on the…

机器学习 · 计算机科学 2025-03-04 Zijian Li , Shunxing Fan , Yujia Zheng , Ignavier Ng , Shaoan Xie , Guangyi Chen , Xinshuai Dong , Ruichu Cai , Kun Zhang

Sparse functional/longitudinal data have attracted widespread interest due to the prevalence of such data in social and life sciences. A prominent scenario where such data are routinely encountered are accelerated longitudinal studies,…

统计方法学 · 统计学 2024-06-24 Yidong Zhou , Hans-Georg Müller

In many real-world regression tasks, the data distribution is heavily skewed, and models learn predominantly from abundant majority samples while failing to predict minority labels accurately. While imbalanced classification has been…

机器学习 · 计算机科学 2025-09-30 Shayan Alahyari

Signal processing is rich in inherently continuous and often nonlinear applications, such as spectral estimation, optical imaging, and super-resolution microscopy, in which sparsity plays a key role in obtaining state-of-the-art results.…

机器学习 · 计算机科学 2020-03-23 Luiz F. O. Chamon , Yonina C. Eldar , Alejandro Ribeiro

Many modern big data applications feature large scale in both numbers of responses and predictors. Better statistical efficiency and scientific insights can be enabled by understanding the large-scale response-predictor association network…

统计方法学 · 统计学 2017-04-28 Yoshimasa Uematsu , Yingying Fan , Kun Chen , Jinchi Lv , Wei Lin

This article develops a novel data assimilation methodology, addressing challenges that are common in real-world settings, such as severe sparsity of observations, lack of reliable models, and non-stationarity of the system dynamics. These…

最优化与控制 · 数学 2024-11-05 David J. Abers , George Hripcsak , Lena Mamykina , Melike Sirlanci , Esteban G. Tabak

Recovering dynamical equations from observed noisy data is the central challenge of system identification. We develop a statistical mechanics approach to analyze sparse equation discovery algorithms, which typically balance data fit and…

统计力学 · 物理学 2025-09-16 Andrei A. Klishin , Joseph Bakarji , J. Nathan Kutz , Krithika Manohar

Datasets with hundreds of variables and many missing values are commonplace. In this setting, it is both statistically and computationally challenging to detect true predictive relationships between variables and also to suppress false…

机器学习 · 统计学 2018-04-03 Feras Saad , Vikash Mansinghka

The paper studies the asymptotic behaviour of weighted functionals of long-range dependent data over increasing observation windows. Various important statistics, including sample means, high order moments, occupation measures can be given…

统计理论 · 数学 2019-05-27 Tareq Alodat , Andriy Olenko

Defect detection is a critical research area in artificial intelligence. Recently, synthetic data-based self-supervised learning has shown great potential on this task. Although many sophisticated synthesizing strategies exist, little…

计算机视觉与模式识别 · 计算机科学 2023-10-12 Yuxuan Cai , Dingkang Liang , Dongliang Luo , Xinwei He , Xin Yang , Xiang Bai

State-level policy evaluations commonly employ a difference-in-differences (DID) study design; yet within this framework, statistical model specification varies notably across studies. Motivated by applied state-level opioid policy…

Approximate Bayesian computation is a statistical framework that uses numerical simulations to calibrate and compare models. Instead of computing likelihood functions, Approximate Bayesian computation relies on numerical simulations, which…

统计方法学 · 统计学 2016-01-19 Louisiane Lemaire , Flora Jay , I-Hung Lee , Katalin Csilléry , Michael G. B. Blum

We develop a framework for learning sparse nonparametric directed acyclic graphs (DAGs) from data. Our approach is based on a recent algebraic characterization of DAGs that led to a fully continuous program for score-based learning of DAG…

机器学习 · 统计学 2020-03-25 Xun Zheng , Chen Dan , Bryon Aragam , Pradeep Ravikumar , Eric P. Xing

Two-sample tests utilizing a similarity graph on observations are useful for high-dimensional and non-Euclidean data due to their flexibility and good performance under a wide range of alternatives. Existing works mainly focused on sparse…

统计理论 · 数学 2023-11-14 Yejiong Zhu , Hao Chen

Methods of performing anomaly detection on high-dimensional data sets are needed, since algorithms which are trained on data are only expected to perform well on data that is similar to the training data. There are theoretical results on…

机器学习 · 计算机科学 2020-11-13 Forrest Laine , Claire Tomlin

Traditionally, the Dirichlet-multinomial distribution has been recognized as a key model for contingency tables generated by cluster sampling schemes. There are, however, other possible distributions appropriate for these contingency…

统计方法学 · 统计学 2016-09-26 Juana M. Alonso-Revenga , Nirian Martin , Leandro Pardo

Statistical matching is an effective method for estimating causal effects in which treated units are paired with control units with ``similar'' values of confounding covariates prior to performing estimation. In this way, matching helps…

统计方法学 · 统计学 2023-09-13 Sanjeewani Weerasingha , Michael J. Higgins

The model interpretation is essential in many application scenarios and to build a classification model with a ease of model interpretation may provide useful information for further studies and improvement. It is common to encounter with a…

机器学习 · 统计学 2019-01-07 Wan-Ping Nicole Chen , Yuan-chin Ivan Chang