中文
相关论文

相关论文: A stochastic second-order generalized estimating e…

200 篇论文

Missing exposure information is a very common feature of many observational studies. Here we study identifiability and efficient estimation of causal effects on vector outcomes, in such cases where treatment is unconfounded but partially…

统计方法学 · 统计学 2020-02-04 Edward H. Kennedy

Randomized Controlled Trials (RCTs) are often considered the gold standard for estimating causal effect, but they may lack external validity when the population eligible to the RCT is substantially different from the target population.…

统计方法学 · 统计学 2023-01-11 Bénédicte Colnet , Julie Josse , Erwan Scornet , Gaël Varoquaux

It is well known that accurate probabilistic predictors can be trained through empirical risk minimisation with proper scoring rules as loss functions. While such learners capture so-called aleatoric uncertainty of predictions, various…

机器学习 · 计算机科学 2023-01-31 Viktor Bengs , Eyke Hüllermeier , Willem Waegeman

In the Correlation Clustering problem, we are given a weighted graph $G$ with its edges labeled as "similar" or "dissimilar" by a binary classifier. The goal is to produce a clustering that minimizes the weight of "disagreements": the sum…

数据结构与算法 · 计算机科学 2021-08-13 Jafar Jafarov , Sanchit Kalhan , Konstantin Makarychev , Yury Makarychev

When analyzing data from randomized clinical trials, covariate adjustment can be used to account for chance imbalance in baseline covariates and to increase precision of the treatment effect estimate. A practical barrier to covariate…

统计方法学 · 统计学 2023-07-04 Chia-Rui Chang , Yue Song , Fan Li , Rui Wang

Here we propose an algorithm, named generalized orthogonal components regression (GOCRE), to explore the relationship between a categorical outcome and a set of massive variables. A set of orthogonal components are sequentially constructed…

统计方法学 · 统计学 2013-04-18 Yanzhu Lin , Min Zhang , Dabao Zhang

We evaluate the misclustering probability of a spectral clustering algorithm under a Gaussian mixture model with a general covariance structure. The algorithm partitions the data into two groups based on the sign of the first principal…

统计理论 · 数学 2026-04-13 Kohei Kawamoto , Yuichi Goto , Koji Tsukuda

In the standard stochastic block model for networks, the probability of a connection between two nodes, often referred to as the edge probability, depends on the unobserved communities each of these nodes belongs to. We consider a flexible…

计量经济学 · 经济学 2024-02-27 Yuichi Kitamura , Louise Laage

Longitudinal studies are often conducted to explore the cohort and age effects in many scientific areas. The within cluster correlation structure plays a very important role in longitudinal data analysis. This is because not only can an…

统计理论 · 数学 2008-12-18 Yan Sun , Wenyang Zhang , Howell Tong

There has been widespread use of causal inference methods for the rigorous analysis of observational studies and to identify policy evaluations. In this article, we consider a class of generalized coarsened procedures for confounding. At a…

统计方法学 · 统计学 2025-07-04 Debashis Ghosh , Lei Wang

In recent years, machine learning and AI have been introduced in many industrial fields. In fields such as finance, medicine, and autonomous driving, where the inference results of a model may have serious consequences, high…

机器学习 · 计算机科学 2021-11-23 Akihisa Watanabe , Michiya Kuramata , Kaito Majima , Haruka Kiyohara , Kensho Kondo , Kazuhide Nakata

Principal Component Analysis (PCA) is a widely used technique in exploratory data analysis, visualization, and data preprocessing, leveraging the concept of variance to identify key dimensions in datasets. In this study, we focus on the…

应用统计 · 统计学 2024-07-03 Reza Dastranj , Martin Kolar

In certain privacy-sensitive scenarios within fields such as clinical trial simulations, federated learning, and distributed learning, researchers often face the challenge of estimating correlations between variables without access to…

统计方法学 · 统计学 2025-08-05 Longwen Shang , Min Tsao , Xuekui Zhang

In this work we introduce a mixture of GPs to address the data association problem, i.e. to label a group of observations according to the sources that generated them. Unlike several previously proposed GP mixtures, the novel mixture has…

机器学习 · 统计学 2011-08-18 Miguel Lázaro-Gredilla , Steven Van Vaerenbergh , Neil Lawrence

First-order stochastic methods are the state-of-the-art in large-scale machine learning optimization owing to efficient per-iteration complexity. Second-order methods, while able to provide faster convergence, have been much less explored…

机器学习 · 统计学 2017-12-01 Naman Agarwal , Brian Bullins , Elad Hazan

Missing data often result in undesirable bias and loss of efficiency. These issues become substantial when the response mechanism is nonignorable, meaning that the response model depends on unobserved variables. To manage nonignorable…

统计方法学 · 统计学 2024-12-30 Kenji Beppu , Jinung Choi , Kosuke Morikawa , Jongho Im

Although randomized experiments are widely regarded as the gold standard for estimating causal effects, missing data of the pretreatment covariates makes it challenging to estimate the subgroup causal effects. When the missing data…

统计理论 · 数学 2014-01-08 Peng Ding , Zhi Geng

Time-course gene expression datasets provide insight into the dynamics of complex biological processes, such as immune response and organ development. It is of interest to identify genes with similar temporal expression patterns because…

统计方法学 · 统计学 2022-01-03 Sara Venkatraman , Sumanta Basu , Andrew G. Clark , Sofie Delbare , Myung Hee Lee , Martin T. Wells

Emotion cause identification aims at identifying the potential causes that lead to a certain emotion expression in text. Several techniques including rule based methods and traditional machine learning methods have been proposed to address…

计算与语言 · 计算机科学 2019-06-05 Zixiang Ding , Huihui He , Mengran Zhang , Rui Xia

Maximum likelihood estimates (MLEs) are asymptotically normally distributed, and this property is used in meta-analyses to test the heterogeneity of estimates, either for a single cluster or for several sub-groups. More recently, MLEs for…

统计理论 · 数学 2022-02-28 Anthony J. Webster