中文
相关论文

相关论文: Fast leave-one-cluster-out cross-validation using …

200 篇论文

Recent literature provides many computational and modeling approaches for covariance matrices estimation in a penalized Gaussian graphical models but relatively little study has been carried out on the choice of the tuning parameter. This…

统计方法学 · 统计学 2009-09-08 Heng Lian

We introduce a novel Information Criterion (IC), termed Learning under Singularity (LS), designed to enhance the functionality of the Widely Applicable Bayes Information Criterion (WBIC) and the Singular Bayesian Information Criterion…

机器学习 · 统计学 2024-02-23 Lirui Liu , Joe Suzuki

The Cox proportional hazards model, commonly used in clinical trials, assumes proportional hazards. However, it does not hold when, for example, there is a delayed onset of the treatment effect. In such a situation, an acute change in the…

统计方法学 · 统计学 2022-04-22 Ryoto Ozaki , Yoshiyuki Ninomiya

Network science investigates methodologies that summarise relational data to obtain better interpretability. Identifying modular structures is a fundamental task, and assessment of the coarse-grain level is its crucial step. Here, we…

社会与信息网络 · 计算机科学 2017-06-13 Tatsuro Kawamoto , Yoshiyuki Kabashima

We consider here a classification method that balances two objectives: large similarity within the samples in the cluster, and large dissimilarity between the cluster and its complement. The method, referred to as HNC or SNC, requires seed…

机器学习 · 计算机科学 2025-03-05 Dorit Hochbaum , Torpong Nitayanont

A measure of distance between two clusterings has important applications, including clustering validation and ensemble clustering. Generally, such distance measure provides navigation through the space of possible clusterings. Mostly used…

社会与信息网络 · 计算机科学 2015-09-01 Reihaneh Rabbany , Osmar R. Zaïane

I.I.D. hypothesis between training and testing data is the basis of numerous image classification methods. Such property can hardly be guaranteed in practice where the Non-IIDness is common, causing instable performances of these models. In…

计算机视觉与模式识别 · 计算机科学 2019-08-15 Yue He , Zheyan Shen , Peng Cui

The Integrated Completed Likelihood (ICL) criterion has been proposed by Biernacki et al. (2000) in the model-based clustering framework to select a relevant number of classes and has been used by statisticians in various application areas.…

统计理论 · 数学 2012-06-01 Jean-Patrick Baudry

Based on the classical Degree Corrected Stochastic Blockmodel (DCSBM) model for network community detection problem, we propose two novel approaches: principal component clustering (PCC) and normalized principal component clustering (NPCC).…

机器学习 · 统计学 2020-11-11 Huan Qing , Jingli Wang

Model selection in mixed models based on the conditional distribution is appropriate for many practical applications and has been a focus of recent statistical research. In this paper we introduce the R-package cAIC4 that allows for the…

统计计算 · 统计学 2018-03-20 Benjamin Säfken , David Rügamer , Thomas Kneib , Sonja Greven

We emphasize that it is possible to improve the principle of unbiased risk estimation for model selection by addressing excess risk deviations in the design of penalization procedures. Indeed, we propose a modification of Akaike's…

统计理论 · 数学 2018-07-23 Adrien Saumard , Fabien Navarro

Background. The reliability paradox describes the empirical observation that cognitive tasks producing robust group-level effects often yield poor between-individual reliability. Existing approaches rely predominantly on the intraclass…

统计方法学 · 统计学 2026-05-26 Maria Westrin

Accurate model selection is a fundamental requirement for statistical analysis. In many real-world applications of graphical modelling, correct model structure identification is the ultimate objective. Standard model validation procedures…

机器学习 · 统计学 2019-08-28 Robert O'Shea

The similarity among samples and the discrepancy between clusters are two crucial aspects of image clustering. However, current deep clustering methods suffer from the inaccurate estimation of either feature similarity or semantic…

计算机视觉与模式识别 · 计算机科学 2022-11-23 Chuang Niu , Hongming Shan , Ge Wang

Bayesian model averaging, model selection and its approximations such as BIC are generally statistically consistent, but sometimes achieve slower rates og convergence than other methods such as AIC and leave-one-out cross-validation. On the…

统计理论 · 数学 2008-09-17 Tim van Erven , Peter Grunwald , Steven de Rooij

Inferring cluster structure in microarray datasets is a fundamental task for the -omic sciences. A fundamental question in Statistics, Data Analysis and Classification, is the prediction of the number of clusters in a dataset, usually…

数据结构与算法 · 计算机科学 2011-02-16 Filippo Utro

In segmented regression, when the regression function is continuous at the change-points that are the boundaries of the segments, it is also called joinpoint regression, and the analysis package developed by \cite{KimFFM00} has become a…

统计方法学 · 统计学 2025-06-11 Kazuki Nakajima , Yoshiyuki Ninomiya

We propose the Sobolev Independence Criterion (SIC), an interpretable dependency measure between a high dimensional random variable X and a response variable Y . SIC decomposes to the sum of feature importance scores and hence can be used…

机器学习 · 计算机科学 2019-11-01 Youssef Mroueh , Tom Sercu , Mattia Rigotti , Inkit Padhi , Cicero Dos Santos

We consider the scenario of deep clustering, in which the available prior knowledge is limited. In this scenario, few existing state-of-the-art deep clustering methods can perform well for both non-complex topology and complex topology…

机器学习 · 统计学 2023-03-07 Yuhui Zhang , Yuichiro Wada , Hiroki Waida , Kaito Goto , Yusaku Hino , Takafumi Kanamori

Cross-validation is a popular non-parametric method for evaluating the accuracy of a predictive rule. The usefulness of cross-validation depends on the task we want to employ it for. In this note, I discuss a simple non-parametric setting,…

统计方法学 · 统计学 2019-09-27 Stefan Wager