中文
相关论文

相关论文: Subsampling Winner Algorithm for Feature Selection…

200 篇论文

Various unsupervised greedy selection methods have been proposed as computationally tractable approximations to the NP-hard subset selection problem. These methods rely on sequentially selecting the variables that best improve performance…

机器学习 · 计算机科学 2025-02-18 Federico Zocco , Marco Maggipinto , Gian Antonio Susto , Seán McLoone

Building compact convolutional neural networks (CNNs) with reliable performance is a critical but challenging task, especially when deploying them in real-world applications. As a common approach to reduce the size of CNNs, pruning methods…

机器学习 · 计算机科学 2020-05-26 Hang Li , Chen Ma , Wei Xu , Xue Liu

Unsupervised feature selection (UFS) is widely applied in machine learning and pattern recognition. However, most of the existing methods only consider a single sparsity, which makes it difficult to select valuable and discriminative…

最优化与控制 · 数学 2025-01-03 Xianchao Xiu , Anning Yang , Chenyi Huang , Xinrong Li , Wanquan Liu

Feature selection methods are widely used to address the high computational overheads and curse of dimensionality in classifying high-dimensional data. Most conventional feature selection methods focus on handling homogeneous features,…

机器学习 · 计算机科学 2021-11-17 Xuyang Yan , Mrinmoy Sarkar , Biniam Gebru , Shabnam Nazmi , Abdollah Homaifar

Feature selection prepares the AI-readiness of data by eliminating redundant features. Prior research falls into two primary categories: i) Supervised Feature Selection, which identifies the optimal feature subset based on their relevance…

机器学习 · 计算机科学 2024-03-08 Xinyuan Wang , Dongjie Wang , Wangyang Ying , Rui Xie , Haifeng Chen , Yanjie Fu

In this paper, we apply the Feature Space Decomposition (FSD) method developed in [LS24, GLS25, LSSW26, ALSS26] to obtain, under fairly general conditions, matching upper and lower bounds for the population excess risk of spectral methods…

统计理论 · 数学 2026-05-18 Guillaume Lecué , Zhifan Li , Zong Shang

Model substructure learning aims to find an invariant network substructure that can have better out-of-distribution (OOD) generalization than the original full structure. Existing works usually search the invariant substructure using…

机器学习 · 计算机科学 2023-06-05 Yingchun Wang , Jingcai Guo , Yi Liu , Song Guo , Weizhan Zhang , Xiangyong Cao , Qinghua Zheng

Decoupling representation learning and classifier learning has been shown to be effective in classification with long-tailed data. There are two main ingredients in constructing a decoupled learning scheme; 1) how to train the feature…

机器学习 · 计算机科学 2023-04-20 Giung Nam , Sunguk Jang , Juho Lee

Feature selection and attribute reduction are crucial problems, and widely used techniques in the field of machine learning, data mining and pattern recognition to overcome the well-known phenomenon of the Curse of Dimensionality, by either…

机器学习 · 计算机科学 2018-11-26 Javad Rahimipour Anaraki , Saeed Samet , Mahdi Eftekhari , Chang Wook Ahn

Many statistical methods have been proposed for variable selection in the past century, but few balance inference and prediction tasks well. Here we report on a novel variable selection approach called Penalized regression with…

统计方法学 · 统计学 2021-06-16 Yi Zuo , Thomas G. Stewart , Jeffrey D. Blume

Stochastic variational inference (SVI) is emerging as the most promising candidate for scaling inference in Bayesian probabilistic models to large datasets. However, the performance of these methods has been assessed primarily in the…

机器学习 · 统计学 2015-06-29 Amar Shah , David A. Knowles , Zoubin Ghahramani

Subsampling is a computationally efficient and scalable method to draw inference in large data settings based on a subset of the data rather than needing to consider the whole dataset. When employing subsampling techniques, a crucial…

统计方法学 · 统计学 2025-10-08 Amalan Mahendran , Helen Thompson , James M. McGree

Classification and probability estimation are fundamental tasks with broad applications across modern machine learning and data science, spanning fields such as biology, medicine, engineering, and computer science. Recent development of…

统计方法学 · 统计学 2026-03-25 Liyun Zeng , Hao Helen Zhang

In this paper, a robust weighted score for unbalanced data (ROWSU) is proposed for selecting the most discriminative feature for high dimensional gene expression binary classification with class-imbalance problem. The method addresses one…

机器学习 · 统计学 2024-01-24 Zardad Khan , Amjad Ali , Saeed Aldahmani

This paper proposes a novel Extended Particle Swarm Optimization model (EPSO) that potentially enhances the search process of PSO for optimization problem. Evidently, gene expression profiles are significantly important measurement factor…

神经与进化计算 · 计算机科学 2020-08-11 Ali Hakem Alsaeedi , Adil L. Albukhnefis , Dhiah Al-Shammary , Muntasir Al-Asfoor

Estimating causal effects from observational data is challenging due to selection bias, which leads to imbalanced covariate distributions across treatment groups. Propensity score-based weighting methods are widely used to address this…

机器学习 · 计算机科学 2025-08-08 Ahmad Saeed Khan , Erik Schaffernicht , Johannes Andreas Stork

The problem of large-scale spatial multiple testing is often encountered in various scientific research fields, where the signals are usually enriched on some regions while sparse on others. To integrate spatial structure information from…

统计方法学 · 统计学 2023-09-28 Pengfei Wang , Pengyu Yan , Canhui Li

We consider applying Bayesian Variable Selection Regression, or BVSR, to genome-wide association studies and similar large-scale regression problems. Currently, typical genome-wide association studies measure hundreds of thousands, or…

应用统计 · 统计学 2011-10-28 Yongtao Guan , Matthew Stephens

The current study proposes a dimension reduction method, stepwise support vector machine (SVM), to reduce the dimensions of large p small n datasets. The proposed method is compared with other dimension reduction methods, namely, the…

应用统计 · 统计学 2017-11-10 Elizabeth P. Chou , Tzu-Wei Ko

Penalization schemes like Lasso or ridge regression are routinely used to regress a response of interest on a high-dimensional set of potential predictors. Despite being decisive, the question of the relative strength of penalization is…

统计方法学 · 统计学 2018-11-08 Britta Velten , Wolfgang Huber