English
Related papers

Related papers: Dimension Reduction for Data with Heterogeneous Mi…

200 papers

Dimensionality reduction is a main step in the learning process which plays an essential role in many applications. The most popular methods in this field like SVD, PCA, and LDA, only can be applied to data with vector format. This means…

Machine Learning · Computer Science 2019-03-01 Soheil Ahmadi , Mansoor Rezghi

Principal component analysis is a useful dimension reduction and data visualization method. However, in high dimension, low sample size asymptotic contexts, where the sample size is fixed and the dimension goes to infinity,a paradox has…

Applications · Statistics 2012-11-21 Dan Shen , Haipeng Shen , Hongtu Zhu , J. S. Marron

One classical canon of statistics is that large models are prone to overfitting, and model selection procedures are necessary for high dimensional data. However, many overparameterized models, such as neural networks, perform very well in…

Machine Learning · Statistics 2021-01-05 Xi Chen , Qiang Liu , Xin T. Tong

The critical behavior of the random field $O(N)$ model driven at a uniform velocity is investigated at zero-temperature. From naive phenomenological arguments, we introduce a dimensional reduction property, which relates the large-scale…

Statistical Mechanics · Physics 2017-11-22 Taiki Haga

Dimensionality reduction (DR) plays a vital role in the visual analysis of high-dimensional data. One main aim of DR is to reveal hidden patterns that lie on intrinsic low-dimensional manifolds. However, DR often overlooks important…

Machine Learning · Computer Science 2023-02-28 Takanori Fujiwara , Yun-Hsin Kuo , Anders Ynnerman , Kwan-Liu Ma

Dimensionality reduction can be applied to hyperspectral images so that the most useful data can be extracted and processed more quickly. This is critical in any situation in which data volume exceeds the capacity of the computational…

Image and Video Processing · Electrical Eng. & Systems 2024-02-27 Daniela Lupu , Joseph L. Garrett , Tor Arne Johansen , Milica Orlandic , Ion Necoara

The unprecedented prowess of measurement techniques provides a detailed, multi-scale look into the depths of living systems. Understanding these avalanches of high-dimensional data -- by distilling underlying principles and mechanisms --…

Other Quantitative Biology · Quantitative Biology 2021-08-16 Jean-Pierre Eckmann , Tsvi Tlusty

Heteroscedasticity testing is of importance in regression analysis. Existing local smoothing tests suffer severely from curse of dimensionality even when the number of covariates is moderate because of use of nonparametric estimation. In…

Methodology · Statistics 2015-10-14 Xuehu Zhu , Fei Chen , Xu Guo , Lixing Zhu

We provide new theoretical results in the field of inverse regression methods for dimension reduction. Our approach is based on the study of some empirical processes that lie close to a certain dimension reduction subspace, called the…

Statistics Theory · Mathematics 2015-06-02 François Portier

Fr\'echet regression is becoming a mainstay in modern data analysis for analyzing non-traditional data types belonging to general metric spaces. This novel regression method is especially useful in the analysis of complex health data such…

Methodology · Statistics 2024-10-23 Abdul-Nasah Soale , Congli Ma , Siyu Chen , Obed Koomson

Heterogeneous datasets emerge in various machine learning and optimization applications that feature different input sources, types or formats. Most models or methods do not natively tackle heterogeneity. Hence, such datasets are often…

Machine Learning · Statistics 2025-08-25 Edward Hallé-Hannan , Charles Audet , Youssef Diouane , Sébastien Le Digabel , Paul Saves

Nowadays, massive datasets are typically dispersed across multiple locations, encountering dual challenges of high dimensionality and huge sample size. Therefore, it is necessary to explore sufficient dimension reduction (SDR) methods for…

Methodology · Statistics 2025-09-16 Hongying Li , Minyi Zhu , Yaqi Cao , Xinyi Xu

This paper presents an extensive empirical study on the integration of dimensionality reduction techniques with advanced unsupervised time series anomaly detection models, focusing on the MUTANT and Anomaly-Transformer models. The study…

Machine Learning · Computer Science 2024-03-08 Mahsun Altin , Altan Cakir

Matrix completion is a modern missing data problem where both the missing structure and the underlying parameter are high dimensional. Although missing structure is a key component to any missing data problems, existing matrix completion…

Machine Learning · Statistics 2020-03-23 Xiaojun Mao , Raymond K. W. Wong , Song Xi Chen

Detection of the number of signals corrupted by high-dimensional noise is a fundamental problem in signal processing and statistics. This paper focuses on a general setting where the high-dimensional noise has an unknown complicated…

Statistics Theory · Mathematics 2022-05-16 Xiucai Ding , Fan Yang

In large-scale, data-driven applications, parameters are often only known approximately due to noise and limited data samples. In this paper, we focus on high-dimensional optimization problems with linear constraints under uncertain…

Optimization and Control · Mathematics 2024-03-01 Naqi Huang , Nestor Parolya , Theresia van Essen

Solutions of symbolic regression problems are expressions that are composed of input variables and operators from a finite set of function symbols. One measure for evaluating symbolic regression algorithms is their ability to recover…

Machine Learning · Computer Science 2025-06-25 Paul Kahlmeyer , Markus Fischer , Joachim Giesen

It is now practically the norm for data to be very high dimensional in areas such as genetics, machine vision, image analysis and many others. When analyzing such data, parametric models are often too inflexible while nonparametric…

Methodology · Statistics 2011-05-31 Abhishek Bhattacharya , Garritt Page , David Dunson

In applications involving ordinal predictors, common approaches to reduce dimensionality are either extensions of unsupervised techniques such as principal component analysis, or variable selection procedures that rely on modeling the…

Statistics Theory · Mathematics 2017-10-13 Liliana Forzani , Rodrigo García Arancibia , Pamela Llop , Diego Tomassi

Suppose the data consist of a set $S$ of points $x_j, 1 \leq j \leq J$, distributed in a bounded domain $D \subset R^N$, where $N$ and $J$ are large numbers. In this paper an algorithm is proposed for checking whether there exists a…

Information Theory · Computer Science 2017-02-02 A. G. Ramm , C. Van
‹ Prev 1 4 5 6 7 8 10 Next ›