中文
相关论文

相关论文: Matrix Healy Plot: A Practical Tool for Visual Ass…

200 篇论文

The monitoring and management of high-volume feature-rich traffic in large networks offers significant challenges in storage, transmission and computational costs. The predominant approach to reducing these costs is based on performing a…

机器学习 · 计算机科学 2016-06-16 Tingshan Huang , Harish Sethu , Nagarajan Kandasamy

In ordinary Dimensionality Reduction (DR), each data instance in a high dimensional space (original space), or on a distance matrix denoting original space distances, is mapped to (projected onto) one point in a low dimensional space…

计算机视觉与模式识别 · 计算机科学 2022-06-28 Farshad Barahimi

One of the challenges in analyzing high-dimensional expression data is the detection of important biological signals. A common approach is to apply a dimension reduction method, such as principal component analysis. Typically, after…

定量方法 · 定量生物学 2012-06-05 Andreas Lehrmann , Michael Huber , Aydin C. Polatkan , Albert Pritzkau , Kay Nieselt

This work approaches the multidimensional scaling problem from a novel angle. We introduce a scalable method based on the h-plot, which inherently accommodates asymmetric proximity data. Instead of embedding the objects themselves, the…

应用统计 · 统计学 2026-01-12 Aleix Alcacer , Irene Epifanio

In Data Science, entities are typically represented by single valued measurements. Symbolic Data Analysis extends this framework to more complex structures, such as intervals and histograms, that express internal variability. We propose an…

机器学习 · 统计学 2025-12-16 Diogo Pinheiro , M. Rosário Oliveira , Igor Kravchenko , Lina Oliveira

Data visualizations like charts are fundamental tools for quantitative analysis and decision-making across fields, requiring accurate interpretation and mathematical reasoning. The emergence of Multimodal Large Language Models (MLLMs)…

人工智能 · 计算机科学 2025-08-26 Anku Rani , Aparna Garimella , Apoorv Saxena , Balaji Vasan Srinivasan , Paul Pu Liang

The use of Hilbert curves to visualize massive vector of data is revisited following previous authors. The Hilbert curve mapping preserves locality and makes meaningful representation of the data. We call such visualization as Hilbert…

数据分析、统计与概率 · 物理学 2015-11-30 E. Estevez-Rams , C. Perez-Davidenko , B. Aragón Fernández , R. Lora-Serrano

For the mean vector test in high dimension, Ayyala et al.(2017,153:136-155) proposed new test statistics when the observational vectors are M dependent. Under certain conditions, the test statistics for one-same and two-sample cases were…

统计理论 · 数学 2019-04-23 Seonghun Cho , Johan Lim , Deepak Nag Ayyala , Junyong Park , Anindya Roy

Data-driven problem solving in many real-world applications involves analysis of time-dependent multivariate data, for which dimensionality reduction (DR) methods are often used to uncover the intrinsic structure and features of the data.…

人机交互 · 计算机科学 2021-10-28 Takanori Fujiwara , Shilpika , Naohisa Sakamoto , Jorji Nonaka , Keiji Yamamoto , Kwan-Liu Ma

We present analytical expressions for the means and covariances of the sample distribution of the cross-validated Mahalanobis distance. This measure has proven to be especially useful in the context of representational similarity analysis…

应用统计 · 统计学 2016-07-06 Jörn Diedrichsen , Serge Provost , Hossein Zareamoghaddam

Time series visualization plays a crucial role in identifying patterns and extracting insights across various domains. However, as datasets continue to grow in size, visualizing them effectively becomes challenging. Downsampling, which…

人机交互 · 计算机科学 2023-04-04 Jonas Van Der Donckt , Jeroen Van Der Donckt , Michael Rademaker , Sofie Van Hoecke

Multidimensional scaling visualizes dissimilarities among objects and reduces data dimensionality. While many methods address symmetric proximity data, asymmetric and especially three-way proximity data (capturing relationships across…

统计方法学 · 统计学 2025-11-21 Aleix Alcacer , Rafael Benitez , Vicente J. Bolos , Irene Epifanio

This work addresses the challenges of robust covariance estimation and interpretable outlier detection for multivariate functional data with separable covariance structure. We develop a method that simultaneously improves robustness and…

统计方法学 · 统计学 2026-05-21 Marcus Mayrhofer , Una Radojičić , Horst Lewitschnig , Peter Filzmoser

Maps have long been been used to visualise estimates of spatial variables, in particular disease burden and risk. Predictions made using a geostatistical model have uncertainty that typically varies spatially. However, this uncertainty is…

应用统计 · 统计学 2020-05-26 Aimee R Taylor , James A Watson , Caroline O Buckee

Recent studies on plant disease diagnosis using machine learning (ML) have highlighted concerns about the overestimated diagnostic performance due to inappropriate data partitioning, where training and test datasets are derived from the…

计算机视觉与模式识别 · 计算机科学 2025-01-03 Yuji Arima , Satoshi Kagiwada , Hitoshi Iyatomi

In this paper, we present Hi-D maps, a novel method for the visualization of multi-dimensional categorical data. Our work addresses the scarcity of techniques for visualizing a large number of data-dimensions in an effective and…

图形学 · 计算机科学 2025-07-11 Radi Muhammad Reza , Benjamin A Watson

Square contingency tables are a special case commonly used in various fields to analyze categorical data. Although several analysis methods have been developed to examine marginal homogeneity (MH) in these tables, existing measures are…

统计方法学 · 统计学 2023-07-04 Satoru Shinoda , Takuya Yoshimoto , Kouji Tahata

Visualization of Machine Learning (ML) models is an important part of the ML process to enhance the interpretability and prediction accuracy of the ML models. This paper proposes a new method SPC-DT to visualize the Decision Tree (DT) as…

机器学习 · 计算机科学 2022-05-10 Alex Worland , Sridevi Wagle , Boris Kovalerchuk

Classical metric and non-metric multidimensional scaling (MDS) variants are widely known manifold learning (ML) methods which enable construction of low dimensional representation (projections) of high dimensional data inputs. However,…

数据分析、统计与概率 · 物理学 2014-06-16 Denis Horvath , Jozef Ulicny , Branislav Brutovsky

One aim of data mining is the identification of interesting structures in data. For better analytical results, the basic properties of an empirical distribution, such as skewness and eventual clipping, i.e. hard limits in value ranges, need…

应用统计 · 统计学 2020-09-08 Michael C. Thrun , Tino Gehlert , Alfred Ultsch