中文
相关论文

相关论文: Imputation of mixed data with multilevel singular …

200 篇论文

Standard approaches for variable selection in linear models are not tailored to deal properly with high-dimensional and incomplete data. Currently, methods dedicated to high-dimensional data handle missing values by ad-hoc strategies, like…

统计方法学 · 统计学 2021-06-09 Avner Bar-Hen , Vincent Audigier

In recent years, many methods have been developed for detecting causal relationships in observational data. Some of them have the potential to tackle large data sets. However, these methods fail to discover a combined cause, i.e. a…

人工智能 · 计算机科学 2015-10-16 Saisai Ma , Jiuyong Li , Lin Liu , Thuc Duy Le

Distributed health data networks that use information from multiple sources have drawn substantial interest in recent years. However, missing data are prevalent in such networks and present significant analytical challenges. The current…

密码学与安全 · 计算机科学 2021-01-01 Yi Deng , Xiaoqian Jiang , Qi Long

Fast computation of singular value decomposition (SVD) is of great interest in various machine learning tasks. Recently, SVD methods based on randomized linear algebra have shown significant speedup in this regime. This paper attempts to…

分布式、并行与集群计算 · 计算机科学 2017-06-23 Yuechao Lu , Fumihiko Ino , Yasuyuki Matsushita

Multiple imputation by chained equations (MICE) has emerged as a popular approach for handling missing data. A central challenge for applying MICE is determining how to incorporate outcome information into covariate imputation models,…

统计方法学 · 统计学 2019-10-11 Lauren Beesley , Jeremy M G Taylor

With the abundance of data in recent years, interesting challenges are posed in the area of recommender systems. Producing high quality recommendations with scalability and performance is the need of the hour. Singular Value…

机器学习 · 计算机科学 2019-07-19 Prasad Bhavana , Vikas Kumar , Vineet Padmanabhan

Efficiently computing a subset of a correlation matrix consisting of values above a specified threshold is important to many practical applications. Real-world problems in genomics, machine learning, finance other applications can produce…

统计计算 · 统计学 2016-03-15 James Baglama , Michael Kane , Bryan Lewis , Alex Poliakov

Coupled tensor decompositions (CTDs) perform data fusion by linking factors from different datasets. Although many CTDs have been already proposed, current works do not address important challenges of data fusion, where: 1) the datasets are…

机器学习 · 计算机科学 2024-12-13 Ricardo Augusto Borsoi , Konstantin Usevich , David Brie , Tülay Adali

In this paper we examine data fusion methods for multi-view data classification. We present a decision concept which explicitly takes into account the input multi-view structure, where for each case there is a different subset of relevant…

计算机视觉与模式识别 · 计算机科学 2018-03-20 Yaniv Shachor , Hayit Greenspan , Jacob Goldberger

Genomics data such as RNA gene expression, methylation and micro RNA expression are valuable sources of information for various clinical predictive tasks. For example, predicting survival outcomes, cancer histology type and other patients'…

基因组学 · 定量生物学 2022-05-26 Sophie Peacock , Etai Jacob , Nikolay Burlutskiy

Public policy-makers use cost-effectiveness analyses (CEA) to decide which health and social care interventions to provide. Appropriate methods have not been developed for handling missing data in complex settings, such as for CEA that use…

统计方法学 · 统计学 2012-06-27 Karla Diaz-Ordaz , Michael G. Kenward , Richard Grieve

Health information is generally fragmented across silos. Though it is technically feasible to unite data for analysis in a manner that underpins a rapid learning healthcare system, privacy concerns and regulatory barriers limit data…

机器学习 · 计算机科学 2020-12-10 Dianbo Liu , Kathe Fox , Griffin Weber , Tim Miller

Handling missing data is crucial in machine learning, but many datasets contain gaps due to errors or non-response. Unlike traditional methods such as listwise deletion, which are simple but inadequate, the literature offers more…

密码学与安全 · 计算机科学 2024-05-30 Julia Jentsch , Ali Burak Ünal , Şeyma Selcan Mağara , Mete Akgün

Hierarchically-organized data arise naturally in many psychology and neuroscience studies. As the standard assumption of independent and identically distributed samples does not hold for such data, two important problems are to accurately…

统计理论 · 数学 2018-09-03 Irene Dowding , Stefan Haufe

Classification methods that leverage the strengths of data from multiple sources (multi-view data) simultaneously have enormous potential to yield more powerful findings than two step methods: association followed by classification. We…

统计方法学 · 统计学 2020-01-16 Sandra E. Safo , Eun Jeong Min , Lillian Haine

In this paper, we show that the SVD of a matrix can be constructed efficiently in a hierarchical approach. Our algorithm is proven to recover the singular values and left singular vectors if the rank of the input matrix $A$ is known.…

数值分析 · 数学 2017-01-09 M. A. Iwen , B. W. Ong

Multi-modal medical images provide complementary soft-tissue characteristics that aid in the screening and diagnosis of diseases. However, limited scanning time, image corruption and various imaging protocols often result in incomplete…

计算机视觉与模式识别 · 计算机科学 2024-07-10 Yue Zhang , Chengtao Peng , Qiuli Wang , Dan Song , Kaiyan Li , S. Kevin Zhou

Federated learning paradigm to utilize datasets across multiple data providers. In FL, cross-silo data providers often hesitate to share their high-quality dataset unless their data value can be fairly assessed. Shapley value (SV) has been…

机器学习 · 计算机科学 2025-04-24 Shuyue Wei , Yongxin Tong , Zimu Zhou , Tianran He , Yi Xu

A versatile medical image segmentation model applicable to images acquired with diverse equipment and protocols can facilitate model deployment and maintenance. However, building such a model typically demands a large, diverse, and fully…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Xiaoyang Chen , Hao Zheng , Yuemeng Li , Yuncong Ma , Liang Ma , Hongming Li , Yong Fan

Motivated by two case studies using primary care records from the Clinical Practice Research Datalink, we describe statistical methods that facilitate the analysis of tall data, with very large numbers of observations. Our focus is on…

统计方法学 · 统计学 2018-05-14 Kirsty Rhodes , Rebecca Turner , Rupert Payne , Ian White