中文
相关论文

相关论文: Multi-sample estimation of centered log-ratio matr…

200 篇论文

Metagenomics sequencing is routinely applied to quantify bacterial abundances in microbiome studies, where the bacterial composition is estimated based on the sequencing read counts. Due to limited sequencing depth and DNA dropouts, many…

统计方法学 · 统计学 2019-04-26 Yuanpei Cao , Anru Zhang , Hongzhe Li

Discrete data such as counts of microbiome taxa resulting from next-generation sequencing are routinely encountered in bioinformatics. Taxa count data in microbiome studies are typically high-dimensional, over-dispersed, and can only reveal…

统计方法学 · 统计学 2022-06-23 Yuan Fang , Sanjeena Subedi

Compositional data, where only relative abundances are available, are common in microbiome and other high-throughput sequencing studies. Log ratios between groups of variables serve as key biomarkers in these settings. However, selecting…

统计方法学 · 统计学 2025-04-02 Jing Ma , Paizhe Xie , Kristyn Pantoja , David E. Jones

In microbiome and genomic studies, the regression of compositional data has been a crucial tool for identifying microbial taxa or genes that are associated with clinical phenotypes. To account for the variation in sequencing depth, the…

统计方法学 · 统计学 2021-03-11 Pixu Shi , Yuchen Zhou , Anru R. Zhang

Microbial communities analysis is drawing growing attention due to the rapid development of high-throughput sequencing techniques nowadays. The observed data has the following typical characteristics: it is high-dimensional, compositional…

统计方法学 · 统计学 2020-04-30 Yong He , Pengfei Liu , Xinsheng Zhang , Wang Zhou

In current applied research the most-used route to an analysis of composition is through log-ratios -- that is, contrasts among log-transformed measurements. Here we argue instead for a more direct approach, using a statistical model for…

统计方法学 · 统计学 2023-12-19 David Firth , Fiona Sammut

High-throughput sequencing technology allows us to test the compositional difference of bacteria in different populations. One important feature of human microbiome data is that it often includes a large number of zeros. Such data can be…

统计方法学 · 统计学 2022-08-23 Wanjie Wang , Eric Z. Chen , Hongzhe Li

Compositional data arise in many areas of research in the natural and biomedical sciences. One prominent example is in the study of the human gut microbiome, where one can measure the relative abundance of many distinct microorganisms in a…

统计方法学 · 统计学 2024-04-26 Aaron J. Molstad , Karl Oskar Ekvall , Piotr M. Suder

Human microbiome studies based on genetic sequencing techniques produce compositional longitudinal data of the relative abundances of microbial taxa over time, allowing to understand, through mixed-effects modeling, how microbial…

统计方法学 · 统计学 2025-12-23 John Barrera , Cristian Meza , Ana Arribas-Gil

This paper investigates the computational and statistical limits in clustering matrix-valued observations. We propose a low-rank mixture model (LrMM), adapted from the classical Gaussian mixture model (GMM) to treat matrix-valued…

统计理论 · 数学 2023-06-08 Zhongyuan Lyu , Dong Xia

Clustering analysis by nonnegative low-rank approximations has achieved remarkable progress in the past decade. However, most approximation approaches in this direction are still restricted to matrix factorization. We propose a new low-rank…

机器学习 · 计算机科学 2012-06-22 Zhirong Yang , Erkki Oja

In natural language processing (NLP), the likelihood ratios (LRs) of N-grams are often estimated from the frequency information. However, a corpus contains only a fraction of the possible N-grams, and most of them occur infrequently. Hence,…

计算与语言 · 计算机科学 2022-04-15 Masato Kikuchi , Mitsuo Yoshida , Kyoji Umemura , Tadachika Ozono

The interactions between microbial taxa in microbiome data has been under great research interest in the science community. In particular, several methods such as SPIEC-EASI, gCoda, and CD-trace have been proposed to model the conditional…

统计方法学 · 统计学 2022-07-05 Chuan Tian , Duo Jiang , Yuan Jiang

Compositional data sets are ubiquitous in science, including geology, ecology, and microbiology. In microbiome research, compositional data primarily arise from high-throughput sequence-based profiling experiments. These data comprise…

统计理论 · 数学 2019-03-05 Patrick L. Combettes , Christian L. Müller

We propose a modification of linear discriminant analysis, referred to as compressive regularized discriminant analysis (CRDA), for analysis of high-dimensional datasets. CRDA is specially designed for feature elimination purpose and can be…

统计方法学 · 统计学 2018-04-12 Muhammad Naveed Tabassum , Esa Ollila

CUR matrix decomposition computes the low rank approximation of a given matrix by using the actual rows and columns of the matrix. It has been a very useful tool for handling large matrices. One limitation with the existing algorithms for…

机器学习 · 计算机科学 2014-11-05 Miao Xu , Rong Jin , Zhi-Hua Zhou

The human microbiome plays an important role in human health and disease status. Next generating sequencing technologies allow for quantifying the composition of the human microbiome. Clustering these microbiome data can provide valuable…

统计方法学 · 统计学 2021-01-07 Wangshu Tu , Sanjeena Subedi

The interpretation of count data originating from the current generation of DNA sequencing platforms requires special attention. In particular, the per-sample library sizes often vary by orders of magnitude from the same sequencing run, and…

定量方法 · 定量生物学 2015-06-17 Paul J. McMurdie , Susan Holmes

With the development of next generation sequencing technology, researchers have now been able to study the microbiome composition using direct sequencing, whose output are bacterial taxa counts for each microbiome sample. One goal of…

应用统计 · 统计学 2013-05-24 Jun Chen , Hongzhe Li

In human microbiome studies, sequencing reads data are often summarized as counts of bacterial taxa at various taxonomic levels specified by a taxonomic tree. This paper considers the problem of analyzing two repeated measurements of…

应用统计 · 统计学 2017-02-17 Pixu Shi , Hongzhe Li
‹ 上一页 1 2 3 10 下一页 ›