English
Related papers

Related papers: Identifying statistical dependence in genomic sequ…

200 papers

Inferring functional relationships within complex networks from static snapshots of a subset of variables is a ubiquitous problem in science. For example, a key challenge of systems biology is to translate cellular heterogeneity data…

Molecular Networks · Quantitative Biology 2024-08-09 Euan Joly-Smith , Zitong Jerry Wang , Andreas Hilfinger

RNA-Seq technology allows for studying the transcriptional state of the cell at an unprecedented level of detail. Beyond quantification of whole-gene expression, it is now possible to disentangle the abundance of individual alternatively…

Genomics · Quantitative Biology 2012-10-11 Barbara Rakitsch , Christoph Lippert , Hande Topa , Karsten Borgwardt , Antti Honkela , Oliver Stegle

Identifying dependency in multivariate data is a common inference task that arises in numerous applications. However, existing nonparametric independence tests typically require computation that scales at least quadratically with the sample…

Methodology · Statistics 2021-07-08 Shai Gorsky , Li Ma

Statistical matching is a technique for integrating two or more data sets when information available for matching records for individual participants across data sets is incomplete. Statistical matching can be viewed as a missing data…

Methodology · Statistics 2015-10-14 Jae-kwang Kim , Emily Berg , Taesung Park

After the completion of human genome sequence was anounced, it is evident that interpretation of DNA sequences is an immediate task to work on. For understanding their signals, improvement of present sequence analysis tools and developing…

Computational Complexity · Computer Science 2007-05-23 Gene Kim , MyungHo Kim

There has been a lot of work fitting Ising models to multivariate binary data in order to understand the conditional dependency relationships between the variables. However, additional covariates are frequently recorded together with the…

Machine Learning · Statistics 2012-09-28 Jie Cheng , Elizaveta Levina , Pei Wang , Ji Zhu

In The Cancer Genome Atlas (TCGA) data set, there are many interesting nonlinear dependencies between pairs of genes that reveal important relationships and subtypes of cancer. Such genomic data analysis requires a rapid, powerful and…

Applications · Statistics 2022-11-30 Siqi Xiang , Wan Zhang , Siyao Liu , Katherine A. Hoadley , Charles M. Perou , Kai Zhang , J. S. Marron

At the core of high throughput DNA sequencing platforms lies a bio-physical surface process that results in a random geometry of clusters of homogenous short DNA fragments typically hundreds of base pairs long - bridge amplification. The…

Genomics · Quantitative Biology 2015-08-13 Eliza O'Reilly , Francois Baccelli , Gustavo de Veciana , Haris Vikalo

High-throughput RNA sequencing (RNA-seq) is now the standard method to determine differential gene expression. Identifying differentially expressed genes crucially depends on estimates of read count variability. These estimates are…

Haplotypes, the global patterns of DNA sequence variation, have important implications for identifying complex traits. Recently, blocks of limited haplotype diversity have been discovered in human chromosomes, intensifying the research on…

Genomics · Quantitative Biology 2012-07-19 Nebojsa Jojic , Vladimir Jojic , David Heckerman

The partial information decomposition (PID) is perhaps the leading proposal for resolving information shared between a set of sources and a target into redundant, synergistic, and unique constituents. Unfortunately, the PID framework has…

Statistical Mechanics · Physics 2018-10-30 Ryan G. James , Jeffrey Emenheiser , James P. Crutchfield

Heterogeneity is a dominant factor in the behaviour of many biological processes. Despite this, it is common for mathematical and statistical analyses to ignore biological heterogeneity as a source of variability in experimental data.…

We use a well known model (T. Vicsek et al. Phys Rev Lett 15, 1226 (1995)) for flocking to test mutual information as a tool for detecting order-disorder transitions, in particular when observations of the system are limited. We show that…

Data Analysis, Statistics and Probability · Physics 2009-11-13 R. T. Wicks , S. C. Chapman , R. O. Dendy

Dependency parsing research, which has made significant gains in recent years, typically focuses on improving the accuracy of single-tree predictions. However, ambiguity is inherent to natural language syntax, and communicating such…

Computation and Language · Computer Science 2018-04-18 Katherine A. Keith , Su Lin Blodgett , Brendan O'Connor

The test of independence is a crucial component of modern data analysis. However, traditional methods often struggle with the complex dependency structures found in high-dimensional data. To overcome this challenge, we introduce a novel…

Methodology · Statistics 2024-09-13 Mingshuo Liu , Doudou Zhou , Hao Chen

The analysis of correlations of amino acid occurrences in globular proteins has led to the development of statistical tools that can identify native contacts -- portions of the chains that come to close distance in folded structural…

Biomolecules · Quantitative Biology 2014-07-28 Rocío Espada , R. Gonzalo Parra , Thierry Mora , Aleksandra M. Walczak , Diego Ferreiro

High dimensional random dynamical systems are ubiquitous, including -- but not limited to -- cyber-physical systems, daily return on different stocks of S&P 1500 and velocity profile of interacting particle systems around McKeanVlasov…

Statistics Theory · Mathematics 2023-10-17 Muhammad Abdullah Naeem , Amir Khazraei , Miroslav Pajic

The robust detection of statistical dependencies between the components of a complex system is a key step in gaining a network-based understanding of the system. Because of their simplicity and low computation cost, pairwise statistics are…

Statistics Theory · Mathematics 2019-08-01 Antoine Messager , Nicos Georgiou , Luc Berthouze

Causal discovery, the task of inferring causal structure from data, has the potential to uncover mechanistic insights from biological experiments, especially those involving perturbations. However, causal discovery algorithms over larger…

Machine Learning · Computer Science 2025-04-01 Menghua Wu , Yujia Bao , Regina Barzilay , Tommi Jaakkola

Statistical methods for analyzing large-scale biomolecular data are commonplace in computational biology. A notable example is phenotype prediction from gene expression data, for instance, detecting human cancers, differentiating subtypes…

Genomics · Quantitative Biology 2014-11-24 Bahman Afsari , Ulisses M. Braga-Neto , Donald Geman
‹ Prev 1 3 4 5 6 7 10 Next ›