中文
相关论文

相关论文: Identifying statistical dependence in genomic sequ…

200 篇论文

Traditional statistical theory assumes that the analysis to be performed on a given data set is selected independently of the data themselves. This assumption breaks downs when data are re-used across analyses and the analysis to be…

机器学习 · 计算机科学 2017-06-06 Adam Smith

We consider the reconstruction of a phylogeny from multiple genes under the multispecies coalescent. We establish a connection with the sparse signal detection problem, where one seeks to distinguish between a distribution and a mixture of…

概率论 · 数学 2017-07-24 Elchanan Mossel , Sebastien Roch

Protein structure prediction is one of the most important problems in computational biology. The most successful computational approach, also called template-based modeling, identifies templates with solved crystal structures for the query…

生物大分子 · 定量生物学 2013-06-20 Jian Peng

We present a new method to extract distance and orientation dependent potentials between amino acid side chains using a database of protein structures and the standard Boltzmann device. The importance of orientation dependent interactions…

化学物理 · 物理学 2016-09-08 N. -V. Buchete , J. E. Straub , D. Thirumalai

Identifying disease-indicative genes is critical for deciphering disease mechanisms and has attracted significant interest in biomedical research. Spatial transcriptomics offers unprecedented insights for the detection of disease-specific…

统计方法学 · 统计学 2024-09-05 Qicheng Zhao , Qihuang Zhang

Modeling biological networks serves as both a major goal and an effective tool of systems biology in studying mechanisms that orchestrate the activities of gene products in cells. Biological networks are context specific and dynamic in…

In this paper, a robust non-parametric measure of statistical dependence, or correlation, between two random variables is presented. The proposed coefficient is a permutation-like statistic that quantifies how much the observed sample S_n :…

统计方法学 · 统计学 2020-07-27 Rami Mahdi

We propose a general method to study dependent data in a binary tree, where an individual in one generation gives rise to two different offspring, one of type 0 and one of type 1, in the next generation. For any specific characteristic of…

概率论 · 数学 2009-09-29 Julien Guyon

Existing sequence alignment algorithms use heuristic scoring schemes which cannot be used as objective distance metrics. Therefore one relies on measures like the p- or log-det distances, or makes explicit, and often simplistic, assumptions…

基因组学 · 定量生物学 2015-05-19 Orion Penner , Peter Grassberger , Maya Paczuski

The bioinformatical methods to detect lateral gene transfer events are mainly based on functional coding DNA characteristics. In this paper, we propose the use of DNA traits not depending on protein coding requirements. We introduce several…

神经与进化计算 · 计算机科学 2012-04-13 C. Calderón , L. Delaye , V. Mireles , P. Miramontes

Mutual information is widely used in artificial intelligence, in a descriptive way, to measure the stochastic dependence of discrete random variables. In order to address questions such as the reliability of the empirical value, one must…

人工智能 · 计算机科学 2008-06-26 Marco Zaffalon , Marcus Hutter

Mutual information is widely used in artificial intelligence, in a descriptive way, to measure the stochastic dependence of discrete random variables. In order to address questions such as the reliability of the empirical value, one must…

人工智能 · 计算机科学 2014-08-08 Marco Zaffalon , Marcus Hutter

Next-generation sequencing technology enables the identification of thousands of gene regulatory sequences in many cell types and organisms. We consider the problem of testing if two such sequences differ in their number of binding site…

基因组学 · 定量生物学 2014-02-04 Dennis Kostka , Tara Friedrich , Alisha K. Holloway , Katherine S. Pollard

Quantification of microbial interactions from 16S rRNA and meta-genomic sequencing data is difficult due to their sparse nature, as well as the fact that the data only provides measures of relative abundance. In this paper, we propose using…

统计方法学 · 统计学 2021-11-04 Rebecca A. Deek , Hongzhe Li

We illustrate the use of tools (asymptotic theories of standard error quantification using appropriate statistical models, bootstrapping, model comparison techniques) in addition to sensitivity that may be employed to determine the…

偏微分方程分析 · 数学 2015-03-17 H. T. Banks , M Doumic , C Kruse , S Prigent , H Rezaei

Advances in data collection are producing growing volumes of temporal count observations, making adapted modeling increasingly necessary. In this work, we introduce a generative framework for independent component analysis of temporal count…

统计方法学 · 统计学 2026-01-30 Alexandre Chaussard , Anna Bonnet , Sylvain Le Corff

The Hill coefficient is often used as a direct measure of the cooperativity of binding processes. It is an essential tool for probing properties of reactions in many biochemical systems. Here we analyze existing experimental data and…

生物物理 · 物理学 2012-08-31 M. Sheinman , Y. Kafri

Consider a random sample $X_1 , X_2 , ..., X_n$ drawn independently and identically distributed from some known sampling distribution $P_X$. Let $X_{(1)} \le X_{(2)} \le ... \le X_{(n)}$ represent the order statistics of the sample. The…

信息论 · 计算机科学 2020-09-28 Alex Dytso , Martina Cardone , Cynthia Rush

Econometricians have usefully separated study of estimation into identification and statistical components. Identification analysis, which assumes knowledge of the probability distribution generating observable data, places an upper bound…

计量经济学 · 经济学 2025-09-03 Charles F. Manski

Profiling is a process that finds similarities between different RNA secondary structures by extracting signals from the Boltzmann sampling. The reproducibility of profiling can be identified by the standard deviation of number of features…

生物大分子 · 定量生物学 2024-03-20 Qiuyun Li , Manda Riehl