English
Related papers

Related papers: Mutation model for nucleotide sequences based on c…

200 papers

Sequencing by Emergence (SEQE) is a new single-molecule nucleic acid (DNA/RNA) sequencing technology that estimates sequence as an emergent property of the binding and localization of a repertoire of short oligonucleotide probes. SEQE…

Genomics · Quantitative Biology 2021-08-04 Nicholas Boyd , Samuel Woodhouse , Kalim Mir

The distribution of bases spacing in human genome was investigated. An analysis of the frequency of occurrence in the human genome of different sequence lengths flanked by one type of nucleotide was carried out showing that the distribution…

Other Quantitative Biology · Quantitative Biology 2019-04-18 Andrzej Z. Górski , Monika Piwowar

It is a well-known fact that genetic sequences may contain sections with repeated units, called repeats, that differ in length over a population, with a length distribution of geometric type. A simple class of recombination models with…

Dynamical Systems · Mathematics 2010-02-09 Michael Baake

A representation of the genetic code as a six-dimensional Boolean hypercube is proposed. It is assumed here that this structure is the result of the hierarchical order of the interaction energies of the bases in codon-anticodon recognition.…

Soft Condensed Matter · Physics 2007-05-23 Miguel A. Jimenez-Montano , Carlos R. de la Mora-Basanez , Thorsten Poeschel

The stability of model proteins with designed sequences is assessed in terms of the number of sequences (obtained from the designed sequence through mutations), which fold into 5the ``native'' conformation. By a complete enumeration of the…

Soft Condensed Matter · Physics 2009-10-31 R. A Broglia , G. Tiana , H. E. Roman , E. Vigezzi , E. I. Shakhnovich

We define the complexity of DNA sequences as the information content per nucleotide, calculated by means of some Lempel-Ziv data compression algorithm. It is possible to use the statistics of the complexity values of the functional regions…

Quantitative Methods · Quantitative Biology 2008-03-05 Giulia Menconi , Vieri Benci , Marcello Buiatti

This paper presents a novel method to segment/decode DNA sequences based on n-grams statistical language model. Firstly, we find the length of most DNA 'words' is 12 to 15 bps by analyzing the genomes of 12 model species. Then we design an…

Genomics · Quantitative Biology 2015-03-13 Wang Liang

We study the statistics of quantum transmission through a one-dimensional disordered system modelled by a sequence of independent scattering units. Each unit is characterized by its length and by its action, which is proportional to the…

Statistical Mechanics · Physics 2007-11-06 D. Boose , J. M. Luck

We study a model of a population with individuals sampled from different species. The Yule-$\Lambda$ nested coalescent describes the genealogy of the sample when each species merges with another randomly chosen species with a constant rate…

Probability · Mathematics 2024-01-05 Toni Gui

A general theoretical framework is put forth to organize and understand various observed phenomena and mathematical relationships in the field of molecular biology. By modeling each cell in eukaryotic organisms as a processor having a…

Other Quantitative Biology · Quantitative Biology 2013-12-18 Barry D. Jacobson

We introduce a family of models incorporating random segmental substitutions and point mutations and demonstrate that such models reproduce algebraic length distributions of exact matches with the slope $-4$ observed earlier in pairwise…

Genomics · Quantitative Biology 2015-07-15 M. V. Koroteev , P. V. Baranov

We explore the large-scale behavior of nucleotide compositional strand asymmetries along human chromosomes. As we observe for 7 of 9 origins of replication experimentally identified so far, the (TA+GC) skew displays rather sharp upward…

A new version of DNA walks, where nucleotides are regarded unequal in their contribution to a walk is introduced, which allows us to study thoroughly the "fine structure" of nucleotide sequences. The approach is based on the assumption that…

Genomics · Quantitative Biology 2007-05-23 Diana Duplij , Steven Duplij

Quantum effects are mainly used for the determination of molecular shapes in molecular biology, but quantum information theory may be a more useful tool to understand the physics of life. Organic molecules and quantum circuits/protocols can…

Quantum Physics · Physics 2011-04-01 Onur Pusuluk , Cemsinan Deliduman

We study permutations over the set of $\ell$-grams, that are feasible in the sense that there is a sequence whose $\ell$-gram frequency has the same ranking as the permutation. Codes, which are sets of feasible permutations, protect…

Information Theory · Computer Science 2021-01-18 Niv Beeri , Moshe Schwartz

The interplay between bending of the molecule axis and appearance of disruptions in circular DNA molecules, with $\sim 100$ base pairs, is addressed. Three minicircles with different radii and almost equal content of AT and GC pairs are…

Soft Condensed Matter · Physics 2014-06-03 Marco Zoli

Several studies suggest strong correlation between different types of cancer and the relative concentration of short circulating RNA sequences (miRNA). Because of short length and low concentration, miRNA detection is not easy. Standard…

Mesoscale and Nanoscale Physics · Physics 2021-10-25 Luyan Yang , Christophe Cullin , Juan Elezgaray

It is shown that metric representation of DNA sequences is one-to-one. By using the metric representation method, suppression of nucleotide strings in the DNA sequences is determined. For a DNA sequence, an optimal string length to display…

Biological Physics · Physics 2007-05-23 Zuo-Bing Wu

The site frequency spectrum describes variation among a set of n DNA sequences. Its i'th entry (i=1,2,...,n-1) is the number of nucleotide sites at which the mutant allele is present in i copies. Under selective neutrality, random mating,…

Populations and Evolution · Quantitative Biology 2021-03-02 Alan R. Rogers , Stephen P. Wooding

We develop statistically based methods to detect single nucleotide DNA mutations in next generation sequencing data. Sequencing generates counts of the number of times each base was observed at hundreds of thousands to billions of genome…

Applications · Statistics 2012-10-01 Omkar Muralidharan , Georges Natsoulis , John Bell , Hanlee Ji , Nancy R. Zhang