English
Related papers

Related papers: Mutually Uncorrelated Primers for DNA-Based Data S…

200 papers

Deep learning models, such as wide neural networks, can be conceptualized as nonlinear dynamical physical systems characterized by a multitude of interacting degrees of freedom. Such systems in the infinite limit, tend to exhibit simplified…

Machine Learning · Computer Science 2024-01-09 Ori Shem-Ur , Yaron Oz

In this paper, we propose the use of a deterministic sequence, known as unique word (UW), instead of the cyclic prefix (CP) in generalized frequency division multiplexing (GFDM) systems. The UW consists of known sequences that, if not null,…

Information Theory · Computer Science 2018-05-29 J. Dias , R. C. de Lamare

Zero mode waveguide technology of next generation sequencing demonstrated sequence-dependence of the enzymatic reaction, incorporating a base into the genomic DNA. We show that these experimental results indicate existence of a previously…

Other Quantitative Biology · Quantitative Biology 2012-06-12 Petr Pancoska , Zdenek Moravek , Uday Kiran Para , Jaroslav Nesetril

In this work, we investigate a challenging problem, which has been considered to be an important criterion in designing codewords for DNA computing purposes, namely secondary structure avoidance in single-stranded DNA molecules. In short,…

Information Theory · Computer Science 2023-02-28 Tuan Thanh Nguyen , Kui Cai , Han Mao Kiah , Duc Tu Dao , Kees A. Schouhamer Immink

Motivation: With the rapid expansion of large-scale biological datasets, DNA and protein sequence alignments have become essential for comparative genomics and proteomics. These alignments facilitate the exploration of sequence similarity…

Genomics · Quantitative Biology 2024-12-02 Michail Patsakis , Kimonas Provatas , Ioannis Mouratidis , Ilias Georgakopoulos-Soares

Minimal absent words (MAW) of a genomic sequence are subsequences that are absent themselves but the subwords of which are all present in the sequence. The characteristic distribution of genomic MAWs as a function of their length has been…

Genomics · Quantitative Biology 2016-05-04 Erik Aurell , Nicolas Innocenti , Hai-Jun-Zhou

Like other pseudorandom sequences, decimal sequences may be used in designing a Code Division Multiple Access (CDMA) system. They appear to be ideally suited for this since the cross-correlation of d-sequences taken over the LCM of their…

Information Theory · Computer Science 2008-04-22 Sandeep Chalasani

Quantum systems with variables in ${\mathbb Z}(d)$ are considered. The properties of lines in the ${\mathbb Z}(d)\times {\mathbb Z}(d)$ phase space of these systems, are studied. Weak mutually unbiased bases in these systems are defined as…

Quantum Physics · Physics 2012-03-06 M. Shalaby , A. Vourdas

The study of correlation structure in the primary sequences of DNA is reviewed. The issues reviewed include: symmetries among 16 base-base correlation functions, accurate estimation of correlation measures, the relationship between $1/f$…

adap-org · Physics 2009-11-10 Wentian Li

We introduce a model of DNA sequence evolution which can account for biases in mutation rates that depend on the identity of the neighboring bases. An analytic solution for this class of non-equilibrium models is developed by adopting…

Biological Physics · Physics 2007-05-23 Peter F. Arndt , Christopher B. Burge , Terence Hwa

The advent of "next-generation" DNA sequencing (NGS) technologies has meant that collections of hundreds of millions of DNA sequences are now commonplace in bioinformatics. Knowing the longest common prefix array (LCP) of such a collection…

Data Structures and Algorithms · Computer Science 2013-05-02 Markus J. Bauer , Anthony J. Cox , Giovanna Rosone , Marinella Sciortino

We introduce a sequence-dependent parametrization for a coarse-grained DNA model [T. E. Ouldridge, A. A. Louis, and J. P. K. Doye, J. Chem. Phys. 134, 085101 (2011)] originally designed to reproduce the properties of DNA molecules with…

In this paper, fundamental limits in sequencing of a set of closely related DNA molecules are addressed. This problem is called pooled-DNA sequencing which encompasses many interesting problems such as haplotype phasing, metageomics, and…

Information Theory · Computer Science 2016-04-20 Amir Najafi , Damoun Nashta-ali , Seyed Abolfazl Motahari , Mehrdad Khani , Babak H. Khalaj , Hamid R. Rabiee

Persistent homology is a leading tool in topological data analysis (TDA). Many problems in TDA can be solved via homological -- and indeed, linear -- algebra. However, matrices in this domain are typically large, with rows and columns…

Algebraic Topology · Mathematics 2021-08-23 Haibin Hang , Chad Giusti , Lori Ziegelmeier , Gregory Henselman-Petrusek

We consider error-correcting coding for deoxyribonucleic acid (DNA)-based storage using nanopore sequencing. We model the DNA storage channel as a sampling noise channel where the input data is chunked into $M$ short DNA strands, which are…

Information Theory · Computer Science 2024-06-21 Lorenz Welter , Roman Sokolovskii , Thomas Heinis , Antonia Wachter-Zeh , Eirik Rosnes , Alexandre Graell i Amat

Owing to its longevity and enormous information density, DNA, the molecule encoding biological information, has emerged as a promising archival storage medium. However, due to technological constraints, data can only be written onto many…

Emerging Technologies · Computer Science 2018-03-12 Reinhard Heckel , Gediminas Mikutis , Robert N. Grass

Deep neural networks (DNNs) have been employed for designing wireless systems in many aspects, say transceiver design, resource optimization, and information prediction. Existing works either use the fully-connected DNN or the DNNs with…

Machine Learning · Computer Science 2020-01-31 Jia Guo , Chenyang Yang

DNA sequences are prone to creating secondary structures by folding back on themselves by non-specific hybridization among its nucleotides. The formation of secondary structures makes the sequences chemically inactive towards synthesis and…

Information Theory · Computer Science 2022-11-30 Siddhartha Siddhiprada Bhoi , Paramapalli Udaya , Abhay Kumar Singh

We live in a period where bio-informatics is rapidly expanding, a significant quantity of genomic data has been produced as a result of the advancement of high-throughput genome sequencing technology, raising concerns about the costs…

Quantitative Methods · Quantitative Biology 2023-03-10 Mehedi Hasan Sarkar , Adnan Ferdous Ashrafi

When building Burrows-Wheeler Transforms (BWTs) of truly huge datasets, prefix-free parsing (PFP) can use an unreasonable amount of memory. In this paper we show how if a dataset can be broken down into small datasets that are not very…

Data Structures and Algorithms · Computer Science 2025-06-09 Diego Diaz-Dominguez , Travis Gagie , Veronica Guerrini , Ben Langmead , Zsuzsanna Liptak , Giovanni Manzini , Francesco Masillo , Vikram Shivakumar
‹ Prev 1 3 4 5 6 7 10 Next ›