Related papers: Models and information-theoretic bounds for nanopo…
As a possible implementation of data storage using DNA, multiple strands of DNA are stored in a liquid container so that, in the future, they can be read by an array of DNA readers in parallel. These readers will sample the strands with…
Sequencing by Emergence (SEQE) is a new single-molecule nucleic acid (DNA/RNA) sequencing technology that estimates sequence as an emergent property of the binding and localization of a repertoire of short oligonucleotide probes. SEQE…
In this work, we analyze the capabilities and practical limitations of neural networks (NNs) for sequence-based signal processing which can be seen as an omnipresent property in almost any modern communication systems. In particular, we…
Recently solid state nanopores nanogaps have generated a lot of interest in ultrafast DNA sequencing. However, there are challenges to slow down the DNA translocation process to achieve a single nucleobase resolution. A series of…
Since the birth of computer and networks, fuelled by pervasive computing and ubiquitous connectivity, the amount of data stored and transmitted has exponentially grown through the years. Due to this demand, new solutions for storing data…
DNA is emerging as an increasingly attractive medium for data storage due to a number of important and unique advantages it offers, most notably the unprecedented durability and density. While the technology is evolving rapidly, the…
Detecting chemical modifications on RNA molecules remains a key challenge in epitranscriptomics. Traditional reverse transcription-based sequencing methods introduce enzyme- and sequence-dependent biases and fragment RNA molecules,…
The threading of a polymer chain through a small pore is a classic problem in polymer dynamics and underlies nanopore sensing technology. However important experimental aspects of the polymer motion in a solid-state nanopore, such as an…
We suggest to discriminate single DNA bases via transverse ionic transport, namely by detecting the ionic current that flows in a channel while a single-stranded DNA is driven through an intersecting nanochannel. Our all-atom molecular…
Nanopore sequencing accuracy is inherently limited by the quality of data from individual molecular translocation events, requiring advances beyond traditional sequencing-by-synthesis methods. We introduce an oxidized pyramidal sub-nm pore…
Double-stranded DNA translocates through sufficiently large nanopores either in a linear, single-file fashion or in a folded hairpin conformation when captured somewhere along its length. We show that the folding state of DNA can be…
DNA-based data storage has been attracting significant attention due to its extremely high data storage density, low power consumption, and long duration compared to conventional data storage media. Despite the recent advancements in DNA…
In this short note, a correction is made to the recently proposed solution [1] to a 1D biased diffusion model for linear DNA translocation and a new analysis will be given to the data in [1]. It was pointed out [2] by us recently that this…
Synthesis of DNA molecules offers unprecedented advances in storage technology. Yet, the microscopic world in which these molecules reside induces error patterns that are fundamentally different from their digital counterparts. Hence, to…
DNA labeling is a powerful tool in molecular biology and biotechnology that allows for the visualization, detection, and study of DNA at the molecular level. Under this paradigm, a DNA molecule is being labeled by specific k patterns and is…
Eukaryotic cell development has been optimized by natural selection to obey maximal intracellular flux of messenger proteins. This, in turn, implies maximum Fisher information on angular position about a target nuclear pore complex (NPR).…
We present a simplified model of the dynamics of translocation of RNA through a nanopore which only allows the passage of unbound nucleotides. In particular, we consider the disorder averaged translocation dynamics of random, two-component,…
Short-read DNA sequencing instruments can yield over 1e+12 bases per run, typically composed of reads 150 bases long. Despite this high throughput, de novo assembly algorithms have difficulty reconstructing contiguous genome sequences using…
DNA sequences encode critical genetic information, yet their variable length and discrete nature impede direct utilization in deep learning models. Existing DNA representation schemes convert sequences into numerical vectors but fail to…
Nanopore resistive pulse techniques are based on analysis of current or voltage spikes in the recorded signal. These spikes result from translocation of nanometer sized analytes through a nanopore. The most important information that needs…