English
Related papers

Related papers: Why high-error-rate random mutagenesis libraries a…

200 papers

Protein folding is the intricate process by which a linear sequence of amino acids self-assembles into a unique three-dimensional structure. Protein folding kinetics is the study of pathways and time-dependent mechanisms a protein undergoes…

Machine Learning · Computer Science 2023-09-19 Vijay Arvind. R , Haribharathi Sivakumar , Brindha. R

Adequate read filtering is critical when processing high-throughput data in marker-gene-based studies. Sequencing errors can cause the mis-clustering of otherwise similar reads, artificially increasing the number of retrieved Operational…

Quantitative Methods · Quantitative Biology 2015-06-02 Fernando Puente-Sánchez , Jacobo Aguirre , Víctor Parro

The most common representation in evolutionary computation are bit strings. This is ideal to model binary decision variables, but less useful for variables taking more values. With very little theoretical work existing on how to use…

Neural and Evolutionary Computing · Computer Science 2016-04-13 Benjamin Doerr , Carola Doerr , Timo Kötzing

Our ability to calculate rates of biochemical processes using molecular dynamics simulations is severely limited by the fact that the time scales for reactions, or changes in conformational state, scale exponentially with the relevant…

Chemical Physics · Physics 2024-03-19 Nicodemo Mazzaferro , Subarna Sasmal , Pilar Cossio , Glen M. Hocky

In this paper, we study error exponents for a concatataned coding based class of DNA storage codes in which the number of reads performed can be variable. That is, the decoder can sequentially perform reads and choose whether to output the…

Information Theory · Computer Science 2025-04-25 Yan Hao Ling , Nir Weinberger , Jonathan Scarlett

When organisms adapt to spatially heterogeneous environments, selection may drive divergence at multiple genes. If populations under divergent selection also exchange migrants, we expect genetic differentiation to be high at selected loci,…

Populations and Evolution · Quantitative Biology 2014-07-21 Simon Aeschbacher , Reinhard Buerger

Synthesis of biopolymers such as DNA, RNA, and proteins are biophysical processes aided by enzymes. Performance of these enzymes is usually characterized in terms of their average error rate and speed. However, because of thermal…

Subcellular Processes · Quantitative Biology 2019-07-24 Davide Chiuchiú , Yuhai Tu , Simone Pigolotti

An accurate estimation of the Protein Space size, in light of the factors that govern it, is a long-standing problem and of paramount importance in evolutionary biology, since it determines the nature of protein evolvability. A simple…

Populations and Evolution · Quantitative Biology 2020-08-27 Jorge A. Vila

We study the impact of mutations (changes in amino acid sequence) on the thermodynamics of simple protein-like heteropolymers consisting of N monomers, representing the amino acid sequence. The sequence is designed to fold into its native…

Condensed Matter · Physics 2009-10-30 G. Tiana , R. A. Broglia , H. E. Roman , E. Vigezzi , E. Shakhnovich

Much recent work has explored molecular and population-genetic constraints on the rate of protein sequence evolution. The best predictor of evolutionary rate is expression level, for reasons which have remained unexplained. Here, we…

Populations and Evolution · Quantitative Biology 2007-05-23 D. Allan Drummond , Jesse D. Bloom , Christoph Adami , Claus O. Wilke , Frances H. Arnold

A significant obstacle in the development of robust machine learning models is covariate shift, a form of distribution shift that occurs when the input distributions of the training and test sets differ while the conditional label…

Machine Learning · Statistics 2021-11-17 Nilesh Tripuraneni , Ben Adlam , Jeffrey Pennington

Genetic fitness optimization using small populations or small population updates across generations generally suffers from randomly diverging evolutions. We propose a notion of highly probable fitness optimization through feasible…

Neural and Evolutionary Computing · Computer Science 2007-05-23 Paul Vitanyi

Translation is a crucial step in gene expression. During translation, macromolecules called ribosomes "read" the mRNA strand in a sequential manner and produce a corresponding protein. Translation is known to consume most of the cell's…

Genomics · Quantitative Biology 2014-10-16 Yoram Zarai , Michael Margaliot

The emerging field of high-throughput compartmentalized in vitro evolution is a promising new approach to protein engineering. In these experiments, libraries of mutant genotypes are randomly distributed and expressed in microscopic…

Populations and Evolution · Quantitative Biology 2019-11-06 Anton S. Zadorin , Yannick Rondelez

The interplay between mutation and selection plays a fundamental role in the behaviour of evolutionary algorithms (EAs). However, this interplay is still not completely understood. This paper presents a rigorous runtime analysis of a…

Neural and Evolutionary Computing · Computer Science 2010-12-15 Per Kristian Lehre , Xin Yao

Virtual screening of large compound libraries to identify potential hit candidates is one of the earliest steps in drug discovery. As the size of commercially available compound collections grows exponentially to the scale of billions,…

Machine Learning · Computer Science 2023-09-22 Zhonglin Cao , Simone Sciabola , Ye Wang

Proteins are polymerized by cyclic machines called ribosome which use their messenger RNA (mRNA) track also as the corresponding template and the process is called translation. We explore, in depth and detail, the stochastic nature of the…

Biological Physics · Physics 2009-11-13 Ashok Garai , Debashish Chowdhury , T. V. Ramakrishnan

We investigated the error-minimization properties of putative primordial codes that consisted of 16 supercodons, with the third base being completely redundant, using a previously derived cost function and the error minimization percentage…

Genomics · Quantitative Biology 2009-08-26 Artem S. Novozhilov , Eugene V. Koonin

Supervised learning algorithms generally assume the availability of enough memory to store data models during the training and test phases. However, this assumption is unrealistic when data comes in the form of infinite data streams, or…

Machine Learning · Computer Science 2022-10-13 Martin Khannouz , Tristan Glatard

High-dimensional data is common in multiple areas, such as health care and genomics, where the number of features can be tens of thousands. In such scenarios, the large number of features often leads to inefficient learning. Constraint…

Machine Learning · Statistics 2023-06-13 Kartheek Bondugula , Santiago Mazuelas , Aritz Pérez