English
Related papers

Related papers: SUMO Substrates and Sites Prediction Combining Pat…

200 papers

Biological screens are plagued by false positive hits resulting from aggregation. Thus, methods to triage small colloidally aggregating molecules (SCAMs) are in high demand. Herein, we disclose a bespoke machine-learning tool to confidently…

Quantitative Methods · Quantitative Biology 2021-05-04 Kuan Lee , Ann Yang , Yen-Chu Lin , Daniel Reker , Goncalo J. L. Bernardes , Tiago Rodrigues

Two-dimensional materials have attracted considerable attention due to their remarkable electronic, mechanical and optical properties, making them prime candidates for next-generation electronic and optoelectronic applications. Despite…

Echoing recent calls to counter reliability and robustness concerns in machine learning via multiverse analysis, we present PRESTO, a principled framework for mapping the multiverse of machine-learning models that rely on latent…

Machine Learning · Computer Science 2024-06-04 Jeremy Wayland , Corinna Coupette , Bastian Rieck

Understanding how protein mutations affect protein-nucleic acid binding is critical for unraveling disease mechanisms and advancing therapies. Current experimental approaches are laborious, and computational methods remain limited in…

Quantitative Methods · Quantitative Biology 2025-05-30 Xiang Liu , Junjie Wee , Guo-Wei Wei

The preservation of soil health is a critical challenge in the 21st century due to its significant impact on agriculture, human health, and biodiversity. We provide the first deep investigation of the predictive potential of machine…

Machine Learning · Statistics 2024-02-20 Rosa Aghdam , Xudong Tang , Shan Shan , Richard Lankau , Claudia Solís-Lemus

Intracellular compartmentalization of proteins underpins their function and the metabolic processes they sustain. Various mass spectrometry-based proteomics methods (subcellular spatial proteomics) now allow high throughput subcellular…

Quantitative Methods · Quantitative Biology 2025-12-10 Ziyue Zheng , Loay J. Jabre , Matthew McIlvin , Mak A. Saito , Sangwon Hyun

Protein language models (PLMs) have enabled advances in structure prediction and de novo protein design, yet they frequently collapse into pathological repetition during generation. Unlike in text, where repetition merely reduces…

Biomolecules · Quantitative Biology 2026-02-03 Jiahao Zhang , Zeqing Zhang , Di Wang , Lijie Hu

We show that macro-molecular self-assembly can recognize and classify high-dimensional patterns in the concentrations of $N$ distinct molecular species. Similar to associative neural networks, the recognition here leverages dynamical…

Disordered Systems and Neural Networks · Physics 2017-04-26 Weishun Zhong , David J. Schwab , Arvind Murugan

Studying the conformations involved in the dimerization of cadherins is highly relevant to understand the development of tissue and its failure, which is associated with tumors and metastases. Experimental techniques, like X-ray…

Biomolecules · Quantitative Biology 2020-02-26 S. Terzoli , G. Tiana

To perform recognition, molecules must locate and specifically bind their targets within a noisy biochemical environment with many look-alikes. Molecular recognition processes, especially the induced-fit mechanism, are known to involve…

Biomolecules · Quantitative Biology 2010-07-27 Yonatan Savir , Tsvi Tlusty

In this paper, we propose a data-driven method to learn interpretable topological features of biomolecular data and demonstrate the efficacy of parsimonious models trained on topological features in predicting the stability of synthetic…

Machine Learning · Statistics 2024-08-12 Amish Mishra , Francis Motta

The design of hybrid peptide-solid interfaces for nanotechnological applications such as biomolecular nanoarrays requires a deep understanding of the basic mechanisms of peptide binding and assembly at solid substrates. Here we show by…

Mesoscale and Nanoscale Physics · Physics 2011-07-07 Michael Bachmann , Karsten Goede , Annette G. Beck-Sickinger , Marius Grundmann , Anders Irbäck , Wolfhard Janke

Generation of drug-like molecules with high binding affinity to target proteins remains a difficult and resource-intensive task in drug discovery. Existing approaches primarily employ reinforcement learning, Markov sampling, or deep…

Machine Learning · Computer Science 2022-06-22 Peter Eckmann , Kunyang Sun , Bo Zhao , Mudong Feng , Michael K. Gilson , Rose Yu

Structure-based molecular ML (SBML) models can be highly sensitive to input geometries and give predictions with large variance. We present an approach to mitigate the challenge of selecting conformations for such models by generating…

Machine Learning · Computer Science 2023-11-08 Michael Maser , Natasa Tagasovska , Jae Hyeon Lee , Andrew Watkins

Genome wide comparisons between enteric bacteria yield large sets of conserved putative regulatory sites on a gene by gene basis that need to be clustered into regulons. Using the assumption that regulatory sites can be represented as…

Biological Physics · Physics 2009-11-07 Erik van Nimwegen , Mihaela Zavolan , Nikolaus Rajewsky , Eric D. Siggia

Predicting compound-protein affinity is critical for accelerating drug discovery. Recent progress made by machine learning focuses on accuracy but leaves much to be desired for interpretability. Through molecular contacts underlying…

Biomolecules · Quantitative Biology 2020-01-01 Mostafa Karimi , Di Wu , Zhangyang Wang , Yang Shen

We introduce ProtoPathway, an interpretable-by-design multimodal framework for cancer survival prediction that unifies whole slide imaging and transcriptomics through encoders producing biologically grounded representations on both sides of…

Computer Vision and Pattern Recognition · Computer Science 2026-05-21 Amaya Gallagher-Syed , Costantino Pitzalis , Myles J. Lewis , Michael R. Barnes , Gregory Slabaugh

Understanding protein sequences is vital and urgent for biology, healthcare, and medicine. Labeling approaches are expensive yet time-consuming, while the amount of unlabeled data is increasing quite faster than that of the labeled data due…

Computation and Language · Computer Science 2021-11-01 Liang He , Shizhuo Zhang , Lijun Wu , Huanhuan Xia , Fusong Ju , He Zhang , Siyuan Liu , Yingce Xia , Jianwei Zhu , Pan Deng , Bin Shao , Tao Qin , Tie-Yan Liu

Protein sequences are abundant in repeating segments, both as exact copies and as approximate segments with mutations. These repeats are important for protein structure and function, motivating decades of algorithmic work on repeat…

Machine Learning · Computer Science 2026-05-26 Gal Pomerants , Yaniv Nikankin , Anja Reusch , Tomer Tsaban , Ora Schueler-Furman , Yonatan Belinkov

How do mammalian cells that share the same genome exist in notably distinct phenotypes, exhibiting differences in morphology, gene expression patterns, and epigenetic chromatin statuses? Furthermore how do cells of different phenotypes…

Molecular Networks · Quantitative Biology 2015-04-27 Jianhua Xing , Jin Yu , Hang Zhang , Xiao-Jun Tian