English
Related papers

Related papers: Smiles2Dock: an open large-scale multi-task datase…

200 papers

In drug-discovery-related tasks such as virtual screening, machine learning is emerging as a promising way to predict molecular properties. Conventionally, molecular fingerprints (numerical representations of molecules) are calculated…

Machine Learning · Computer Science 2019-11-13 Shion Honda , Shoi Shi , Hiroki R. Ueda

Scientific databases aggregate vast amounts of quantitative data alongside descriptive text. In biochemistry, molecule screening assays evaluate candidate molecules' functional responses against disease targets. Unstructured text that…

Machine Learning · Computer Science 2025-11-18 Yifan Deng , Spencer S. Ericksen , Anthony Gitter

The accurate prediction of protein-ligand binding affinity is important for drug discovery yet remains challenging for multi-domain proteins, where inter-domain dynamics and flexible linkers govern molecular recognition. Current geometric…

Quantitative Methods · Quantitative Biology 2026-01-27 Shuo Zhang , Jian K. Liu

A fine-grained data recipe is crucial for pre-training large language models, as it can significantly enhance training efficiency and model performance. One important ingredient in the recipe is to select samples based on scores produced by…

Computation and Language · Computer Science 2026-01-01 Ziqing Fan , Yuqiao Xian , Yan Sun , Li Shen

The majority of machine learning scoring functions used in drug discovery for predicting protein-ligand binding poses and affinities have been trained on the PDBBind dataset. However, it is unclear whether these new scoring functions are…

Biological Physics · Physics 2026-01-13 Jie Li , Xingyi Guan , Oufan Zhang , Kunyang Sun , Yingze Wang , Dorian Bagni , Teresa Head-Gordon

Molecular docking is an essential step in the drug discovery process involving the detection of three-dimensional poses of a ligand inside the active site of the protein. In this paper, we address the Molecular Docking search phase by…

Understanding how protein mutations affect protein-nucleic acid binding is critical for unraveling disease mechanisms and advancing therapies. Current experimental approaches are laborious, and computational methods remain limited in…

Quantitative Methods · Quantitative Biology 2025-05-30 Xiang Liu , Junjie Wee , Guo-Wei Wei

Enzyme mining is rapidly evolving as a data-driven strategy to identify biocatalysts with tailored functions from the vast landscape of uncharacterized proteins. The integration of machine learning into these workflows enables…

Biomolecules · Quantitative Biology 2025-07-11 Yanzi Zhang , Felix Moorhoff , Sizhe Qiu , Wenjuan Dong , David Medina-Ortiz , Jing Zhao , Mehdi D. Davari

Small molecules are essential to drug discovery, and graph-language models hold promise for learning molecular properties and functions from text. However, existing molecule-text datasets are limited in scale and informativeness,…

Biomolecules · Quantitative Biology 2025-06-03 Yihan Zhu , Gang Liu , Eric Inae , Meng Jiang

The rapid evolution of artificial intelligence in drug discovery encounters challenges with generalization and extensive training, yet Large Language Models (LLMs) offer promise in reshaping interactions with complex molecular data. Our…

Biomolecules · Quantitative Biology 2024-12-20 He Cao , Zijing Liu , Xingyu Lu , Yuan Yao , Yu Li

Binding affinity optimization is crucial in early-stage drug discovery. While numerous machine learning methods exist for predicting ligand potency, their comparative efficacy remains unclear. This study evaluates the performance of…

Biomolecules · Quantitative Biology 2024-07-30 Nikolai Schapin , Carles Navarro , Albert Bou , Gianni De Fabritiis

Artificial intelligence (AI) is increasingly used in every stage of drug development. One challenge facing drug discovery AI is that drug pharmacokinetic (PK) datasets are often collected independently from each other, often with limited…

Quantitative Methods · Quantitative Biology 2025-07-03 Bing Hu , Anita Layton , Helen Chen

The performance of machine learning models in drug discovery is highly dependent on the quality and consistency of the underlying training data. Due to limitations in dataset sizes, many models are trained by aggregating bioactivity data…

Machine Learning · Computer Science 2025-11-21 Vincent Fan , Regina Barzilay

Light-activated drugs are a promising way to localize biological activity and minimize side effects. However, their development is complicated by the numerous photophysical and biological properties that must be simultaneously optimized. To…

Chemical Physics · Physics 2023-02-23 Simon Axelrod , Eugene Shakhnovich , Rafael Gómez-Bombarelli

Models based on machine learning can enable accurate and fast molecular property predictions, which is of interest in drug discovery and material design. Various supervised machine learning models have demonstrated promising performance,…

Machine Learning · Computer Science 2022-12-15 Jerret Ross , Brian Belgodere , Vijil Chenthamarakshan , Inkit Padhi , Youssef Mroueh , Payel Das

Prediction of protein-ligand interactions (PLI) plays a crucial role in drug discovery as it guides the identification and optimization of molecules that effectively bind to target proteins. Despite remarkable advances in deep…

Biomolecules · Quantitative Biology 2023-07-18 Seokhyun Moon , Sang-Yeon Hwang , Jaechang Lim , Woo Youn Kim

Virtual screening of large compound libraries to identify potential hit candidates is one of the earliest steps in drug discovery. As the size of commercially available compound collections grows exponentially to the scale of billions,…

Machine Learning · Computer Science 2023-09-22 Zhonglin Cao , Simone Sciabola , Ye Wang

Protein-peptide interactions play a key role in cell functions. Their structural characterization, though challenging, is important for the discovery of new drugs. The CABS-dock web server provides an interface for modeling protein-peptide…

Biomolecules · Quantitative Biology 2015-07-08 Mateusz Kurcinski , Michal Jamroz , Maciej Blaszczyk , Andrzej Kolinski , Sebastian Kmiecik

With the advancing capabilities of computational methodologies and resources, ultra-large-scale virtual screening via molecular docking has emerged as a prominent strategy for in silico hit discovery. Given the exhaustive nature of…

Machine Learning · Computer Science 2024-06-21 Jeonghyeon Kim , Juno Nam , Seongok Ryu

Protein-peptide molecular docking is a difficult modeling problem. It is even more challenging when significant conformational changes that may occur during the binding process need to be predicted. In this chapter, we demonstrate the…

Biomolecules · Quantitative Biology 2017-01-03 Maciej Pawel Ciemny , Mateusz Kurcinski , Konrad Jakub Kozak , Andrzej Kolinski , Sebastian Kmiecik
‹ Prev 1 4 5 6 7 8 10 Next ›