English
Related papers

Related papers: DOCKSTRING: easy molecular docking yields better b…

200 papers

With the development of computer-assisted techniques, research communities including biochemistry and deep learning have been devoted into the drug discovery field for over a decade. Various applications of deep learning have drawn great…

Machine Learning · Computer Science 2023-03-07 Wenhao Hu , Yingying Liu , Xuanyu Chen , Wenhao Chai , Hangyue Chen , Hongwei Wang , Gaoang Wang

The rapid evolution of artificial intelligence in drug discovery encounters challenges with generalization and extensive training, yet Large Language Models (LLMs) offer promise in reshaping interactions with complex molecular data. Our…

Biomolecules · Quantitative Biology 2024-12-20 He Cao , Zijing Liu , Xingyu Lu , Yuan Yao , Yu Li

Prediction of protein-ligand interactions (PLI) plays a crucial role in drug discovery as it guides the identification and optimization of molecules that effectively bind to target proteins. Despite remarkable advances in deep…

Biomolecules · Quantitative Biology 2023-07-18 Seokhyun Moon , Sang-Yeon Hwang , Jaechang Lim , Woo Youn Kim

Structure-based drug design involves finding ligand molecules that exhibit structural and chemical complementarity to protein pockets. Deep generative methods have shown promise in proposing novel molecules from scratch (de-novo design),…

Quantitative Methods · Quantitative Biology 2021-11-09 Pavol Drotár , Arian Rokkum Jamasb , Ben Day , Cătălina Cangea , Pietro Liò

Is it feasible to create an analysis paradigm that can analyze and then accurately and quickly predict known drugs from experimental data? PharML.Bind is a machine learning toolkit which is able to accomplish this feat. Utilizing deep…

Biomolecules · Quantitative Biology 2019-11-15 Aaron D. Vose , Jacob Balma , Damon Farnsworth , Kaylie Anderson , Yuri K. Peterson

For large libraries of small molecules, exhaustive combinatorial chemical screens become infeasible to perform when considering a range of disease models, assay conditions, and dose ranges. Deep learning models have achieved state of the…

Protein-protein docking is crucial for understanding how proteins interact. Numerous docking tools have been developed to discover possible conformations of two interacting proteins. However, the reliability and success of these docking…

Biomolecules · Quantitative Biology 2025-11-18 Azam Shirali , Vitalii Stebliankin , Jimeng Shi , Prem Chapagain , Giri Narasimhan

Significant interests have recently risen in leveraging sequence-based large language models (LLMs) for drug design. However, most current applications of LLMs in drug discovery lack the ability to comprehend three-dimensional (3D)…

Prediction of ligand binding sites of proteins is a fundamental and important task for understanding the function of proteins and screening potential drugs. Most existing methods require experimentally determined protein holo-structures as…

Quantitative Methods · Quantitative Biology 2023-12-07 Shuo Zhang , Lei Xie

Recent advancements in biology and chemistry have leveraged multi-modal learning, integrating molecules and their natural language descriptions to enhance drug discovery. However, current pre-training frameworks are limited to two…

Machine Learning · Computer Science 2025-02-05 Teng Xiao , Chao Cui , Huaisheng Zhu , Vasant G. Honavar

In this work, we propose a deep learning approach to improve docking-based virtual screening. The introduced deep neural network, DeepVS, uses the output of a docking program and learns how to extract relevant features from basic data such…

Quantitative Methods · Quantitative Biology 2016-11-22 Janaina Cruz Pereira , Ernesto Raul Caffarena , Cicero dos Santos

Drug development is an expensive and time-consuming process where thousands of chemical compounds are being tested in order to find those possessing drug-like properties while being safe and effective. One of key parts of the early drug…

Quantitative Methods · Quantitative Biology 2022-02-15 Josip Mesarić

Molecular structures are always depicted as 2D printed form in scientific documents like journal papers and patents. However, these 2D depictions are not machine-readable. Due to a backlog of decades and an increasing amount of these…

Computer Vision and Pattern Recognition · Computer Science 2022-05-24 Youjun Xu , Jinchuan Xiao , Chia-Han Chou , Jianhang Zhang , Jintao Zhu , Qiwan Hu , Hemin Li , Ningsheng Han , Bingyu Liu , Shuaipeng Zhang , Jinyu Han , Zhen Zhang , Shuhao Zhang , Weilin Zhang , Luhua Lai , Jianfeng Pei

Simulations of biological macromolecules play an important role in understanding the physical basis of a number of complex processes such as protein folding. Even with increasing computational power and evolution of specialized…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-09-18 Hyungro Lee , Heng Ma , Matteo Turilli , Debsindhu Bhowmik , Shantenu Jha , Arvind Ramanathan

Design of new drugs is a challenging process: a candidate molecule should satisfy multiple conditions to act properly and make the least side-effect -- perfect candidates selectively attach to and influence only targets, leaving off-targets…

Biomolecules · Quantitative Biology 2024-05-07 Andrij Rovenchak , Maksym Druchok

Generating molecules with high binding affinities to target proteins (a.k.a. structure-based drug design) is a fundamental and challenging task in drug discovery. Recently, deep generative models have achieved remarkable success in…

Biomolecules · Quantitative Biology 2023-05-24 Zaixi Zhang , Qi Liu

Fragment-Based Drug Discovery (FBDD) is a popular approach in early drug development, but designing effective linkers to combine disconnected molecular fragments into chemically and pharmacologically viable candidates remains challenging.…

Machine Learning · Computer Science 2025-09-24 Xuefeng Liu , Songhao Jiang , Qinan Huang , Tinson Xu , Ian Foster , Mengdi Wang , Hening Lin , Rick Stevens

Pre-training datasets are typically collected from web content and lack inherent domain divisions. For instance, widely used datasets like Common Crawl do not include explicit domain labels, while manually curating labeled datasets such as…

Computation and Language · Computer Science 2025-12-02 Shizhe Diao , Yu Yang , Yonggan Fu , Xin Dong , Dan Su , Markus Kliegl , Zijia Chen , Peter Belcak , Yoshi Suhara , Hongxu Yin , Mostofa Patwary , Yingyan , Lin , Jan Kautz , Pavlo Molchanov

Protein-ligand complex structures have been utilised to design benchmark machine learning methods that perform important tasks related to drug design such as receptor binding site detection, small molecule docking and binding affinity…

Biomolecules · Quantitative Biology 2021-08-26 Rishal Aggarwal , Akash Gupta , U Deva Priyakumar

Existing benchmarking methods are time consuming processes as they typically benchmark the entire Virtual Machine (VM) in order to generate accurate performance data, making them less suitable for real-time analytics. The research in this…

Distributed, Parallel, and Cluster Computing · Computer Science 2016-11-17 Blesson Varghese , Lawan Thamsuhang Subba , Long Thai , Adam Barker