中文
相关论文

相关论文: FIT: Tag based method for fusion proteins identifi…

200 篇论文

Genomic signal processing has been used successfully in bioinformatics to analyze biomolecular sequences and gain varied insights into DNA structure, gene organization, protein binding, sequence evolution, etc. But challenges remain in…

基因组学 · 定量生物学 2022-11-04 Saish Jaiswal , Shreya Nema , Hema A Murthy , Manikandan Narayanan

User identity linkage across social networks is an essential problem for cross-network data mining. Since network structure, profile and content information describe different aspects of users, it is critical to learn effective user…

社会与信息网络 · 计算机科学 2020-03-17 Siyuan Chen , Jiahai Wang , Xin Du , Yanqing Hu

Accurate prediction of drug-target binding affinity can accelerate drug discovery by prioritizing promising compounds before costly wet-lab screening. While deep learning has advanced this task, most models fuse ligand and protein…

机器学习 · 计算机科学 2025-09-26 Mohammadsaleh Refahi , Bahrad A. Sokhansanj , James R. Brown , Gail Rosen

In this study, we present a method of pattern mining based on network theory that enables the identification of protein structures or complexes from synthetic volume densities, without the knowledge of predefined templates or human biases…

定量方法 · 定量生物学 2022-10-18 August George , Doo Nam Kim , Trevor Moser , Ian T. Gildea , James E. Evans , Margaret S. Cheung

We introduce a protein language model for determining the complete sequence of a peptide based on measurement of a limited set of amino acids. To date, protein sequencing relies on mass spectrometry, with some novel edman degregation based…

Time series data are ubiquitous in real-world applications. However, one of the most common problems is that the time series data could have missing values by the inherent nature of the data collection process. So imputing missing values…

机器学习 · 计算机科学 2022-09-23 Eunkyu Oh , Taehun Kim , Yunhu Ji , Sushil Khyalia

Modern genomic analyses increasingly rely on pangenomes, that is, representations of the genome of entire populations. The simplest representation of a pangenome is a set of individual genome sequences. Compared to e.g. sequence graphs,…

基因组学 · 定量生物学 2025-11-25 Jannik Olbrich , Enno Ohlebusch

Traditional drug discovery relies on rounds of screening millions of candidate molecules with low success rates, making drug discovery time and resource intensive. To overcome this screening bottleneck, we introduce Latent-X, an all-atom…

This paper describes a method to efficiently retrieve protein database sequences similar to a query sequence, while allowing for significant numbers of mutations. We call this method SEQR for SEQuence Retrieval. This approach increases the…

基因组学 · 定量生物学 2018-11-05 David I. Hurwitz , Lianyi Han , Lewis Y. Geer

Due to the lack of a method to efficiently represent the multimodal information of a protein, including its structure and sequence information, predicting compound-protein binding affinity (CPA) still suffers from low accuracy when applying…

生物大分子 · 定量生物学 2022-11-28 Binjie Guo , Hanyu Zheng , Haohan Jiang , Xiaodan Li , Naiyu Guan , Yanming Zuo , Yicheng Zhang , Hengfu Yang , Xuhua Wang

Autonomous driving demands accurate perception and safe decision-making. To achieve this, automated vehicles are now equipped with multiple sensors (e.g., camera, Lidar, etc.), enabling them to exploit complementary environmental context by…

计算机视觉与模式识别 · 计算机科学 2022-02-24 Xiaoming Zeng , Zhendong Wang , Yang Hu

Structured data is widely used in domains such as healthcare, finance, and scientific data management. Recent studies on structured data foundation models (SFMs) aim to support data analysis and mining tasks over such data, but still face…

机器学习 · 计算机科学 2026-05-21 Zhenghang Song , Tang Qian , Lu Chen , Yushuai Li , Zhengke Hu , Bingbing Fang , Yumeng Song , Junbo Zhao , Sheng Zhang , Tianyi Li

Pedestrian detection plays a critical role in computer vision as it contributes to ensuring traffic safety. Existing methods that rely solely on RGB images suffer from performance degradation under low-light conditions due to the lack of…

计算机视觉与模式识别 · 计算机科学 2024-08-28 Xue Zhang , Xiaohan Zhang , Jiangtao Wang , Jiacheng Ying , Zehua Sheng , Heng Yu , Chunguang Li , Hui-Liang Shen

Problems of search and recognition appear over different scales in biological systems. In this review we focus on the challenges posed by interactions between proteins, in particular transcription factors, and DNA and possible mechanisms…

生物大分子 · 定量生物学 2015-03-19 M. Sheinman , O. Bénichou , Y. Kafri , R. Voituriez

The Gene Ontology (GO) provides a knowledge base to effectively describe proteins. However, measuring similarity between proteins based on GO remains a challenge. In this paper, we propose a new similarity measure, information coefficient…

计算工程、金融与科学 · 计算机科学 2010-01-07 Bo Li , James Z. Wang , F. Alex Feltus , Jizhong Zhou , Feng Luo

Accurate identification of protein-nucleotide binding sites is fundamental to deciphering molecular mechanisms and accelerating drug discovery. However, current computational methods often struggle with suboptimal performance due to…

机器学习 · 计算机科学 2026-03-17 Yiming Gao , Liuyi Xu , Pengshan Cui , Yining Qian , An-Yang Lu , Xianpeng Wang

Tandem mass spectrometry has played a pivotal role in advancing proteomics, enabling the high-throughput analysis of protein composition in biological tissues. Many deep learning methods have been developed for \emph{de novo} peptide…

定量方法 · 定量生物学 2024-11-01 Jingbo Zhou , Shaorong Chen , Jun Xia , Sizhe Liu , Tianze Ling , Wenjie Du , Yue Liu , Jianwei Yin , Stan Z. Li

Pathology foundation models (PFMs) have demonstrated strong representational capabilities through self-supervised pre-training on large-scale, unannotated histopathology image datasets. However, their diverse yet opaque pretraining…

计算机视觉与模式识别 · 计算机科学 2025-09-15 Yuxiang Xiao , Yang Hu , Bin Li , Tianyang Zhang , Zexi Li , Huazhu Fu , Jens Rittscher , Kaixiang Yang

We present ensemble methods in a machine learning (ML) framework combining predictions from five known motif/binding site exploration algorithms. For a given TF the ensemble starts with position weight matrices (PWM's) for the motif,…

基因组学 · 定量生物学 2018-05-11 Yue Fan , Mark Kon , Charles DeLisi

Drug discovery represents a time-consuming and financially intensive process, and virtual screening can accelerate it. Scoring functions, as one of the tools guiding virtual screening, have their precision closely tied to screening…

机器学习 · 计算机科学 2026-01-13 Haotian Gao , Xiangying Zhang , Jingyuan Li , Xinchong Chen , Haojie Wang , Yifei Qi , Renxiao Wang