中文
相关论文

相关论文: Application of Data mining in Protein sequence Cla…

200 篇论文

We introduce a new model of proteins, which extends and enhances the traditional graphical representation by associating a combinatorial object called a fatgraph to any protein based upon its intrinsic geometry. Fatgraphs can easily be…

生物大分子 · 定量生物学 2009-05-30 R. C. Penner , Michael Knudsen , Carsten Wiuf , Joergen Ellegaard Andersen

Protein structural classification (PSC) is a supervised problem of assigning proteins into pre-defined structural (e.g., CATH or SCOPe) classes based on the proteins' sequence or 3D structural features. We recently proposed PSC approaches…

分子网络 · 定量生物学 2021-05-18 Khalique Newaz , Jacob Piland , Patricia L. Clark , Scott J. Emrich , Jun Li , Tijana Milenkovic

DNA sequence alignment is important today as it is usually the first step in finding gene mutation, evolutionary similarities, protein structure, drug development and cancer treatment. Covid-19 is one recent example. There are many…

基因组学 · 定量生物学 2023-06-01 Suchindra , Preetam Nagaraj

In the present work, we review the fundamental methods which have been developed in the last few years for classifying into families and clans the distribution of amino acids in protein databases. This is done through functions of random…

生物大分子 · 定量生物学 2018-06-15 R. P. Mondaini , S. C. de Albuquerque Neto

Classification of seizure type is a key step in the clinical process for evaluating an individual who presents with seizures. It determines the course of clinical diagnosis and treatment, and its impact stretches beyond the clinical domain…

信号处理 · 电气工程与系统科学 2024-03-06 David Ahmedt-Aristizabal , Tharindu Fernando , Simon Denman , Lars Petersson , Matthew J. Aburn , Clinton Fookes

Protein inverse folding, the design of an amino acid sequence based on a target protein structure, is a fundamental problem of computational protein engineering. Existing methods either generate sequences without leveraging external…

定量方法 · 定量生物学 2026-03-10 Jin Han , Tianfan Fu , Wu-Jun Li

Tokenization is a promising path to multi-modal models capable of jointly understanding protein sequences, structure, and function. Existing protein structure tokenizers create tokens by pooling information from local neighborhoods, an…

机器学习 · 计算机科学 2026-02-09 Rohit Dilip , Ayush Varshney , David Van Valen

Cutting planes (cuts) are crucial for solving Mixed Integer Linear Programming (MILP) problems. Advanced MILP solvers typically rely on manually designed heuristic algorithms for cut selection, which require much expert experience and…

最优化与控制 · 数学 2024-12-11 Xuefeng Zhang , Liangyu Chen , Zhengfeng Yang , Zhenbing Zeng

We propose an optimized parameter set for protein secondary structure prediction using three layer feed forward back propagation neural network. The methodology uses four parameters viz. encoding scheme, window size, number of neurons in…

生物大分子 · 定量生物学 2018-02-02 Jyotshna Dongardivev , Siby Abraham

Understanding the structure and dynamics of biological networks is one of the important challenges in system biology. In addition, increasing amount of experimental data in biological networks necessitate the use of efficient methods to…

人工智能 · 计算机科学 2012-07-17 Mohammadreza Keyvanpour , Fereshteh Azizani

Fingerprint classification is one of the most common approaches to accelerate the identification in large databases of fingerprints. Fingerprints are grouped into disjoint classes, so that an input fingerprint is compared only with those…

计算机视觉与模式识别 · 计算机科学 2017-05-16 Daniel Peralta , Isaac Triguero , Salvador García , Yvan Saeys , Jose M. Benitez , Francisco Herrera

Optical Character Recognition software (OCR) are important tools for obtaining accessible texts. We propose the use of artificial neural networks (ANN) in order to develop pattern recognition algorithms capable of recognizing both normal…

神经与进化计算 · 计算机科学 2016-07-08 Giuseppe Airò Farulla , Tiziana Armano , Anna Capietto , Nadir Murru , Rosaria Rossini

Data discretization, also known as binning, is a frequently used technique in computer science, statistics, and their applications to biological data analysis. We present a new method for the discretization of real-valued data into a finite…

其他定量生物学 · 定量生物学 2007-05-23 Elena S. Dimitrova , John J. McGee , Reinhard C. Laubenbacher

Predicting ATP-Protein Binding sites in genes is of great significance in the field of Biology and Medicine. The majority of research in this field has been conducted through time- and resource-intensive 'wet experiments' in laboratories.…

生物大分子 · 定量生物学 2024-02-06 Shreyas V , Swati Agarwal

Prediction of protein secondary structure from the amino acid sequence is a classical bioinformatics problem. Common methods use feed forward neural networks or SVMs combined with a sliding window, as these models does not naturally handle…

定量方法 · 定量生物学 2015-01-06 Søren Kaae Sønderby , Ole Winther

Attention-based models trained on protein sequences have demonstrated incredible success at classification and generation tasks relevant for artificial intelligence-driven protein design. However, we lack a sufficient understanding of how…

机器学习 · 计算机科学 2022-06-29 Erik Nijkamp , Jeffrey Ruffolo , Eli N. Weinstein , Nikhil Naik , Ali Madani

Inferring the structural properties of a protein from its amino acid sequence is a challenging yet important problem in biology. Structures are not known for the vast majority of protein sequences, but structure is critical for…

机器学习 · 计算机科学 2019-10-17 Tristan Bepler , Bonnie Berger

Multiple technologies that measure expression levels of protein mixtures in the human body offer a potential for detection and understanding the disease. The recent increase of these technologies prompts researchers to evaluate the…

机器学习 · 计算机科学 2026-05-12 Michal Valko , Richard Pelikan , Miloš Hauskrecht

This paper presents a novel meta learning framework for feature selection (FS) based on fuzzy similarity. The proposed method aims to recommend the best FS method from four candidate FS methods for any given dataset. This is achieved by…

机器学习 · 计算机科学 2020-05-22 Zixiao Shen , Xin Chen , Jonathan M. Garibaldi

Motivation: Drug discovery demands rapid quantification of compound-protein interaction (CPI). However, there is a lack of methods that can predict compound-protein affinity from sequences alone with high applicability, accuracy, and…

生物大分子 · 定量生物学 2020-12-17 Mostafa Karimi , Di Wu , Zhangyang Wang , Yang Shen
‹ 上一页 1 8 9 10 下一页 ›