中文
相关论文

相关论文: Pattern recognition on random trees associated to …

200 篇论文

Pattern recognition constitutes a particularly important task underlying a great deal of scientific and technologica activities. At the same time, pattern recognition involves several challenges, including the choice of features to…

机器学习 · 计算机科学 2024-09-04 Alexandre Benatti , Luciano da F. Costa

Thermostability is an important prerequisite for enzymes employed for industrial applications. Several machine learning based models have thus been formulated for protein classification based on this particular trait. These models have…

定量方法 · 定量生物学 2021-03-08 Jithin S. Sunny , Lilly M. Saleena

Matrices are two-dimensional data structures allowing one to conceptually organize information. For example, adjacency matrices are useful to store the links of a network; correlation matrices are simple ways to arrange gene co-expression…

无序系统与神经网络 · 物理学 2022-09-29 Flaviano Morone

Function of proteins or a network of interacting proteins often involves communication between residues that are well separated in sequence. The classic example is the participation of distant residues in allosteric regulation.…

生物大分子 · 定量生物学 2007-05-23 Ruxandra I. Dima , D. Thirumalai

Research in bioinformatics is a complex phenomenon as it overlaps two knowledge domains, namely, biological and computer sciences. This paper has tried to introduce an efficient data mining approach for classifying proteins into some useful…

计算工程、金融与科学 · 计算机科学 2011-11-11 Muhammad Mahbubur Rahman , Arif Ul Alam , Abdullah-Al-Mamun , Tamnun E Mursalin

We consider clustering in group decision making where the opinions are given by pairwise comparison matrices. In particular, the k-medoids model is suggested to classify the matrices since it has a linear programming problem formulation…

最优化与控制 · 数学 2025-04-17 Kolos Csaba Ágoston , Sándor Bozóki , László Csató

In data containing heterogeneous subpopulations, classification performance benefits from incorporating the knowledge of cluster structure in the classifier. Previous methods for such combined clustering and classification either 1) are…

机器学习 · 计算机科学 2023-01-04 Shivin Srivastava , Siddharth Bhatia , Lingxiao Huang , Lim Jun Heng , Kenji Kawaguchi , Vaibhav Rajan

The sequence of amino acids in a protein is believed to determine its native state structure, which in turn is related to the functionality of the protein. In addition, information pertaining to evolutionary relationships is contained in…

定量方法 · 定量生物学 2008-06-17 Kyung Dae Ko , Yoojin Hong , Gue Su Chang , Gaurav Bhardwaj , Damian B. van Rossum , Randen L. Patterson

Given a point set S and an unknown metric d on S, we study the problem of efficiently partitioning S into k clusters while querying few distances between the points. In our model we assume that we have access to one versus all queries that…

数据结构与算法 · 计算机科学 2011-05-10 Konstantin Voevodski , Maria-Florina Balcan , Heiko Roglin , Shang-Hua Teng , Yu Xia

Statistical analysis of alignments of large numbers of protein sequences has revealed "sectors" of collectively coevolving amino acids in several protein families. Here, we show that selection acting on any functional property of a protein,…

生物大分子 · 定量生物学 2019-04-26 Shou-Wen Wang , Anne-Florence Bitbol , Ned S. Wingreen

Fixed effects models are very flexible because they do not make assumptions on the distribution of effects and can also be used if the heterogeneity component is correlated with explanatory variables. A disadvantage is the large number of…

统计方法学 · 统计学 2015-12-17 Moritz Berger , Gerhard Tutz

Alignment-based sequence similarity searches, while accurate for some type of sequences, can produce incorrect results when used on more divergent but functionally related sequences that have undergone the sequence rearrangements observed…

基因组学 · 定量生物学 2015-01-21 Ivan Borozan , Stuart Watt , Vincent Ferretti

Recent years have seen tremendous developments in the use of machine learning models to link amino acid sequence, structure and function of folded proteins. These methods are, however, rarely applicable to the wide range of proteins and…

生物大分子 · 定量生物学 2025-02-27 Sören von Bülow , Giulio Tesei , Kresten Lindorff-Larsen

Experimental determination of protein function is resource-consuming. As an alternative, computational prediction of protein function has received attention. In this context, protein structural classification (PSC) can help, by allowing for…

分子网络 · 定量生物学 2020-03-17 Khalique Newaz , Mahboobeh Ghalehnovi , Arash Rahnama , Panos J. Antsaklis , Tijana Milenkovic

We present a technique for clustering categorical data by generating many dissimilarity matrices and averaging over them. We begin by demonstrating our technique on low dimensional categorical data and comparing it to several other…

机器学习 · 统计学 2017-09-20 Saeid Amiri , Bertrand Clarke , Jennifer Clarke

This report investigates an unsupervised, feature-based image matching pipeline for the novel application of identifying individual k\=ak\=a. Applied with a similarity network for clustering, this addresses a weakness of current supervised…

计算机视觉与模式识别 · 计算机科学 2023-01-25 Fintan O'Sullivan , Kirita-Rose Escott , Rachael C. Shaw , Andrew Lensen

Protein structure prediction remains to be an open problem in bioinformatics. There are two main categories of methods for protein structure prediction: Free Modeling (FM) and Template Based Modeling (TBM). Protein threading, belonging to…

生物大分子 · 定量生物学 2015-09-14 Haicang Zhang , Mingfu Shao , Chao Wang , Jianwei Zhu , Wei-Mou Zheng , Dongbo Bu

The goal of protein representation learning is to extract knowledge from protein databases that can be applied to various protein-related downstream tasks. Although protein sequence, structure, and function are the three key modalities for…

生物大分子 · 定量生物学 2024-05-14 Eunji Ko , Seul Lee , Minseon Kim , Dongki Kim

Clustering is a popular form of unsupervised learning for geometric data. Unfortunately, many clustering algorithms lead to cluster assignments that are hard to explain, partially because they depend on all the features of the data in a…

机器学习 · 计算机科学 2020-09-23 Sanjoy Dasgupta , Nave Frost , Michal Moshkovitz , Cyrus Rashtchian

We investigate active learning by pairwise similarity over the leaves of trees originating from hierarchical clustering procedures. In the realizable setting, we provide a full characterization of the number of queries needed to achieve…

机器学习 · 计算机科学 2019-10-15 Fabio Vitale , Anand Rajagopalan , Claudio Gentile