English
Related papers

Related papers: Mass spectrometry based protein identification wit…

200 papers

De novo peptide sequencing from mass spectrometry data is an important method for protein identification. Recently, various deep learning approaches were applied for de novo peptide sequencing and DeepNovoV2 is one of the represetative…

Quantitative Methods · Quantitative Biology 2022-03-18 Cheng Ge , Yi Lu , Jia Qu , Liangxu Xie , Feng Wang , Hong Zhang , Ren Kong , Shan Chang

Proteins are the main workhorses of biological functions in a cell, a tissue, or an organism. Identification and quantification of proteins in a given sample, e.g. a cell type under normal/disease conditions, are fundamental tasks for the…

Computational Engineering, Finance, and Science · Computer Science 2017-10-10 Ngoc Hieu Tran , Zachariah Levine , Lei Xin , Baozhen Shan , Ming Li

Nuclear magnetic resonance (NMR) spectroscopy is one of the leading techniques for protein studies. The method features a number of properties, allowing to explain macromolecular interactions mechanistically and resolve structures with…

Quantitative Methods · Quantitative Biology 2018-08-03 Piotr Klukowski , Adam Gonczarek

Mass spectrometry-based metabolomic analysis depends upon the identification of spectral peaks by their mass and retention time. Statistical analysis that follows the identification currently relies on one main peak of each compound.…

Quantitative Methods · Quantitative Biology 2014-03-20 Tommi Suvitaival , Simon Rogers , Samuel Kaski

Compound-Protein Interaction (CPI) prediction aims to predict the pattern and strength of compound-protein interactions for rational drug discovery. Existing deep learning-based methods utilize only the single modality of protein sequences…

Biomolecules · Quantitative Biology 2024-02-14 Lirong Wu , Yufei Huang , Cheng Tan , Zhangyang Gao , Bozhen Hu , Haitao Lin , Zicheng Liu , Stan Z. Li

Evolution in its course found a variety of solutions to the same optimisation problem. The advent of high-throughput genomic sequencing has made available extensive data from which, in principle, one can infer the underlying structure on…

Quantitative Methods · Quantitative Biology 2016-04-12 Silvia Grigolon , Silvio Franz , Matteo Marsili

Multiple technologies that measure expression levels of protein mixtures in the human body offer a potential for detection and understanding the disease. The recent increase of these technologies prompts researchers to evaluate the…

Machine Learning · Computer Science 2026-05-12 Michal Valko , Richard Pelikan , Miloš Hauskrecht

Proteins are arguably the most important class of biomarkers for health diagnostic purposes. Label-free solid-state nanopore sensing is a versatile technique for sensing and analysing biomolecules such as proteins at single-molecule level.…

Data mining techniques have been used by researchers for analyzing protein sequences. In protein analysis, especially in protein sequence classification, selection of feature is most important. Popular protein sequence classification…

Databases · Computer Science 2012-11-22 Suprativ Saha , Rituparna Chaki

The identification of compound-protein interactions (CPI) plays a critical role in drug screening, drug repurposing, and combination therapy studies. The effectiveness of CPI prediction relies heavily on the features extracted from both…

Biomolecules · Quantitative Biology 2023-06-16 Li Zhang , Wenhao Li , Haotian Guan , Zhiquan He , Mingjun Cheng , Han Wang

Systematic identification of protein function is a key problem in current biology. Most traditional methods fail to identify functionally equivalent proteins if they lack similar sequences, structural data or extensive manual annotations.…

Genomics · Quantitative Biology 2016-03-08 Dan Ofer

Breast cancer's complexity and variability pose significant challenges in understanding its progression and guiding effective treatment. This study aims to integrate protein sequence data with expression levels to improve the molecular…

Biomolecules · Quantitative Biology 2025-10-31 Hossein Sholehrasa , Majid Jaberi-Douraki

In comparative proteomics studies, LC-MS/MS data is generally quantified using one or both of two measures: the spectral count, derived from the identification of MS/MS spectra, or some measure of ion abundance derived from the LC-MS data.…

Applications · Statistics 2015-03-19 Thomas I. Milac , Timothy W. Randolph , Pei Wang

High-throughput spectrometers are capable of producing data sets containing thousands of spectra for a single biological sample. These data sets contain a substantial amount of redundancy from peptides that may get selected multiple times…

Data Structures and Algorithms · Computer Science 2013-01-08 Fahad Saeed , Trairak Pisitkun , Mark A. Knepper , Jason D. Hoffert

Protein activity is a significant characteristic for recombinant proteins which can be used as biocatalysts. High activity of proteins reduces the cost of biocatalysts. A model that can predict protein activity from amino acid sequence is…

Quantitative Methods · Quantitative Biology 2018-07-23 X. Han , X. Wang , K. Zhou

With the advent of high-throughput wet lab technologies the amount of protein interaction data available publicly has increased substantially, in turn spurring a plethora of computational methods for in silico knowledge discovery from this…

Molecular Networks · Quantitative Biology 2015-05-06 Sriganesh Srihari , Hon Wai Leong

Current metagenomic analysis algorithms require significant computing resources, can report excessive false positives (type I errors), may miss organisms (type II errors / false negatives), or scale poorly on large datasets. This paper…

Databases · Computer Science 2015-01-23 Ashley Mae Conard , Stephanie Dodson , Jeremy Kepner , Darrell Ricke

Composed of amino acid chains that influence how they fold and thus dictating their function and features, proteins are a class of macromolecules that play a central role in major biological processes and are required for the structure,…

Quantitative Methods · Quantitative Biology 2022-07-15 Aaron Wang

In post genomic era with the advent of new technologies a huge amount of complex molecular data are generated with high throughput. The management of this biological data is definitely a challenging task due to complexity and heterogeneity…

Databases · Computer Science 2014-03-13 Ananya Bose , Suprativ Saha

High-quality training datasets are crucial for the development of effective protein design models, but existing synthetic datasets often include unfavorable sequence-structure pairs, impairing generative model performance. We leverage…