English
Related papers

Related papers: Fast Proteome Identification and Quantification fr…

200 papers

To obtain an electron-density map from a macromolecular crystal the phase-problem needs to be solved, which often involves the use of heavy-atom derivative crystals and concomitantly the determination of the heavy atom substructure. This is…

Biomolecules · Quantitative Biology 2018-08-20 Bjørn Panyella Pedersen , Pontus Gourdon , Xiangyu Liu , Jesper Lykkegaard Karlsen , Poul Nissen

AI-assisted imaging made substantial advances in tumor diagnosis and management. However, a major barrier to developing robust oncology foundation models is the scarcity of large-scale, high-quality annotated datasets, which are limited by…

DNA methylation (DNAme) is a critical component of the epigenetic regulatory machinery and aberrations in DNAme patterns occur in many diseases, such as cancer. Mapping and understanding DNAme profiles offers considerable promise for…

Over the last years, the SWATH data-independent acquisition protocol (Sequential Window acquisition of All THeoretical mass spectra) has become a cornerstone for the worldwide proteomics community. In this approach, a high-resolution…

Quantitative Methods · Quantitative Biology 2019-03-18 Clarissa Braccia , Meritxell Pons Espinal , Mattia Pini , Davide De Pietri Tonelli , Andrea Armirotti

The Protein Data Bank (PDB) contains the atomic structures of over 105 biomolecules with better than 2.8A resolution. The listing of the identities and coordinates of the atoms comprising each macromolecule permits an analysis of the…

Biomolecules · Quantitative Biology 2018-12-05 Hyuntae Na , Daniel ben-Avraham , Monique M. Tirion

Current metagenomic analysis algorithms require significant computing resources, can report excessive false positives (type I errors), may miss organisms (type II errors / false negatives), or scale poorly on large datasets. This paper…

Databases · Computer Science 2015-01-23 Ashley Mae Conard , Stephanie Dodson , Jeremy Kepner , Darrell Ricke

Unsupervised Domain Adaptation (UDA) aims at classifying unlabeled target images leveraging source labeled ones. In this work, we consider the Partial Domain Adaptation (PDA) variant, where we have extra source classes not present in the…

Computer Vision and Pattern Recognition · Computer Science 2022-10-05 Tiago Salvador , Kilian Fatras , Ioannis Mitliagkas , Adam Oberman

Large scale initiatives such as the Human Genome Project, Structural Genomics, and individual research teams have provided large deposits of genomic and proteomic data. The transfer of data to knowledge has become one of the existing…

Databases · Computer Science 2019-11-21 Casey A Cole , Christopher Ott , Diego Valdes , Homayoun Valafar

Topological data analysis (TDA) detects geometric structure in biological data. However, many TDA algorithms are memory intensive and impractical for massive datasets. Here, we introduce a statistical protocol that reduces TDA's memory…

Quantitative Methods · Quantitative Biology 2025-09-05 Andrew J. Stier , Naichen Shi , Raed Al Kontar , Chad Giusti , Marc G. Berman

Large-scale proteomic analysis is emerging as a powerful technique in biology and relies heavily on data acquired by state-of-the-art mass spectrometers. As with any other field in Systems Biology, computational tools are required to deal…

Quantitative Methods · Quantitative Biology 2011-05-02 Fahad Saeed , Trairak Pisitkun , Mark A. Knepper , Jason D. Hoffert

Data-dependent metrics are powerful tools for learning the underlying structure of high-dimensional data. This article develops and analyzes a data-dependent metric known as diffusion state distance (DSD), which compares points using a…

Machine Learning · Statistics 2020-03-10 Lenore Cowen , Kapil Devkota , Xiaozhe Hu , James M. Murphy , Kaiyi Wu

Computer-aided detection systems based on deep learning have shown good performance in breast cancer detection. However, high-density breasts show poorer detection performance since dense tissues can mask or even simulate masses. Therefore,…

Image and Video Processing · Electrical Eng. & Systems 2023-01-25 Lidia Garrucho , Kaisar Kushibar , Richard Osuala , Oliver Diaz , Alessandro Catanese , Javier del Riego , Maciej Bobowicz , Fredrik Strand , Laura Igual , Karim Lekadir

The question of how best to estimate a continuous probability density from finite data is an intriguing open problem at the interface of statistics and physics. Previous work has argued that this problem can be addressed in a natural way…

Data Analysis, Statistics and Probability · Physics 2014-07-16 Justin B. Kinney

Supervised fine-tuning (SFT) is a standard approach for adapting large language models to specialized domains, yet its application to protein sequence modeling and protein language models (PLMs) remains ad hoc. This is in part because…

Machine Learning · Computer Science 2025-12-11 Amin Tavakoli , Raswanth Murugan , Ozan Gokdemir , Arvind Ramanathan , Frances Arnold , Anima Anandkumar

Digital PCR (dPCR) has revolutionized nucleic acid diagnostics by enabling absolute quantification of rare mutations and target sequences. However, current detection methodologies face challenges, as flow cytometers are costly and complex,…

Quantitative Methods · Quantitative Biology 2024-03-29 Yuanyuan Wei , Shanhang Luo , Changran Xu , Yingqi Fu , Qingyue Dong , Yi Zhang , Fuyang Qu , Guangyao Cheng , Yi-Ping Ho , Ho-Pui Ho , Wu Yuan

Unsupervised domain adaptation (UDA) is one of the key technologies to solve a problem where it is hard to obtain ground truth labels needed for supervised learning. In general, UDA assumes that all samples from source and target domains…

Image and Video Processing · Electrical Eng. & Systems 2022-09-07 Satoshi Kondo

Proteomics can be defined as the large-scale analysis of proteins. Due to the complexity of biological systems, it is required to concatenate various separation techniques prior to mass spectrometry. These techniques, dealing with proteins…

Protein structure reconstruction from Nuclear Magnetic Resonance (NMR) experiments largely relies on computational algorithms. Recently, some effective low-rank matrix completion (MC) methods, such as ASD and ScaledASD, have been…

Biological Physics · Physics 2018-09-24 Z. Li , S. Li , X. Wei , X. Peng , Q. Zhao

High-throughput genetic and epigenetic data are often screened for associations with an observed phenotype. For example, one may wish to test hundreds of thousands of genetic variants, or DNA methylation sites, for an association with…

Methodology · Statistics 2017-10-20 Eric F. Lock , David B. Dunson

Life science is entering a new era of petabyte-level sequencing data. Converting such big data to biological insights represents a huge challenge for computational analysis. To this end, we developed DeepMetabolism, a biology-guided deep…

Genomics · Quantitative Biology 2017-05-10 Weihua Guo , You Xu , Xueyang Feng