English
Related papers

Related papers: Aligning 415 519 proteins in less than two hours o…

200 papers

Motivation: Recent advances in sequencing technologies promise ultra-long reads of $\sim$100 kilo bases (kb) in average, full-length mRNA or cDNA reads in high throughput and genomic contigs over 100 mega bases (Mb) in length. Existing…

Genomics · Quantitative Biology 2018-09-17 Heng Li

A method to search for local structural similarities in proteins at atomic resolution is presented. It is demonstrated that a huge amount of structural data can be handled within a reasonable CPU time by using a conventional relational…

Biomolecules · Quantitative Biology 2007-12-28 Akira R. Kinjo , Haruki Nakamura

A new method for the Automated Protein Structure Analysis (APSA) is derived, which simplifies the protein backbone to a smooth curve in 3-dimensional space. For the purpose of obtaining this smooth line each amino acid is represented by its…

Quantitative Methods · Quantitative Biology 2008-11-24 Sushilee Raganathan , Dmitry Izotov , Elfi Kraka , Dieter Cremer

Background: The increasing volume and variety of genotypic and phenotypic data is a major defining characteristic of modern biomedical sciences. At the same time, the limitations in technology for generating data and the inherently…

Quantitative Methods · Quantitative Biology 2016-12-07 Yuxiang Jiang , Tal Ronnen Oron , Wyatt T Clark , Asma R Bankapur , Daniel D'Andrea , Rosalba Lepore , Christopher S Funk , Indika Kahanda , Karin M Verspoor , Asa Ben-Hur , Emily Koo , Duncan Penfold-Brown , Dennis Shasha , Noah Youngs , Richard Bonneau , Alexandra Lin , Sayed ME Sahraeian , Pier Luigi Martelli , Giuseppe Profiti , Rita Casadio , Renzhi Cao , Zhaolong Zhong , Jianlin Cheng , Adrian Altenhoff , Nives Skunca , Christophe Dessimoz , Tunca Dogan , Kai Hakala , Suwisa Kaewphan , Farrokh Mehryary , Tapio Salakoski , Filip Ginter , Hai Fang , Ben Smithers , Matt Oates , Julian Gough , Petri Törönen , Patrik Koskinen , Liisa Holm , Ching-Tai Chen , Wen-Lian Hsu , Kevin Bryson , Domenico Cozzetto , Federico Minneci , David T Jones , Samuel Chapman , Dukka B K. C. , Ishita K Khan , Daisuke Kihara , Dan Ofer , Nadav Rappoport , Amos Stern , Elena Cibrian-Uhalte , Paul Denny , Rebecca E Foulger , Reija Hieta , Duncan Legge , Ruth C Lovering , Michele Magrane , Anna N Melidoni , Prudence Mutowo-Meullenet , Klemens Pichler , Aleksandra Shypitsyna , Biao Li , Pooya Zakeri , Sarah ElShal , Léon-Charles Tranchevent , Sayoni Das , Natalie L Dawson , David Lee , Jonathan G Lees , Ian Sillitoe , Prajwal Bhat , Tamás Nepusz , Alfonso E Romero , Rajkumar Sasidharan , Haixuan Yang , Alberto Paccanaro , Jesse Gillis , Adriana E Sedeño-Cortés , Paul Pavlidis , Shou Feng , Juan M Cejuela , Tatyana Goldberg , Tobias Hamp , Lothar Richter , Asaf Salamov , Toni Gabaldon , Marina Marcet-Houben , Fran Supek , Qingtian Gong , Wei Ning , Yuanpeng Zhou , Weidong Tian , Marco Falda , Paolo Fontana , Enrico Lavezzo , Stefano Toppo , Carlo Ferrari , Manuel Giollo , Damiano Piovesan , Silvio Tosatto , Angela del Pozo , José M Fernández , Paolo Maietta , Alfonso Valencia , Michael L Tress , Alfredo Benso , Stefano Di Carlo , Gianfranco Politano , Alessandro Savino , Hafeez Ur Rehman , Matteo Re , Marco Mesiti , Giorgio Valentini , Joachim W Bargsten , Aalt DJ van Dijk , Branislava Gemovic , Sanja Glisic , Vladmir Perovic , Veljko Veljkovic , Nevena Veljkovic , Danillo C Almeida-e-Silva , Ricardo ZN Vencio , Malvika Sharan , Jörg Vogel , Lakesh Kansakar , Shanshan Zhang , Slobodan Vucetic , Zheng Wang , Michael JE Sternberg , Mark N Wass , Rachael P Huntley , Maria J Martin , Claire O'Donovan , Peter N Robinson , Yves Moreau , Anna Tramontano , Patricia C Babbitt , Steven E Brenner , Michal Linial , Christine A Orengo , Burkhard Rost , Casey S Greene , Sean D Mooney , Iddo Friedberg , Predrag Radivojac

Principal component analysis (PCA) is widely used for dimension reduction and embedding of real data in social network analysis, information retrieval, and natural language processing, etc. In this work we propose a fast randomized PCA…

Machine Learning · Computer Science 2018-10-17 Xu Feng , Yuyang Xie , Mingye Song , Wenjian Yu , Jie Tang

We improve on GenASM, a recent algorithm for genomic sequence alignment, by significantly reducing its memory footprint and bandwidth requirement. Our algorithmic improvements reduce the memory footprint by 24$\times$ and the number of…

Hardware Architecture · Computer Science 2022-03-30 Joël Lindegger , Damla Senol Cali , Mohammed Alser , Juan Gómez-Luna , Onur Mutlu

Predicting which proteins interact together from amino-acid sequences is an important task. We develop a method to pair interacting protein sequences which leverages the power of protein language models trained on multiple sequence…

Biomolecules · Quantitative Biology 2024-12-30 Umberto Lupo , Damiano Sgarbossa , Anne-Florence Bitbol

Data mining, particularly the analysis of multivariate time series data, plays a crucial role in extracting insights from complex systems and supporting informed decision-making across diverse domains. However, assessing the similarity of…

Machine Learning · Computer Science 2025-07-15 Franck Tonle , Henri Tonnang , Milliam Ndadji , Maurice Tchendji , Armand Nzeukou , Kennedy Senagi , Saliou Niassy

Genome sequence analysis has enabled significant advancements in medical and scientific areas such as personalized medicine, outbreak tracing, and the understanding of evolution. Unfortunately, it is currently bottlenecked by the…

Molecular similarity search has been widely used in drug discovery to identify structurally similar compounds from large molecular databases rapidly. With the increasing size of chemical libraries, there is growing interest in the efficient…

Hardware Architecture · Computer Science 2021-09-15 Hongwu Peng , Shiyang Chen , Zhepeng Wang , Junhuan Yang , Scott A. Weitze , Tong Geng , Ang Li , Jinbo Bi , Minghu Song , Weiwen Jiang , Hang Liu , Caiwen Ding

We investigate parameter-efficient fine-tuning (PEFT) methods that can provide good accuracy under limited computational and memory budgets in the context of large language models (LLMs). We present a new PEFT method called Robust…

Computation and Language · Computer Science 2024-06-04 Mahdi Nikdan , Soroush Tabesh , Elvir Crnčević , Dan Alistarh

Tensor methods have gained increasingly attention from various applications, including machine learning, quantum chemistry, healthcare analytics, social network analysis, data mining, and signal processing, to name a few. Sparse tensors and…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-02-12 Jiajia Li , Yuchen Ma , Xiaolong Wu , Ang Li , Kevin Barker

The analysis of the three-dimensional structure of proteins is an important topic in molecular biochemistry. Structure plays a critical role in defining the function of proteins and is more strongly conserved than amino acid sequence over…

Applications · Statistics 2015-01-19 Abel Rodriguez , Scott C. Schmidler

Frameshift mutations in protein-coding DNA sequences produce a drastic change in the resulting protein sequence, which prevents classic protein alignment methods from revealing the proteins' common origin. Moreover, when a large number of…

Quantitative Methods · Quantitative Biology 2011-01-18 Marta L. Gîrdea , Laurent Noé , Gregory Kucherov

Classification of proteins based on their structure provides a valuable resource for studying protein structure, function and evolutionary relationships. With the rapidly increasing number of known protein structures, manual and…

Computational Engineering, Finance, and Science · Computer Science 2009-07-14 Oktie Hassanzadeh

We present Fast Approximate Minimum Spanning Tree (FAMST), a novel algorithm that addresses the computational challenges of constructing Minimum Spanning Trees (MSTs) for large-scale and high-dimensional datasets. FAMST utilizes a…

Data Structures and Algorithms · Computer Science 2025-07-22 Mahmood K. M. Almansoori , Miklos Telek

Time series analysis is a key technique for extracting and predicting events in domains as diverse as epidemiology, genomics, neuroscience, environmental sciences, economics, and more. Matrix profile, the state-of-the-art algorithm to…

Genome sequence alignment is the core of many biological applications. The advancement of sequencing technologies produces a tremendous amount of data, making sequence alignment a critical bottleneck in bioinformatics analysis. The existing…

Hardware Architecture · Computer Science 2023-01-26 Weihong Xu , Saransh Gupta , Niema Moshiri , Tajana Rosing

Genome sequence analysis plays a pivotal role in enabling many medical and scientific advancements in personalized medicine, outbreak tracing, and forensics. However, the analysis of genome sequencing data is currently bottlenecked by the…

Hardware Architecture · Computer Science 2021-11-04 Damla Senol Cali

Background:Prediction of protein three-dimensional structures from amino acid sequences is a long-standing goal in computational/molecular biology. The successful discrimination of protein folds would help to improve the accuracy of protein…

Biomolecules · Quantitative Biology 2007-05-23 Y-h. Taguchi , M. Michael Gromiha
‹ Prev 1 3 4 5 6 7 10 Next ›