English
Related papers

Related papers: Information profiles for DNA pattern discovery

200 papers

Protein motifs are conserved fragments occurred frequently in protein sequences. They have significant functions, such as active site of an enzyme. Search and clustering protein sequence motifs are computational intensive. Most existing…

Genomics · Quantitative Biology 2017-01-03 Haifeng Chen , Ting Chen

In this paper, we introduce the Phantom Domain Finite Element Method (PDFEM), a novel computational approach tailored for the efficient analysis of heterogeneous and composite materials. Inspired by fictitious domain methods, this method…

Numerical Analysis · Mathematics 2025-05-06 Tianlong He , Philippe Karamian-Surville , Daniel Choï

We introduce a novel method to analyse complete genomes and recognise some distinctive features by means of an adaptive compression algorithm, which is not DNA-oriented. We study the Information Content as a function of the number of…

Genomics · Quantitative Biology 2007-05-23 Giulia Menconi

It has been shown that a random-effects framework can be used to test the association between a gene's expression level and the number of DNA copies of a set of genes. This gene-set modelling framework was later applied to find associations…

Methodology · Statistics 2015-10-09 Renée Menezes , Leila Mohammadi , Jelle Goeman , Judith Boer

We propose a solution on the stopping criterion in segmenting inhomogeneous DNA sequences with complex statistical patterns. This new stopping criterion is based on Bayesian Information Criterion (BIC) in the model selection framework. When…

Biological Physics · Physics 2009-11-07 Wentian Li

Previous divide-and-conquer segmentation analyses of DNA sequences do not provide a satisfactory stopping criterion for the recursion. This paper proposes that segmentation be considered as a model selection process. Using the tools in…

Biological Physics · Physics 2007-05-23 Wentian Li

Finite mixture models are useful in applied econometrics. They can be used to model unobserved heterogeneity, which plays major roles in labor economics, industrial organization and other fields. Mixtures are also convenient in dealing with…

Econometrics · Economics 2018-11-08 Yuichi Kitamura , Louise Laage

In this paper, we explore the determination of a spectral emissivity profile that closely matches real data, intended for use as an initial guess and/or a-priori information in a retrieval code. Our approach employs a Bayesian method that…

Applications · Statistics 2024-07-11 Luca Sgheri , Cristina Sgattoni , Chiara Zugarini

Additionally, the strong dependency among in-context examples makes it an NP-hard combinatorial optimization problem and enumerating all permutations is infeasible. Hence we propose LENS, a fiLter-thEN-Search method to tackle this challenge…

Computation and Language · Computer Science 2023-10-10 Xiaonan Li , Xipeng Qiu

High-throughput data analyses are becoming common in biology, communications, economics and sociology. The vast amounts of data are usually represented in the form of matrices and can be considered as knowledge networks. Spectra-based…

Quantitative Methods · Quantitative Biology 2010-01-06 Viet-Anh Nguyen , Zdena Koukolikova-Nicola , Franco Bagnoli , Pietro Lio

Autonomous synthesis and characterization of inorganic materials requires the automatic and accurate analysis of X-ray diffraction spectra. For this task, we designed a probabilistic deep learning algorithm to identify complex multi-phase…

Materials Science · Physics 2021-05-27 Nathan J. Szymanski , Christopher J. Bartel , Yan Zeng , Qingsong Tu , Gerbrand Ceder

Deep Generative Models (DGMs) are versatile tools for learning data representations while adequately incorporating domain knowledge such as the specification of conditional probability distributions. Recently proposed DGMs tackle the…

Machine Learning · Computer Science 2024-01-30 Romain Lopez , Jan-Christian Huetter , Ehsan Hajiramezanali , Jonathan Pritchard , Aviv Regev

The Fuzzy Gene Filter (FGF) is an optimised Fuzzy Inference System designed to rank genes in order of differential expression, based on expression data generated in a microarray experiment. This paper examines the effectiveness of the FGF…

Machine Learning · Computer Science 2011-08-24 Meir Perez , Tshilidzi Marwala

Gene expression and regulation rely on an apparently finely tuned set of reactions between some proteins and DNA. Such DNA-binding proteins have to find specific sequences on very long DNA molecules and they mostly do so in absence of any…

Biological Physics · Physics 2013-12-02 Maria Barbi , Fabien Paillusson

Nowadays data sets are available in very complex and heterogeneous ways. Mining of such data collections is essential to support many real-world applications ranging from healthcare to marketing. In this work, we focus on the analysis of…

Artificial Intelligence · Computer Science 2015-04-10 Aleksey Buzmakov , Elias Egho , Nicolas Jay , Sergei O. Kuznetsov , Amedeo Napoli , Chedy Raïssi

The paper focuses on Image Compression, explaining efficient approaches based on Frequent Pattern Mining(FPM). The proposed compression mechanism is based on clustering similar pixels in the image and thus using cluster identifiers in image…

Image and Video Processing · Electrical Eng. & Systems 2026-02-03 Avinash Kadimisetty , C. Oswald , B. Sivalselvan

Beyond the genetic code, there is another layer of information encoded as chemical modifications on histone proteins positioned along the DNA. Maintaining these modifications is crucial for survival and identity of cells. How the…

Genomics · Quantitative Biology 2020-05-15 Nithya Ramakrishnan , Sibi Raj B Pillai , Ranjith Padinhateeri

A Profile Mixture Model is a model of protein evolution, describing sequence data in which sites are assumed to follow many related substitution processes on a single evolutionary tree. The processes depend in part on different amino acid…

Populations and Evolution · Quantitative Biology 2020-07-07 Samaneh Yourdkhani , Elizabeth S. Allman , John A. Rhodes

Motivation: Many inference tools use the Perfect Phylogeny Model (PPM) to learn trees from noisy variant allele frequency (VAF) data. Learning in this setting is hard, and existing tools use approximate or heuristic algorithms. An…

Quantitative Methods · Quantitative Biology 2019-08-26 Surjyendu Ray , Bei Jia , Sam Safavi , Tim van Opijnen , Ralph Isberg , Jason Rosch , José Bento

We address the problem of learning a syntactic profile for a collection of strings, i.e. a set of regex-like patterns that succinctly describe the syntactic variations in the strings. Real-world datasets, typically curated from multiple…

Machine Learning · Computer Science 2019-04-17 Saswat Padhi , Prateek Jain , Daniel Perelman , Oleksandr Polozov , Sumit Gulwani , Todd Millstein
‹ Prev 1 3 4 5 6 7 10 Next ›