中文
相关论文

相关论文: Identifying DNA motifs based on match and mismatch…

200 篇论文

In semi-supervised segmentation, capturing meaningful semantic structures from unlabeled data is essential. This is particularly challenging in histopathology image analysis, where objects are densely distributed. To address this issue, we…

计算机视觉与模式识别 · 计算机科学 2025-10-03 Meilong Xu , Xiaoling Hu , Shahira Abousamra , Chen Li , Chao Chen

Traditional methods for matching in causal inference are impractical for high-dimensional datasets. They suffer from the curse of dimensionality: exact matching and coarsened exact matching find exponentially fewer matches as the input…

机器学习 · 统计学 2026-02-12 Oscar Clivio , Fabian Falck , Brieuc Lehmann , George Deligiannidis , Chris Holmes

This paper focuses on pattern matching in the DNA sequence. It was inspired by a previously reported method that proposes encoding both pattern and sequence using prime numbers. Although fast, the method is limited to rather small pattern…

计算机视觉与模式识别 · 计算机科学 2016-11-21 Janja Paliska Soldo , Ana Sovic Krzic , and Damir Sersic

Sequence discovery tools play a central role in several fields of computational biology. In the framework of Transcription Factor binding studies, motif finding algorithms of increasingly high performance are required to process the big…

定量方法 · 定量生物学 2014-08-27 Nicolò Colombo , Nikos Vlassis

If the probability model is correctly specified, then we can estimate the covariance matrix of the asymptotic maximum likelihood estimate distribution using either the first or second derivatives of the likelihood function. Therefore, if…

统计方法学 · 统计学 2024-11-06 Reyhaneh Hosseinpourkhoshkbari , Richard M. Golden

With the advent of large-scale heterogeneous search engines comes the problem of unified search control resulting in mismatches that could have otherwise avoided. A mechanism is needed to determine exact patterns in web mining and…

密码学与安全 · 计算机科学 2019-01-28 Nazim Uddin Sheikh , Hasina Rahman , Hamid Al-Qahtani

Each human genome is a 3 billion base pair set of encoding instructions. Decoding the genome using deep learning fundamentally differs from most tasks, as we do not know the full structure of the data and therefore cannot design…

机器学习 · 计算机科学 2016-05-24 Laura Deming , Sasha Targ , Nate Sauder , Diogo Almeida , Chun Jimmie Ye

Gene annotation has traditionally required direct comparison of DNA sequences between an unknown gene and a database of known ones using string comparison methods. However, these methods do not provide useful information when a gene does…

机器学习 · 计算机科学 2019-09-17 James K. Senter , Taylor M. Royalty , Andrew D. Steen , Amir Sadovnik

Identifying the complete set of functional elements within the human genome would be a windfall for multiple areas of biological research including medicine, molecular biology, and evolution. Complete knowledge of function would aid in the…

种群与进化 · 定量生物学 2015-11-24 Daniel R. Schrider , Andrew D. Kern

Next-generation sequencing technology enables the identification of thousands of gene regulatory sequences in many cell types and organisms. We consider the problem of testing if two such sequences differ in their number of binding site…

基因组学 · 定量生物学 2014-02-04 Dennis Kostka , Tara Friedrich , Alisha K. Holloway , Katherine S. Pollard

Motivation: Spliced alignment refers to the alignment of messenger RNA (mRNA) or protein sequences to eukaryotic genomes. It plays a critical role in gene annotation and the study of gene functions. Accurate spliced alignment demands…

基因组学 · 定量生物学 2025-09-23 Siying Yang , Neng Huang , Heng Li

The automatic assignment of species information to the corresponding genes in a research article is a critically important step in the gene normalization task, whereby a gene mention is normalized and linked to a database record or…

计算与语言 · 计算机科学 2022-10-17 Ling Luo , Chih-Hsuan Wei , Po-Ting Lai , Qingyu Chen , Rezarta Islamaj Doğan , Zhiyong Lu

A number of signal processing and statistical methods can be used in analyzing either pieces of text or DNA sequences. These techniques can be used in a number of ways, such as determining authorship of documents, finding genes in DNA, and…

数据分析、统计与概率 · 物理学 2007-05-23 Matthew J. Berryman , Andrew Allison , Pedro Carpena , Derek Abbott

Computational methods are needed to differentiate the small fraction of missense mutations that contribute to disease by disrupting protein function from neutral variants. We describe several complementary methods using large-scale homology…

生物大分子 · 定量生物学 2013-08-22 Andrew J. Bordner , Barry Zorman

Much of the on-going statistical analysis of DNA sequences is focused on the estimation of characteristics of coding and non-coding regions that would possibly allow discrimination of these regions. In the current approach, we concentrate…

基因组学 · 定量生物学 2009-11-10 D. Kugiumtzis , A. Provata

We propose a novel framework for combining datasets via alignment of their intrinsic geometry. This alignment can be used to fuse data originating from disparate modalities, or to correct batch effects while preserving intrinsic data…

机器学习 · 计算机科学 2020-01-31 Jay S. Stanley , Scott Gigante , Guy Wolf , Smita Krishnaswamy

Detection of false-positive motifs is one of the main causes of low performance in motif finding methods. It is generally assumed that false-positives are mostly due to algorithmic weakness of motif-finders. Here, however, we derive the…

基因组学 · 定量生物学 2010-12-23 Amin Zia , Alan M. Moses

It is well known that the general transcription factors (GTF) specifically recognize correct TATA boxes, distinguishing them from many others. Employing the principles of determinacy analysis (mathematical theory of rules) we analyzed a…

基因组学 · 定量生物学 2012-03-30 Sergey V. Chesnokov , Lina G. Chesnokov , Viktor Wixler

We present an online algorithm to deal with pattern matching in strings. The problem we investigate is commonly known as string matching with mismatches in which the objective is to report the number of characters that match when a pattern…

数据结构与算法 · 计算机科学 2016-03-11 Vinodprasad P

This paper addresses the problem of handling spatial misalignments due to camera-view changes or human-pose variations in person re-identification. We first introduce a boosting-based approach to learn a correspondence structure which…

计算机视觉与模式识别 · 计算机科学 2016-04-28 Yang Shen , Weiyao Lin , Junchi Yan , Mingliang Xu , Jianxin Wu , Jingdong Wang