中文
相关论文

相关论文: Shotgun DNA sequencing for human identification: D…

200 篇论文

Ever since deoxyribonucleic acid (DNA) was considered as a next-generation data-storage medium, lots of research efforts have been made to correct errors occurred during the synthesis, storage, and sequencing processes using error…

信息论 · 计算机科学 2023-06-14 Jaeho Jeong , Hosung Park , Hee-Youl Kwak , Jong-Seon No , Hahyeon Jeon , Jeong Wook Lee , Jae-Won Kim

A DNA palindrome is a segment of double-stranded DNA sequence with inver- sion symmetry which may form secondary structures conferring significant biolog- ical functions ranging from RNA transcription to DNA replication. To test if the…

应用统计 · 统计学 2011-04-28 I-Ping Tu , Yuan-Fu Huang , Shao-Hsuan Wang

Segmental duplications (SDs), or low-copy repeats (LCR), are segments of DNA greater than 1 Kbp with high sequence identity that are copied to other regions of the genome. SDs are among the most important sources of evolution, a common…

数据结构与算法 · 计算机科学 2018-09-25 Ibrahim Numanagić , Alim S. Gökkaya , Lillian Zhang , Bonnie Berger , Can Alkan , Faraz Hach

Training datasets are crucial for convolutional neural network-based algorithms, which directly impact their overall performance. As such, using a well-structured dataset that has minimum level of bias is always desirable. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2021-06-29 Ekberjan Derman

Classic concepts of genetic (gene) diversity (heterozygosity) such as Nei (1973: PNAS) and Nei and Li (1979: PNAS) nucleotide diversity were defined within the context of populations. Although variations are often measured in population…

种群与进化 · 定量生物学 2019-03-13 Zhanshan , Ma , Lianwei Li , Ya-Ping Zhang

Single-Cell RNA sequencing (scRNA-seq) measurements have facilitated genome-scale transcriptomic profiling of individual cells, with the hope of deconvolving cellular dynamic changes in corresponding cell sub-populations to better…

基因组学 · 定量生物学 2021-04-06 Seyednami Niyakan , Ehsan Hajiramezanali , Shahin Boluki , Siamak Zamani Dadaneh , Xiaoning Qian

Motivation: Whole-genome high-coverage sequencing has been widely used for personal and cancer genomics as well as in various research areas. However, in the lack of an unbiased whole-genome truth set, the global error rate of variant calls…

基因组学 · 定量生物学 2018-07-27 Heng Li

In 2016, the European Network of Forensic Science Institutes (ENFSI) published guidelines for the evaluation, interpretation and reporting of scientific evidence. In the guidelines, ENFSI endorsed the use of the likelihood ratio (LR) as a…

应用统计 · 统计学 2020-05-25 Nathaniel Garton , Danica Ommen , Jarad Niemi , Alicia Carriquiry

Motivation: Genome-Wide Association Studies (GWAS) seek to identify causal genomic variants associated with rare human diseases. The classical statistical approach for detecting these variants is based on univariate hypothesis testing, with…

统计方法学 · 统计学 2018-10-22 Florent Guinot , Marie Szafranski , Christophe Ambroise , Franck Samson

We propose sequenced-replacement sampling (SRS) for training deep neural networks. The basic idea is to assign a fixed sequence index to each sample in the dataset. Once a mini-batch is randomly drawn in each training iteration, we refill…

机器学习 · 计算机科学 2018-10-22 Chiu Man Ho , Dae Hoon Park , Wei Yang , Yi Chang

A common task in forensic biology is to interpret and evaluate short tandem repeat DNA profiles. The first step in these interpretations is to assign a number of contributors to the profiles, a task that is most often performed manually by…

机器学习 · 计算机科学 2024-12-16 Duncan Taylor , Melissa A. Humphries

Single nucleotide polymorphism (SNP) datasets are fundamental to genetic studies but pose significant privacy risks when shared. The correlation of SNPs with each other makes strong adversarial attacks such as masked-value reconstruction,…

机器学习 · 计算机科学 2025-10-08 Shadi Rahimian , Mario Fritz

The alignment of biological sequences such as DNA, RNA, and proteins, is one of the basic tools that allow to detect evolutionary patterns, as well as functional/structural characterizations between homologous sequences in different…

定量方法 · 定量生物学 2023-05-01 Louise Budzynski , Andrea Pagnani

In this paper, association results from genome-wide association studies (GWAS) are combined with a deep learning framework to test the predictive capacity of statistically significant single nucleotide polymorphism (SNPs) associated with…

计算机与社会 · 计算机科学 2018-08-27 Casimiro Adays Curbelo Montañez , Paul Fergus , Almudena Curbelo Montañez , Carl Chalmers

The rapidly changing landscape of sequencing technologies brings new opportunities to genomics research. Longer sequence reads and higher sequence throughput coupled with ever-improving base accuracy and decreasing per-base cost is now…

基因组学 · 定量生物学 2022-09-20 René L. Warren

Storing digital data in synthetic DNA faces challenges in ensuring data reliability in the presence of edit errors--deletions, insertions, and substitutions--that occur randomly during various stages of the storage process. Current…

信息论 · 计算机科学 2025-09-11 Serge Kas Hanna

Despite much progress over the past decade, current Single Nucleotide Polymorphism (SNP) genotyping technologies still offer an insufficient degree of multiplexing when required to handle user-selected sets of SNPs. In this paper we propose…

数据结构与算法 · 计算机科学 2007-05-23 Ion I. Mandoiu , Claudia Prajescu

At the core of high throughput DNA sequencing platforms lies a bio-physical surface process that results in a random geometry of clusters of homogenous short DNA fragments typically hundreds of base pairs long - bridge amplification. The…

基因组学 · 定量生物学 2015-08-13 Eliza O'Reilly , Francois Baccelli , Gustavo de Veciana , Haris Vikalo

The advent of "next-generation" DNA sequencing (NGS) technologies has meant that collections of hundreds of millions of DNA sequences are now commonplace in bioinformatics. Knowing the longest common prefix array (LCP) of such a collection…

数据结构与算法 · 计算机科学 2013-05-02 Markus J. Bauer , Anthony J. Cox , Giovanna Rosone , Marinella Sciortino

In this paper, we present a novel unsupervised algorithm for word sense disambiguation (WSD) at the document level. Our algorithm is inspired by a widely-used approach in the field of genetics for whole genome sequencing, known as the…

计算与语言 · 计算机科学 2017-07-26 Andrei M. Butnaru , Radu Tudor Ionescu , Florentina Hristea