English
Related papers

Related papers: Probability that a chromosome is lost without trac…

200 papers

A probabilistic query may not be estimable from observed data corrupted by missing values if the data are not missing at random (MAR). It is therefore of theoretical interest and practical importance to determine in principle whether a…

Machine Learning · Statistics 2016-11-16 Jin Tian

Motivated by DNA-based storage applications, we study the problem of reconstructing a coded sequence from multiple traces. We consider the model where the traces are outputs of independent deletion channels, where each channel deletes each…

Information Theory · Computer Science 2022-07-13 Serge Kas Hanna

Using techniques from Poisson approximation, we prove explicit error bounds on the number of permutations that avoid any pattern. Most generally, we bound the total variation distance between the joint distribution of pattern occurrences…

Combinatorics · Mathematics 2023-06-22 Harry Crane , Stephen DeSalvo

Under the effect of strong genetic drift, it is highly probable to observe gene fixation or gene loss in a population, shown by infinite peaks on a coherently constructed potential energy landscape. It is then important to ask what such…

Populations and Evolution · Quantitative Biology 2015-06-15 Song Xu , Shuyun Jiao , Pengyao Jiang , Ping Ao

Hi-C experiments are used to infer the contact probabilities between loci separated by varying genome lengths. Contact probability should decrease as the spatial distance between two loci increases. However, studies comparing Hi-C and FISH…

Soft Condensed Matter · Physics 2020-05-25 Guang Shi , D. Thirumalai

For a genetic locus carrying a strongly beneficial allele which has just fixed in a large population, we study the ancestry at a linked neutral locus. During this ``selective sweep'' the linkage between the two loci is broken up by…

Probability · Mathematics 2007-05-23 Alison Etheridge , Peter Pfaffelhuber , Anton Wakolbinger

We consider the maximum coding rate achievable by uniformly-random codes for the deletion channel. We prove an upper bound that's within 0.1 of the best known lower bounds for all values of the deletion probability $d,$ and much closer for…

Information Theory · Computer Science 2022-10-17 Berivan Isik , Francisco Pernice , Tsachy Weissman

In the trace reconstruction problem, the goal is to reconstruct an unknown string $x$ of length $n$ from multiple traces obtained by passing $x$ through the deletion channel. In the relaxed problem of $approximate$ trace reconstruction, the…

Probability · Mathematics 2021-07-15 Zachary Chase , Yuval Peres

Mixture models have received considerable attention recently and Newton [Sankhy\={a} Ser. A 64 (2002) 306--322] proposed a fast recursive algorithm for estimating a mixing distribution. We prove almost sure consistency of this recursive…

Statistics Theory · Mathematics 2009-08-25 Surya T. Tokdar , Ryan Martin , Jayanta K. Ghosh

The paper suggests a frequency criterion of error-free recoverability of a missing value for sequences, i.e. discrete time processes, in a pathwise setting without probabilistic assumptions. The paper establishes error-free recoverability…

Information Theory · Computer Science 2018-09-19 Nikolai Dokuchaev

Probability estimation is an elementary building block of every statistical data compression algorithm. In practice probability estimation is often based on relative letter frequencies which get scaled down, when their sum is too large.…

Information Theory · Computer Science 2015-01-12 Christopher Mattern

The fitness of a biological strategy is typically measured by its expected reproductive rate, the first moment of its offspring distribution. However, strategies with high expected rates can also have high probabilities of extinction. A…

Populations and Evolution · Quantitative Biology 2013-05-17 Sterling Sawaya , Steffen Klaere

The stationary sampling distribution of a neutral decoupled Moran or Wright-Fisher diffusion with neutral mutations is known to first order for a general rate matrix with small but otherwise unconstrained mutation rates. Using this…

Populations and Evolution · Quantitative Biology 2020-05-07 Claus Vogl , Lynette C. Mikula , Conrad J. Burden

We propose methods for estimating correspondence between two point sets under the presence of outliers in both the source and target sets. The proposed algorithms expand upon the theory of the regression without correspondence problem to…

Machine Learning · Statistics 2019-10-29 Amin Nejatbakhsh , Erdem Varol

In this paper, we derive an expression for the expected number of runs in a trace of a binary sequence $x \in \{0,1\}^n$ obtained by passing $x$ through a deletion channel that independently deletes each bit with probability $q$. We use…

Information Theory · Computer Science 2025-11-06 Shiv Pratap Singh Rathore , Navin Kashyap

We consider models of nucleotidic substitution processes where the rate of substitution at a given site depends on the state of its neighbours. For a wide class of such nonreversible models, we show how to compute consistent, mathematically…

Probability · Mathematics 2010-03-25 Mikael Falconnet

We propose a modification to the random destruction of graphs: Given a finite network with a distinguished set of sources and targets, remove (cut) vertices at random, discarding components that do not contain a source node. We investigate…

Probability · Mathematics 2022-01-05 Fabian Burghart

We consider a general statistical learning problem where an unknown fraction of the training data is corrupted. We develop a robust learning method that only requires specifying an upper bound on the corrupted data fraction. The method…

Machine Learning · Statistics 2020-02-10 Muhammad Osama , Dave Zachariah , Peter Stoica

Many random combinatorial objects have a component structure whose joint distribution is equal to that of a process of mutually independent random variables, conditioned on the value of a weighted sum of the variables. It is interesting to…

Probability · Mathematics 2013-08-16 Richard Arratia , Simon Tavare

Based on a discrete version of the Pollaczeck-Khinchine formula, a general method to calculate the ultimate ruin probability in the Gerber-Dickson risk model is provided when claims follow a negative binomial mixture distribution. The…

Probability · Mathematics 2020-06-03 David J. Santana , Luis Rincon
‹ Prev 1 3 4 5 6 7 10 Next ›