English
Related papers

Related papers: The probability of finding a fixed pattern in rand…

200 papers

We present algorithms for topic modeling based on the geometry of cross-document word-frequency patterns. This perspective gains significance under the so called separability condition. This is a condition on existence of novel-words that…

Machine Learning · Statistics 2013-03-19 Weicong Ding , Mohammad H. Rohban , Prakash Ishwar , Venkatesh Saligrama

A sorting network is a shortest path from 12..n to n..21 in the Cayley graph of the symmetric group S(n) generated by nearest-neighbor swaps. A pattern is a sequence of swaps that forms an initial segment of some sorting network. We prove…

Probability · Mathematics 2012-11-21 Omer Angel , Vadim Gorin , Alexander E. Holroyd

We say that a random integer variable $X$ is monotone if the modulus of the characteristic function of $X$ is decreasing on $[0,\pi]$. This is the case for many commonly encountered variables, e.g., Bernoulli, Poisson and geometric random…

Probability · Mathematics 2021-04-14 Anders Aamand , Noga Alon , Jakob Bæk Tejs Knudsen , Mikkel Thorup

We consider a minimal model of persistent random searcher with short range memory. We calculate exactly for such searcher the mean first-passage time to a target in a bounded domain and find that it admits a non trivial minimum as function…

Statistical Mechanics · Physics 2012-02-28 V. Tejedor , R. Voituriez , O. Bénichou

We provide finite-sample distribution approximations, that are uniform in the parameter, for inference in linear mixed models. Focus is on variances and covariances of random effects in cases where existing theory fails because their…

Statistics Theory · Mathematics 2025-07-29 Karl Oskar Ekvall , Matteo Bottai

We report the existence of deterministic patterns in plots showing the relationship between the mean and the Fano factor (ratio of variance and mean) of stochastic count data. These patterns are found in a wide variety of datasets,…

Quantitative Methods · Quantitative Biology 2024-05-29 Zhixing Cao , Yiling Wang , Ramon Grima

We consider variations on the following problem: given an NFA M and a pattern p, does there exist an x in L(M) such that p matches x? We consider the restricted problem where M only accepts a finite language. We also consider the variation…

Formal Languages and Automata Theory · Computer Science 2009-06-18 Narad Rampersad , Jeffrey Shallit

Two results on palindromicity of bi-infinite words in a finite alphabet are presented. The first is a simple, but efficient criterion to exclude palindromicity of minimal sequences and applies, in particular, to the Rudin-Shapiro sequence.…

Mathematical Physics · Physics 2019-07-17 Michael Baake

We propose a randomized greedy search algorithm to find a point estimate for a random partition based on a loss function and posterior Monte Carlo samples. Given the large size and awkward discrete nature of the search space, the…

Methodology · Statistics 2021-05-11 David B. Dahl , Devin J. Johnson , Peter Mueller

As is the case of many signals produced by complex systems, language presents a statistical structure that is balanced between order and disorder. Here we review and extend recent results from quantitative characterisations of the degree of…

Computation and Language · Computer Science 2015-03-05 Marcelo A Montemurro , Damián H Zanette

We consider the following generalization of binary search in sorted arrays to tree domains. In each step of the search, an algorithm is querying a vertex $q$, and as a reply, it receives an answer, which either states that $q$ is the…

Data Structures and Algorithms · Computer Science 2024-01-26 Dariusz Dereniowski , Izajasz Wrosz

We explain how certain tools from convex analysis and probability theory may be used in order to obtain counting results for the number of words with prescribed frequencies of letters in regular languages.

Combinatorics · Mathematics 2023-11-20 Rostislav Grigorchuk , Jean-François Quint

In this paper, we introduce a tag recommendation algorithm that mimics the way humans draw on items in their long-term memory. This approach uses the frequency and recency of previous tag assignments to estimate the probability of reusing a…

Information Retrieval · Computer Science 2013-12-19 Dominik Kowald , Paul Seitlinger , Christoph Trattner , Tobias Ley

For a sample of Exponentially distributed durations we aim at point estimation and a confidence interval for its parameter. A duration is only observed if it has ended within a certain time interval, determined by a Uniform distribution.…

Methodology · Statistics 2021-10-19 Rafael Weißbach , Dominik Wied

Synonyms and homonyms appear in all natural languages. We analyse their evolution within the framework of the signaling game. Agents in our model use reinforcement learning, where probabilities of selection of a communicated word or of its…

Physics and Society · Physics 2022-01-28 Dorota Lipowski , Adam Lipowski

We characterize the minimum-length sequences of independent lazy simple transpositions whose composition is a uniformly random permutation. For every reduced word of the reverse permutation there is exactly one valid way to assign…

Probability · Mathematics 2018-03-09 Omer Angel , Alexander E Holroyd

When flipping a fair coin, let $W = L_1L_2...L_N$ with $L_i\in\{H,T\}$ be a binary word of length $N=2$ or $N=3$. In this paper, we establish second- and third-order linear recurrence relations and their generating functions to discuss the…

The escape probability $\xi_{x}$ from a site $x$ of a one-dimensional disordered lattice with trapping is treated as a discrete dynamical evolution by random iterations over nonlinear maps parametrized by the right and left jump…

Condensed Matter · Physics 2016-08-31 Thomas Wichmann , Achille Giacometti , K. P. N. Murthy

A key issue in the handling of temporal data is the treatment of persistence; in most approaches it consists in inferring defeasible confusions by extrapolating from the actual knowledge of the history of the world; we propose here a…

Artificial Intelligence · Computer Science 2013-03-08 Dimiter Driankov , Jerome Lang

The key point limits to define the {\it statistical model} describing the data distribution. Hence, it turns out that the characteristics related to the so-called. Inverse Tully-Fisher relation and the Direct relation are maximum likelyhood…

Astrophysics · Physics 2007-05-23 R. Triay