English
Related papers

Related papers: A Comparison of Two Proximity Catch Digraph Famili…

200 papers

We prove that if a rectangular matrix with uniformly small entries and approximately orthogonal rows is applied to the independent standardized random variables with uniformly bounded third moments, then the empirical CDF of the resulting…

Probability · Mathematics 2007-06-14 Bernard Bercu , Wlodzimierz Bryc

The pair contact process with diffusion (PCPD) is studied with a standard Monte Carlo approach and with simulations at fixed densities. A standard analysis of the simulation results, based on the particle densities or on the pair densities,…

Statistical Mechanics · Physics 2009-11-13 F. Smallenburg , G. T. Barkema

Propensity score matching is a common tool for adjusting for observed confounding in observational studies, but is known to have limitations in the presence of unmeasured confounding. In many settings, researchers are confronted with…

Methodology · Statistics 2017-12-08 Georgia Papadogeorgou , Christine Choirat , Corwin Zigler

Margin-based classifiers have been popular in both machine learning and statistics for classification problems. Since a large number of classifiers are available, one natural question is which type of classifiers should be used given a…

Machine Learning · Statistics 2021-10-19 Hanwen Huang , Qinglong Yang

We introduce a new approach for comparing the predictive accuracy of two nested models that bypasses the difficulties caused by the degeneracy of the asymptotic variance of forecast error loss differentials used in the construction of…

Econometrics · Economics 2023-10-17 Jean-Yves Pitarakis

We propose a bootstrap procedure for data that may exhibit clustering in two or more dimensions. We use insights from the theory of generalized U-statistics to analyze the large-sample properties of statistics that are sample averages from…

Methodology · Statistics 2017-12-06 Konrad Menzel

We consider DNA codes based on the nearest-neighbor (stem) similarity model which adequately reflects the "hybridization potential" of two DNA sequences. Our aim is to present a survey of bounds on the rate of DNA codes with respect to a…

Information Theory · Computer Science 2016-11-17 A. D'yachkov , A. Voronina , A. Macula , T. Renz , V. Rykov

We investigate the pairwise negative correlation (p-NC) property for uniform probability measures on several families of spanning subgraphs of the complete graph $K_n$. Motivated by conjectured negative dependence properties of the…

Probability · Mathematics 2026-03-12 Pengfei Tang , Zibo Zhang

Understanding the metric structure of permutation families is fundamental to combinatorics and has applications in social choice theory, bioinformatics, and coding theory. We study permutation families defined by restriction…

Discrete Mathematics · Computer Science 2025-07-16 Danylo Tymoshenko , Leonhard Nagel

It is well known that the clustering of galaxies depends on galaxy type.Such relative bias complicates the inference of cosmological parameters from galaxy redshift surveys, and is a challenge to theories of galaxy formation and evolution.…

I present the Phase Distance Correlation (PDC) periodogram -- a new periodicity metric, based on the Distance Correlation concept of G\'abor Sz\'ekely. For each trial period PDC calculates the distance correlation between the data samples…

Instrumentation and Methods for Astrophysics · Physics 2019-01-01 Shay Zucker

We study random points on the real line generated by the eigenvalues in unitary invariant random matrix ensembles or by more general repulsive particle systems. As the number of points tends to infinity, we prove convergence of the…

Probability · Mathematics 2015-11-11 Kristina Schubert , Martin Venker

We study the problem of clustering with relative constraints, where each constraint specifies relative similarities among instances. In particular, each constraint $(x_i, x_j, x_k)$ is acquired by posing a query: is instance $x_i$ more…

Machine Learning · Computer Science 2015-01-05 Yuanli Pei , Xiaoli Z. Fern , Rómer Rosales , Teresa Vania Tjahja

We derive asymptotic expansions up to order $n^{-1/2}$ for the nonnull distribution functions of the likelihood ratio, Wald, score and gradient test statistics in the class of dispersion models, under a sequence of Pitman alternatives. The…

Statistics Theory · Mathematics 2011-02-23 Artur J. Lemonte , Silvia L. P. Ferrari

In condensed-matter, level statistics has long been used to characterize the phases of a disordered system. We provide evidence within the context of a simple model that in a disordered large-N gauge theory with a gravity dual, there exist…

High Energy Physics - Theory · Physics 2012-06-12 Omid Saremi

Compression-based similarity measures are effectively employed in applications on diverse data types with a basically parameter-free approach. Nevertheless, there are problems in applying these techniques to medium-to-large datasets which…

Machine Learning · Statistics 2012-10-03 Daniele Cerra , Mihai Datcu

Nonparametric tests for equality of multivariate distributions are frequently desired in research. It is commonly required that test-procedures based on relatively small samples of vectors accurately control the corresponding Type I Error…

Methodology · Statistics 2021-01-14 Ablert Vexler , Gregory Gurevich , Li Zou

When we represent a network of sensors in Euclidean space by a graph, there are two distances between any two nodes that we may consider. One of them is the Euclidean distance. The other is the distance between the two nodes in the graph,…

Networking and Internet Architecture · Computer Science 2009-06-10 Rodrigo S. C. Leao , Valmir C. Barbosa

We treat the problem of testing independence between m continuous variables when m can be larger than the available sample size n. We consider three types of test statistics that are constructed as sums or sums of squares of pairwise rank…

Statistics Theory · Mathematics 2016-12-05 Dennis Leung , Mathias Drton

In this paper, we present a novel way to summarize the structure of large graphs, based on non-parametric estimation of edge density in directed multigraphs. Following coclustering approach, we use a clustering of the vertices, with a…

Social and Information Networks · Computer Science 2015-08-07 Marc Boullé