English
Related papers

Related papers: Finding Non-Redundant Simpson's Paradox from Multi…

200 papers

A fundamental problem of statistical data analysis, distribution density estimation by experimental data, is considered. A new method with optimal asymptotic behavior, the root density estimator, is developed. The method proposed may be…

Data Analysis, Statistics and Probability · Physics 2007-05-23 Yu. I. Bogdanov

Nonuniform subsampling methods are effective to reduce computational burden and maintain estimation efficiency for massive data. Existing methods mostly focus on subsampling with replacement due to its high computational efficiency. If the…

Methodology · Statistics 2021-07-06 Jun Yu , HaiYing Wang , Mingyao Ai , Huiming Zhang

Addressing health disparities among different demographic groups is a key challenge in public health. Despite many efforts, there is still a gap in understanding how these disparities unfold over time. Our paper focuses on this overlooked…

Applications · Statistics 2024-04-19 Sang Kyu Lee , Seonjin Kim , Mi-Ok Kim , Katherine L. Grantz , Hyokyoung G. Hong

We give a method for proactively identifying small, plausible shifts in distribution which lead to large differences in model performance. These shifts are defined via parametric changes in the causal mechanisms of observed variables, where…

Machine Learning · Computer Science 2023-01-18 Nikolaj Thams , Michael Oberst , David Sontag

Statistical inference problems arising within signal processing, data mining, and machine learning naturally give rise to hard combinatorial optimization problems. These problems become intractable when the dimensionality of the data is…

Statistical Mechanics · Physics 2017-04-27 Adel Javanmard , Andrea Montanari , Federico Ricci-Tersenghi

Subsampling is a general statistical method developed in the 1990s aimed at estimating the sampling distribution of a statistic $\hat \theta _n$ in order to conduct nonparametric inference such as the construction of confidence intervals…

Statistics Theory · Mathematics 2021-12-14 Dimitris N. Politis

In statistical inference, retrodiction is the act of inferring potential causes in the past based on knowledge of the effects in the present and the dynamics leading to the present. Retrodiction is applicable even when the dynamics is not…

Category Theory · Mathematics 2024-02-01 Arthur J. Parzygnat

Nonlocality is a quintessential signature of nonclassical behaviour and a resource for quantum advantages in communication and computation. The paradoxical correlations witnessed by strong nonlocality undergird the standard probabilistic…

Quantum Physics · Physics 2025-08-21 Nadish de Silva , Santanil Jana , Ming Yin

Consider a set of order statistics that arise from sorting samples from two different populations, each with their own, possibly different distribution function. The probability that these order statistics fall in disjoint, ordered…

Computation · Statistics 2007-06-26 Deborah H. Glueck , Anis Karimpour-Fard , Jan Mandel , Keith E. Muller

Many ecological studies and conservation policies are based on field observations of species, which can be affected by systematic variability introduced by the observation process. A recently introduced causal modeling technique called…

Methodology · Statistics 2021-01-05 Shiv Shankar , Daniel Sheldon , Tao Sun , John Pickering , Thomas G. Dietterich

The problem of sequential change diagnosis is considered, where observations are obtained on-line, an abrupt change occurs in their distribution, and the goal is to quickly detect the change and accurately identify the post-change…

Statistics Theory · Mathematics 2022-11-24 Austin Warner , Georgios Fellouris

One of the most important empirical findings in microeconometrics is the pervasiveness of heterogeneity in economic behaviour (cf. Heckman 2001). This paper shows that cumulative distribution functions and quantiles of the nonparametric…

Econometrics · Economics 2020-05-19 Juan Carlos Escanciano

The friendship paradox is the observation that the degrees of the neighbors of a node in any network will, on average, be greater than the degree of the node itself. In common parlance, your friends have more friends than you do. In this…

Social and Information Networks · Computer Science 2021-10-26 George T. Cantwell , Alec Kirkley , M. E. J. Newman

A common limitation of diagnostic tests for detecting social biases in NLP models is that they may only detect stereotypic associations that are pre-specified by the designer of the test. Since enumerating all possible problematic…

Computation and Language · Computer Science 2023-02-17 Haozhe An , Zongxia Li , Jieyu Zhao , Rachel Rudinger

Finding the most similar subsequences between two multidimensional time series has many applications: e.g. capturing dependency in stock market or discovering coordinated movement of baboons. Considering one pattern occurring in one time…

Machine Learning · Computer Science 2025-05-19 Thanadej Rattanakornphan , Piyanon Charoenpoonpanich , Chainarong Amornbunchornvej

The Gibbs Paradox is essentially a set of open questions as to how sameness of gases or fluids (or masses, more generally) are to be treated in thermodynamics and statistical mechanics. They have a variety of answers, some restricted to…

History and Philosophy of Physics · Physics 2018-08-07 Simon Saunders

Consider the problem where a statistician in a two-node system receives rate-limited information from a transmitter about marginal observations of a memoryless process generated from two possible distributions. Using its own observations,…

Information Theory · Computer Science 2017-03-02 Gil Katz , Pablo Piantanida , Mérouane Debbah

With the emergence of graph databases, the task of frequent subgraph discovery has been extensively addressed. Although the proposed approaches in the literature have made this task feasible, the number of discovered frequent subgraphs is…

Databases · Computer Science 2013-08-16 Wajdi Dhifli , Mohamed Moussaoui , Rabie Saidi , Engelbert Mephu Nguifo

Semi-Poisson statistics are shown to be obtained by removing every other number from a random sequence. Retaining every (r+1)th level we obtain a family of secuences which we call daisy models. Their statistical properties coincide with…

chao-dyn · Physics 2009-10-31 H. Hernandez-Saldaña , J. Flores , T. H. Seligman

Modern biomedical, behavioral and psychological inference about cause-effect relationships respects an ergodic assumption, that is, that mean response of representative samples allow predictions about individual members of those samples.…

Neurons and Cognition · Quantitative Biology 2021-05-31 Madhur Mangalam , Damian G. Kelty-Stephen
‹ Prev 1 8 9 10 Next ›