English
Related papers

Related papers: The Top-K Tau-Path Screen for Monotone Association

200 papers

We present the symmetric thermal optimal path (TOPS) method to determine the time-dependent lead-lag relationship between two stochastic time series. This novel version of the previously introduced TOP method alleviates some inconsistencies…

Statistical Finance · Quantitative Finance 2018-02-27 Hao Meng , Hai-Chuan Xu , Wei-Xing Zhou , Didier Sornette

In shape-constrained nonparametric inference, it is often necessary to perform preliminary tests to verify whether a probability mass function (p.m.f.) satisfies qualitative constraints such as monotonicity, convexity, or in general…

Statistics Theory · Mathematics 2025-12-23 Fadoua Balabdaoui , Antonio Di Noia

MixUp is a data augmentation strategy where additional samples are generated during training by combining random pairs of training samples and their labels. However, selecting random pairs is not potentially an optimal choice. In this work,…

Computation and Language · Computer Science 2022-05-09 Seo Yeon Park , Cornelia Caragea

Often, it is required to estimate the probability that a quantity such as toxicity level, plutonium, temperature, rainfall, damage, wind speed, wave size, earthquake magnitude, risk, etc., exceeds an unsafe high threshold. The probability…

Methodology · Statistics 2019-07-24 Benjamin Kedem , Lemeng Pan , Paul Smith , Chen Wang

Graph association rule mining is a data mining technique used for discovering regularities in graph data. In this study, we propose a novel concept, {\it path association rule mining}, to discover the correlations of path patterns that…

Databases · Computer Science 2022-10-25 Yuya Sasaki

The adaptive voter model allows for studying the interplay between homophily, the tendency of like-minded individuals to attract each other, and social influence, the tendency for connected individuals to influence each other. However, it…

Physics and Society · Physics 2023-03-21 Nikos Papanikolaou , Renaud Lambiotte , Giacomo Vaccario

We develop a unified mathematical framework for certified Top-$k$ attention truncation that quantifies approximation error at both the distribution and output levels. For a single attention distribution $P$ and its Top-$k$ truncation $\hat…

Machine Learning · Computer Science 2025-12-09 Georgios Tzachristas , Lei Deng , Ioannis Tzachristas , Gong Zhang , Renhai Chen

We consider the problem of similarity search within a set of top-k lists under the Kendall's Tau distance function. This distance describes how related two rankings are in terms of concordantly and discordantly ordered items. As top-k lists…

Databases · Computer Science 2014-09-03 Koninika Pal , Sebastian Michel

The comparison of different medical treatments from observational studies or across different clinical studies is often biased by confounding factors such as systematic differences in patient demographics or in the inclusion criteria for…

Methodology · Statistics 2025-05-15 Ekkehard Glimm , Lillian Yau

Generating paired sequences with maximal compatibility from a given set is one of the most important challenges in various applications, including information and communication technologies. However, the number of possible pairings explodes…

Data Structures and Algorithms · Computer Science 2022-05-10 Naoki Fujita , Nicolas Chauvet , Andre Roehm , Ryoichi Horisaki , Aohan Li , Mikio Hasegawa , Makoto Naruse

Graph based entropy, an index of the diversity of events in their distribution to parts of a co-occurrence graph, is proposed for detecting signs of structural changes in the data that are informative in explaining latent dynamics of…

Social and Information Networks · Computer Science 2019-05-03 Yukio Ohsawa

In this paper, we propose a novel approach that employs kinetic equations to describe the collective dynamics emerging from graph-mediated pairwise interactions in multi-agent systems. We formally show that for large graphs and specific…

Physics and Society · Physics 2026-05-15 Marco Nurisso , Matteo Raviola , Andrea Tosin

The paper considers the problem of finding the number of dominant voters in two-level voting procedures. At the first stage, voting is conducted among local groups of voters, and at the second stage, the results are aggregated to form a…

Discrete Mathematics · Computer Science 2025-06-11 N. I. Shushko , D. V. Lemtyuzhnikova

High-dimensional k-sample comparison is a common applied problem. We construct a class of easy-to-implement nonparametric distribution-free tests based on new tools and unexplored connections with spectral graph theory. The test is shown to…

Methodology · Statistics 2019-08-12 Subhadeep , Mukhopadhyay , Kaijun Wang

We consider the problem of determining the top-$k$ largest measurements from a dataset distributed among a network of $n$ agents with noisy communication links. We show that this scenario can be cast as a distributed convex optimization…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-12-02 Xu Zhang , Marcos Vasconcelos

Community detection is a fundamental problem in network analysis which is made more challenging by overlaps between communities which often occur in practice. Here we propose a general, flexible, and interpretable generative model for…

Machine Learning · Statistics 2015-03-16 Yuan Zhang , Elizaveta Levina , Ji Zhu

Community detection is a central task in graph analytics. Given the substantial growth in graph size, scalability in community detection continues to be an unresolved challenge. Recently, alongside established methods like Louvain and…

Social and Information Networks · Computer Science 2024-12-18 Tianyi Chen , Charalampos E. Tsourakakis

We propose the Heterogeneous Thurstone Model (HTM) for aggregating ranked data, which can take the accuracy levels of different users into account. By allowing different noise distributions, the proposed HTM model maintains the generality…

Machine Learning · Computer Science 2019-12-04 Tao Jin , Pan Xu , Quanquan Gu , Farzad Farnoud

Introduction The tau statistic is a recent second-order correlation function that can assess the magnitude and range of global spatiotemporal clustering from epidemiological data containing geolocations of individual cases and, usually,…

The paper investigates the problem of performing correlation analysis when the number of observations is very large. In such a case, it is often necessary to combine the random observations to achieve dimensionality reduction of the…

Information Theory · Computer Science 2020-10-19 Pavel Loskot
‹ Prev 1 4 5 6 7 8 10 Next ›