English
Related papers

Related papers: Fast Simulation of Size-Constrained Multitype Bien…

200 papers

We introduce a random finite rooted tree $\mathcal{C}$, the steady state cluster, characterized by a recursive description: $\mathcal{C}$ is a singleton with probability $1/2$ and otherwise is obtained by joining by an edge the roots of two…

Probability · Mathematics 2018-09-11 Edward Crane

We present improved learning-augmented algorithms for finding an approximate minimum spanning tree (MST) for points in an arbitrary metric space. Our work follows a recent framework called metric forest completion (MFC), where the learned…

Data Structures and Algorithms · Computer Science 2026-03-02 Nate Veldt , Thomas Stanley , Benjamin W. Priest , Trevor Steil , Keita Iwabuchi , T. S. Jayram , Grace J. Li , Geoffrey Sanders

Random forests are popular methods for regression and classification analysis, and many different variants have been proposed in recent years. One interesting example is the Mondrian random forest, in which the underlying constituent trees…

Statistics Theory · Mathematics 2025-11-10 Matias D. Cattaneo , Jason M. Klusowski , William G. Underwood

Gaussian Graphical Models (GGMs) or Gauss Markov random fields are widely used in many applications, and the trade-off between the modeling capacity and the efficiency of learning and inference has been an important research problem. In…

Machine Learning · Computer Science 2013-11-12 Ying Liu , Alan S. Willsky

Gradient boosted decision trees (GBDTs) are widely used in machine learning, and the output of current GBDT implementations is a single variable. When there are multiple outputs, GBDT constructs multiple trees corresponding to the output…

Computer Vision and Pattern Recognition · Computer Science 2020-01-01 Zhendong Zhang , Cheolkon Jung

In this paper, we modify the proof methods of some previously weakly consistent variants of random forests into strongly consistent proof methods, and improve the data utilization of these variants in order to obtain better theoretical…

Machine Learning · Computer Science 2023-10-17 JunHao Chen

In this paper, we study the fundamental statistical efficiency of Reinforcement Learning in Mean-Field Control (MFC) and Mean-Field Game (MFG) with general model-based function approximation. We introduce a new concept called Mean-Field…

Machine Learning · Computer Science 2024-10-04 Jiawei Huang , Batuhan Yardim , Niao He

In this article, we focus on Bienaym\'e-Galton-Watson processes with linear-fractional offspring distributions. At a fixed generation, we consider a sample of the individuals alive, drawn in two different ways: either through Bernoulli…

Probability · Mathematics 2025-06-24 Natalia Cardona-Tobón , Sandra Palau

We show that large critical multi-type Galton-Watson trees, when conditioned to be large, converge locally in distribution to an infinite tree which is analoguous to Kesten's infinite monotype Galton-Watson tree. This is proven when we…

Probability · Mathematics 2016-08-02 Robin Stephenson

Big Data is one of the major challenges of statistical science and has numerous consequences from algorithmic and theoretical viewpoints. Big Data always involve massive data but they also often include online data and data heterogeneity.…

Machine Learning · Statistics 2017-03-23 Robin Genuer , Jean-Michel Poggi , Christine Tuleau-Malot , Nathalie Villa-Vialaneix

The problem of learning forest-structured discrete graphical models from i.i.d. samples is considered. An algorithm based on pruning of the Chow-Liu tree through adaptive thresholding is proposed. It is shown that this algorithm is both…

Information Theory · Computer Science 2011-02-15 Vincent Y. F. Tan , Animashree Anandkumar , Alan S. Willsky

Mixed-effects models are among the most commonly used statistical methods for the exploration of multispecies data. In recent years, also Joint Species Distribution Models and Generalized Linear Latent Variale Models have gained in…

Computation · Statistics 2025-01-31 Bert van der Veen , Robert Brian O'Hara

We address the problem of finding influential training samples for a particular case of tree ensemble-based models, e.g., Random Forest (RF) or Gradient Boosted Decision Trees (GBDT). A natural way of formalizing this problem is studying…

Machine Learning · Computer Science 2018-03-14 Boris Sharchilev , Yury Ustinovsky , Pavel Serdyukov , Maarten de Rijke

We observe $n$ sequences at each of $m$ sites, and assume that they have evolved from an ancestral sequence that forms the root of a binary tree of known topology and branch lengths, but the sequence states at internal nodes are unknown.…

Computation · Statistics 2014-08-28 Adam Persing , Ajay Jasra , Alexandros Beskos , David Balding , Maria De Iorio

We consider the problem of collaborative tree exploration posed by Fraigniaud, Gasieniec, Kowalski, and Pelc where a team of $k$ agents is tasked to collectively go through all the edges of an unknown tree as fast as possible. Denoting by…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-02-01 Romain Cosson , Laurent Massoulié , Laurent Viennot

We discuss various forms of convergence of the vicinity of a uniformly at random selected vertex in random simply generated trees, as the size tends to infinity. For the standard case of a critical Galton-Watson tree conditioned to be large…

Probability · Mathematics 2018-02-09 Benedikt Stufler

This paper presents a computational framework for the Wasserstein auto-encoding of merge trees (MT-WAE), a novel extension of the classical auto-encoder neural network architecture to the Wasserstein metric space of merge trees. In contrast…

Machine Learning · Computer Science 2023-11-13 Mahieu Pont , Julien Tierny

We develop a rigorous and implementable framework for Gibbs sampling of infinite-dimensional quantum systems governed by unbounded Hamiltonians. Extending dissipative Gibbs samplers beyond finite dimensions raises fundamental obstacles,…

Quantum Physics · Physics 2026-04-02 Simon Becker , Cambyse Rouzé , Robert Salzmann

We consider the problem of learning an $\varepsilon$-optimal policy in a general class of continuous-space Markov decision processes (MDPs) having smooth Bellman operators. Given access to a generative model, we achieve rate-optimal sample…

Machine Learning · Computer Science 2024-05-13 Davide Maran , Alberto Maria Metelli , Matteo Papini , Marcello Restelli

We give an invariance principle for very general additive functionals of conditioned Bienaym{\'e}-Galton-Watson trees in the global regime when the offspring distribution lies in the domain of attraction of a stable distribution, the limit…

Probability · Mathematics 2020-09-18 Romain Abraham , Jean-François Delmas , Michel Nassif