English
Related papers

Related papers: Reading policies for joins: An asymptotic analysis

200 papers

While many multiagent algorithms are designed for homogeneous systems (i.e. all agents are identical), there are important applications which require an agent to coordinate its actions without knowing a priori how the other agents behave.…

Artificial Intelligence · Computer Science 2019-07-17 Stefano V. Albrecht , Subramanian Ramamoorthy

Let $\{X_i\}$ be a sequence of independent identically distributed random variables with an intermediate regularly varying (IR) right tail $\bar{F}$. Let $(N, C_1, ..., C_N)$ be a nonnegative random vector independent of the $\{X_i\}$ with…

Probability · Mathematics 2012-04-18 Mariana Olvera-Cravioto

Adaptive experiments such as multi-arm bandits adapt the treatment-allocation policy and/or the decision to stop the experiment to the data observed so far. This has the potential to improve outcomes for study participants within the…

Methodology · Statistics 2024-05-03 Aurélien Bibaut , Nathan Kallus

We report on work towards flexible algorithms for solving decision problems represented as influence diagrams. An algorithm is given to construct a tree structure for each decision node in an influence diagram. Each tree represents a…

Artificial Intelligence · Computer Science 2013-02-18 Michael C. Horsch , David L. Poole

In a model of network communication based on a random walk in an undirected graph, what subset of nodes (subject to constraints on the set size), enables the fastest spread of information? In this paper, we assume the dynamics of spread is…

Discrete Mathematics · Computer Science 2017-04-11 F. Y. Hunt

We investigate the problem of active learning on a given tree whose nodes are assigned binary labels in an adversarial way. Inspired by recent results by Guillory and Bilmes, we characterize (up to constant factors) the optimal placement of…

Machine Learning · Computer Science 2013-01-23 Nicolo Cesa-Bianchi , Claudio Gentile , Fabio Vitale , Giovanni Zappella

In collaborative learning with streaming data, nodes (e.g., organizations) jointly and continuously learn a machine learning (ML) model by sharing the latest model updates computed from their latest streaming data. For the more resourceful…

Machine Learning · Computer Science 2023-06-12 Xiaoqiang Lin , Xinyi Xu , See-Kiong Ng , Chuan-Sheng Foo , Bryan Kian Hsiang Low

Policy gradient methods are among the most effective methods in challenging reinforcement learning problems with large state and/or action spaces. However, little is known about even their most basic theoretical convergence properties,…

Machine Learning · Computer Science 2020-10-16 Alekh Agarwal , Sham M. Kakade , Jason D. Lee , Gaurav Mahajan

We consider goodness-of-fit tests for uniformity of a multinomial distribution by means of tests based on a class of symmetric statistics, defined as the sum of some function of cell-frequencies. We are dealing with an asymptotic regime,…

Statistics Theory · Mathematics 2022-11-03 Sherzod M Mirakhmedov

A multi-agent system operates in an uncertain environment about which agents have different and time varying beliefs that, as time progresses, converge to a common belief. A global utility function that depends on the realized state of the…

Computer Science and Game Theory · Computer Science 2016-02-08 Ceyhun Eksin , Alejandro Ribeiro

In this paper, we study the large $n$ asymptotics of the expected maximum of an $n$-step random walk/L\'evy flight (characterized by a L\'evy index $1<\mu\leq 2$) on a line, in the presence of a constant drift $c$. For $0<\mu\leq 1$, the…

Statistical Mechanics · Physics 2018-09-03 Philippe Mounaix , Satya N. Majumdar , Gregory Schehr

We consider the infinite-horizon, average-reward restless bandit problem in discrete time. We propose a new class of policies that are designed to drive a progressively larger subset of arms toward the optimal distribution. We show that our…

Machine Learning · Computer Science 2026-03-31 Yige Hong , Qiaomin Xie , Yudong Chen , Weina Wang

Graph data are pervasive in many real-world applications. Recently, increasing attention has been paid on graph neural networks (GNNs), which aim to model the local graph structures and capture the hierarchical patterns by aggregating the…

Machine Learning · Computer Science 2020-06-29 Kwei-Herng Lai , Daochen Zha , Kaixiong Zhou , Xia Hu

We consider sequential selection of an alternating subsequence from a sequence of independent, identically distributed, continuous random variables, and we determine the exact asymptotic behavior of an optimal sequentially selected…

Probability · Mathematics 2011-08-15 Alessandro Arlotto , Robert W. Chen , Lawrence A. Shepp , J. Michael Steele

The optimal policy of a reinforcement learning problem is often discontinuous and non-smooth. I.e., for two states with similar representations, their optimal policies can be significantly different. In this case, representing the entire…

Machine Learning · Computer Science 2020-02-10 Zhimin Hou , Kuangen Zhang , Yi Wan , Dongyu Li , Chenglong Fu , Haoyong Yu

We study optimal policy learning under combined budget and minimum coverage constraints. We show that the problem admits a knapsack-type structure and that the optimal policy can be characterized by an affine threshold rule involving both…

Machine Learning · Statistics 2026-05-13 Giovanni Cerulli

This paper investigates the asymptotic behaviour of solutions to certain infinite systems of coupled recurrence relations. In particular, we obtain a characterisation of those initial values which lead to a convergent solution, and for…

Functional Analysis · Mathematics 2019-02-14 L. Paunonen , D. Seifert

We study the selection of covariate adjustment sets for estimating the value of point exposure dynamic policies, also known as dynamic treatment regimes, assuming a non-parametric causal graphical model with hidden variables, in which at…

Statistics Theory · Mathematics 2020-05-27 Ezequiel Smucler , Facundo Sapienza , Andrea Rotnitzky

Controllable Markov chains describe the dynamics of sequential decision making tasks and are the central component in optimal control and reinforcement learning. In this work, we give the general form of an optimal policy for learning…

Machine Learning · Computer Science 2025-12-24 Peter N. Loxley

Given a countably infinite group $G$ acting on some space $X$, an increasing family of finite subsets $G_n$ and $x\in X$, a natural question to ask is what asymptotical distribution the sets $G_nx$ form. More formally, we define for a…

Dynamical Systems · Mathematics 2020-09-23 Uriya Pumerantz
‹ Prev 1 8 9 10 Next ›