中文
相关论文

相关论文: Reading policies for joins: An asymptotic analysis

200 篇论文

While many multiagent algorithms are designed for homogeneous systems (i.e. all agents are identical), there are important applications which require an agent to coordinate its actions without knowing a priori how the other agents behave.…

人工智能 · 计算机科学 2019-07-17 Stefano V. Albrecht , Subramanian Ramamoorthy

Let $\{X_i\}$ be a sequence of independent identically distributed random variables with an intermediate regularly varying (IR) right tail $\bar{F}$. Let $(N, C_1, ..., C_N)$ be a nonnegative random vector independent of the $\{X_i\}$ with…

概率论 · 数学 2012-04-18 Mariana Olvera-Cravioto

Adaptive experiments such as multi-arm bandits adapt the treatment-allocation policy and/or the decision to stop the experiment to the data observed so far. This has the potential to improve outcomes for study participants within the…

统计方法学 · 统计学 2024-05-03 Aurélien Bibaut , Nathan Kallus

We report on work towards flexible algorithms for solving decision problems represented as influence diagrams. An algorithm is given to construct a tree structure for each decision node in an influence diagram. Each tree represents a…

人工智能 · 计算机科学 2013-02-18 Michael C. Horsch , David L. Poole

In a model of network communication based on a random walk in an undirected graph, what subset of nodes (subject to constraints on the set size), enables the fastest spread of information? In this paper, we assume the dynamics of spread is…

离散数学 · 计算机科学 2017-04-11 F. Y. Hunt

We investigate the problem of active learning on a given tree whose nodes are assigned binary labels in an adversarial way. Inspired by recent results by Guillory and Bilmes, we characterize (up to constant factors) the optimal placement of…

机器学习 · 计算机科学 2013-01-23 Nicolo Cesa-Bianchi , Claudio Gentile , Fabio Vitale , Giovanni Zappella

In collaborative learning with streaming data, nodes (e.g., organizations) jointly and continuously learn a machine learning (ML) model by sharing the latest model updates computed from their latest streaming data. For the more resourceful…

机器学习 · 计算机科学 2023-06-12 Xiaoqiang Lin , Xinyi Xu , See-Kiong Ng , Chuan-Sheng Foo , Bryan Kian Hsiang Low

Policy gradient methods are among the most effective methods in challenging reinforcement learning problems with large state and/or action spaces. However, little is known about even their most basic theoretical convergence properties,…

机器学习 · 计算机科学 2020-10-16 Alekh Agarwal , Sham M. Kakade , Jason D. Lee , Gaurav Mahajan

We consider goodness-of-fit tests for uniformity of a multinomial distribution by means of tests based on a class of symmetric statistics, defined as the sum of some function of cell-frequencies. We are dealing with an asymptotic regime,…

统计理论 · 数学 2022-11-03 Sherzod M Mirakhmedov

A multi-agent system operates in an uncertain environment about which agents have different and time varying beliefs that, as time progresses, converge to a common belief. A global utility function that depends on the realized state of the…

计算机科学与博弈论 · 计算机科学 2016-02-08 Ceyhun Eksin , Alejandro Ribeiro

In this paper, we study the large $n$ asymptotics of the expected maximum of an $n$-step random walk/L\'evy flight (characterized by a L\'evy index $1<\mu\leq 2$) on a line, in the presence of a constant drift $c$. For $0<\mu\leq 1$, the…

统计力学 · 物理学 2018-09-03 Philippe Mounaix , Satya N. Majumdar , Gregory Schehr

We consider the infinite-horizon, average-reward restless bandit problem in discrete time. We propose a new class of policies that are designed to drive a progressively larger subset of arms toward the optimal distribution. We show that our…

机器学习 · 计算机科学 2026-03-31 Yige Hong , Qiaomin Xie , Yudong Chen , Weina Wang

Graph data are pervasive in many real-world applications. Recently, increasing attention has been paid on graph neural networks (GNNs), which aim to model the local graph structures and capture the hierarchical patterns by aggregating the…

机器学习 · 计算机科学 2020-06-29 Kwei-Herng Lai , Daochen Zha , Kaixiong Zhou , Xia Hu

We consider sequential selection of an alternating subsequence from a sequence of independent, identically distributed, continuous random variables, and we determine the exact asymptotic behavior of an optimal sequentially selected…

The optimal policy of a reinforcement learning problem is often discontinuous and non-smooth. I.e., for two states with similar representations, their optimal policies can be significantly different. In this case, representing the entire…

机器学习 · 计算机科学 2020-02-10 Zhimin Hou , Kuangen Zhang , Yi Wan , Dongyu Li , Chenglong Fu , Haoyong Yu

We study optimal policy learning under combined budget and minimum coverage constraints. We show that the problem admits a knapsack-type structure and that the optimal policy can be characterized by an affine threshold rule involving both…

机器学习 · 统计学 2026-05-13 Giovanni Cerulli

This paper investigates the asymptotic behaviour of solutions to certain infinite systems of coupled recurrence relations. In particular, we obtain a characterisation of those initial values which lead to a convergent solution, and for…

泛函分析 · 数学 2019-02-14 L. Paunonen , D. Seifert

We study the selection of covariate adjustment sets for estimating the value of point exposure dynamic policies, also known as dynamic treatment regimes, assuming a non-parametric causal graphical model with hidden variables, in which at…

统计理论 · 数学 2020-05-27 Ezequiel Smucler , Facundo Sapienza , Andrea Rotnitzky

Controllable Markov chains describe the dynamics of sequential decision making tasks and are the central component in optimal control and reinforcement learning. In this work, we give the general form of an optimal policy for learning…

机器学习 · 计算机科学 2025-12-24 Peter N. Loxley

Given a countably infinite group $G$ acting on some space $X$, an increasing family of finite subsets $G_n$ and $x\in X$, a natural question to ask is what asymptotical distribution the sets $G_nx$ form. More formally, we define for a…

动力系统 · 数学 2020-09-23 Uriya Pumerantz
‹ 上一页 1 8 9 10 下一页 ›