中文
相关论文

相关论文: Multi-Hop Network with Multiple Decision Centers u…

200 篇论文

This paper investigates the problem of distributed stochastic approximation in multi-agent systems. The algorithm under study consists of two steps: a local stochastic approximation step and a diffusion step which drives the network to a…

多智能体系统 · 计算机科学 2014-10-28 Gemma Morral , Pascal Bianchi , Gersende Fort

We study model-based reinforcement learning (RL) for episodic Markov decision processes (MDP) whose transition probability is parametrized by an unknown transition core with features of state and action. Despite much recent progress in…

机器学习 · 统计学 2024-11-19 Taehyun Hwang , Min-hwan Oh

This paper proposes a distributed version of Determinant Point Processing (DPP) inference to enhance multi-source data diversification under limited communication bandwidth. DPP is a popular probabilistic approach that improves data…

机器学习 · 计算机科学 2023-11-21 Xiwen Chen , Huayu Li , Rahul Amin , Abolfazl Razi

In this paper, we consider a network of processors aiming at cooperatively solving mixed-integer convex programs subject to uncertainty. Each node only knows a common cost function and its local uncertain constraint set. We propose a…

最优化与控制 · 数学 2022-07-19 Mohammadreza Chamanbaz , Giuseppe Notarstefano , Francesco Sasso , Roland Bouffanais

Meta reinforcement learning sets a distribution over a set of tasks on which the agent can train at will, then is asked to learn an optimal policy for any test task efficiently. In this paper, we consider a finite set of tasks modeled…

机器学习 · 计算机科学 2024-06-05 Mirco Mutti , Aviv Tamar

Increasing integration of renewable generation poses significant challenges to ensure robustness guarantees in real-time energy system decision-making. This work aims to develop a robust optimal transmission switching (OTS) framework that…

最优化与控制 · 数学 2022-09-01 Yuqi Zhou , Hao Zhu , Grani A. Hanasusanto

This paper studies the problem of sequential Gaussian shift-in-mean hypothesis testing in a distributed multi-agent network. A sequential probability ratio test (SPRT) type algorithm in a distributed framework of the…

最优化与控制 · 数学 2015-09-02 Anit Kumar Sahu , Soummya Kar

In confirmatory clinical trials with small sample sizes, hypothesis tests based on asymptotic distributions are often not valid and exact non-parametric procedures are applied instead. However, the latter are based on discrete test…

统计方法学 · 统计学 2018-02-22 Robin Ristl , Dong Xi , Ekkehard Glimm , Martin Posch

We consider the classical sequential binary hypothesis testing problem in which there are two hypotheses governed respectively by distributions $P_0$ and $P_1$ and we would like to decide which hypothesis is true using a sequential test. It…

信息论 · 计算机科学 2020-07-01 Yonglong Li , Vincent Y. F. Tan

We establish a collection of closed-loop guarantees and propose a scalable optimization algorithm for distributionally robust model predictive control (DRMPC) applied to linear systems, convex constraints, and quadratic costs. Via standard…

最优化与控制 · 数学 2024-11-13 Robert D. McAllister , Peyman Mohajerin Esfahani

Markov chains are the de facto finite-state model for stochastic dynamical systems, and Markov decision processes (MDPs) extend Markov chains by incorporating non-deterministic behaviors. Given an MDP and rewards on states, a classical…

计算机科学中的逻辑 · 计算机科学 2024-11-13 Krishnendu Chatterjee , Laurent Doyen

This study demonstrates that double descent can be mitigated by adding a dropout layer adjacent to the fully connected linear layer. The unexpected double-descent phenomenon garnered substantial attention in recent years, resulting in…

机器学习 · 计算机科学 2025-08-08 Tian-Le Yang , Joe Suzuki

This paper studies the estimation of low-rank Markov chains from empirical trajectories. We propose a non-convex estimator based on rank-constrained likelihood maximization. Statistical upper bounds are provided for the Kullback-Leiber…

机器学习 · 统计学 2018-07-20 Xudong Li , Mengdi Wang , Anru Zhang

Let $A$ be a transition probability kernel on a finite state space $\Delta^o =\{1, \ldots , d\}$ such that $A(x,y)>0$ for all $x,y \in \Delta^o$. Consider a reinforced chain given as a sequence $\{X_n, \; n \in \mathbb{N}_0\}$ of…

概率论 · 数学 2022-05-20 Amarjit Budhiraja , Adam Waterbury

We study the composite sequential quantum hypothesis testing (SQHT) problem, where the objective is to distinguish a null quantum state from a set of alternative quantum states. We propose a mixture-sequential quantum probability ratio test…

量子物理 · 物理学 2026-05-12 Jacob Paul Simpson , Efstratios Palias , Sharu Theresa Jose

We consider the problem of learning to behave optimally in a Markov Decision Process when a reward function is not specified, but instead we have access to a set of demonstrators of varying performance. We assume the demonstrators are…

机器学习 · 计算机科学 2019-08-01 Pablo Samuel Castro , Shijian Li , Daqing Zhang

In a recent study, (Jain et al 2007 Phys. Rev. Lett. 99 190601), a symmetric exclusion process with time-dependent hopping rates was introduced. Using simulations and a perturbation theory, it was shown that if the hopping rates at two…

统计力学 · 物理学 2008-11-17 Rahul Marathe , Kavita Jain , Abhishek Dhar

The maximum type-I and type-II error exponents associated with the newly introduced almost-fixed-length hypothesis testing is characterized. In this class of tests, the decision-maker declares the true hypothesis almost always after…

信息论 · 计算机科学 2016-05-18 Anusha Lalitha , Tara Javidi

We consider a linear multi-hop network composed of multi-state discrete-time memoryless channels over each hop, with orthogonal time-sharing across hops under a half-duplex relaying protocol. We analyze the probability of error and…

信息论 · 计算机科学 2008-10-29 Ozgur Oyman

Motivated by differential co-expression analysis in genomics, we consider in this paper estimation and testing of high-dimensional differential correlation matrices. An adaptive thresholding procedure is introduced and theoretical…

统计方法学 · 统计学 2015-10-22 T. Tony Cai , Anru Zhang