中文
相关论文

相关论文: Real-time Remote Reconstruction of a Markov Source…

200 篇论文

Remote tracking systems play a critical role in applications such as IoT, monitoring, surveillance and healthcare. In such systems, maintaining both real-time state awareness (for online decision making) and accurate reconstruction of…

系统与控制 · 电气工程与系统科学 2025-05-20 Sunjung Kang , Vishrant Tripathi , Christopher G. Brinton

In this paper, we study a remote monitoring system where a receiver observes a remote binary Markov source and decides whether to sample and transmit the state through a randomly delayed channel. We adopt uncertainty of information (UoI),…

信息论 · 计算机科学 2024-05-20 Xiaomeng Chen , Aimin Li , Shaohua Wu

We study the offline data-driven sequential decision making problem in the framework of Markov decision process (MDP). In order to enhance the generalizability and adaptivity of the learned policy, we propose to evaluate each policy by a…

统计理论 · 数学 2021-11-11 Zhengling Qi , Peng Liao

The goal of this paper is to analyze distributional Markov Decision Processes as a class of control problems in which the objective is to learn policies that steer the distribution of a cumulative reward toward a prescribed target law,…

最优化与控制 · 数学 2026-02-09 Nicole Bäuerle , Athanasios Vasileiadis

This paper investigates MDPs with intermittent state information. We consider a scenario where the controller perceives the state information of the process via an unreliable communication channel. The transmissions of state information…

人工智能 · 计算机科学 2025-02-17 Gongpu Chen , Soung-Chang Liew

In many Cyber-Physical Systems, we encounter the problem of remote state estimation of geographically distributed and remote physical processes. This paper studies the scheduling of sensor transmissions to estimate the states of multiple…

系统与控制 · 计算机科学 2020-05-28 Alex S. Leong , Arunselvan Ramaswamy , Daniel E. Quevedo , Holger Karl , Ling Shi

We consider a multi-process remote estimation system observing $K$ independent Ornstein-Uhlenbeck processes. In this system, a shared sensor samples the $K$ processes in such a way that the long-term average sum mean square error (MSE) is…

信息论 · 计算机科学 2023-11-01 Karim Banawan , Ahmed Arafa , Karim G. Seddik

Remote state estimation, where sensors send their measurements of distributed dynamic plants to a remote estimator over shared wireless resources, is essential for mission-critical applications of Industry 4.0. Existing algorithms on…

系统与控制 · 电气工程与系统科学 2022-05-26 Gaoyang Pang , Wanchun Liu , Yonghui Li , Branka Vucetic

Information source sampling and update scheduling have been treated separately in the context of real-time status update for age of information optimization. In this paper, a unified sampling and scheduling ($\mathcal{S}^2$) approach is…

信息论 · 计算机科学 2018-12-14 Zhiyuan Jiang , Sheng Zhou , Zhisheng Niu , Yu Cheng

Recent theoretical work studies sample-efficient reinforcement learning (RL) extensively in two settings: learning interactively in the environment (online RL), or learning from an offline dataset (offline RL). However, existing algorithms…

机器学习 · 计算机科学 2022-02-14 Tengyang Xie , Nan Jiang , Huan Wang , Caiming Xiong , Yu Bai

We present algorithms to effectively represent a set of Markov decision processes (MDPs), whose optimal policies have already been learned, by a smaller source subset for lifelong, policy-reuse-based transfer learning in reinforcement…

人工智能 · 计算机科学 2016-05-03 M. M. Hassan Mahmud , Majd Hawasly , Benjamin Rosman , Subramanian Ramamoorthy

Motivated by the recent success of Machine Learning (ML) tools in wireless communications, the idea of semantic communication by Weaver from 1949 has gained attention. It breaks with Shannon's classic design paradigm by aiming to transmit…

信息论 · 计算机科学 2023-07-14 Edgar Beck , Carsten Bockelmann , Armin Dekorsy

We propose a novel randomized linear programming algorithm for approximating the optimal policy of the discounted Markov decision problem. By leveraging the value-policy duality and binary-tree data structures, the algorithm adaptively…

最优化与控制 · 数学 2019-06-04 Mengdi Wang

This paper addresses the challenge of packet-based information routing in large-scale wireless communication networks. The problem is framed as a constrained statistical learning task, where each network node operates using only local…

信号处理 · 电气工程与系统科学 2025-04-15 Sourajit Das , Kirtan Gopal Panda , Navid NaderiAlizadeh

We consider an underlay cognitive radio network where the secondary user (SU) harvests energy from the environment. We consider a slotted-mode of operation where each slot of SU is used for either energy harvesting or data transmission.…

信息论 · 计算机科学 2019-06-04 Kalpant Pathak , Adrish Banerjee

We consider a remote state estimation problem in the presence of an eavesdropper over packet dropping links. A smart sensor transmits its local estimates to a legitimate remote estimator, in the course of which an eavesdropper can randomly…

系统与控制 · 电气工程与系统科学 2019-10-10 Jingyi Lu , Alex S. Leong , Daniel E. Quevedo

The future wireless networks envision ultra-reliable communication with efficient use of the limited wireless channel resources. Closed-loop repetition protocols where retransmission of a packet is enabled using a feedback channel has been…

信息论 · 计算机科学 2017-10-03 Saeed R. Khosravirad , Harish Viswanathan

A setup involving zero-delay sequential transmission of a vector Markov source over a burst erasure channel is studied. A sequence of source vectors is compressed in a causal fashion at the encoder, and the resulting output is transmitted…

信息论 · 计算机科学 2014-10-10 Farrokh Etezadi , Ashish Khisti , Mitchell Trott

Jointly optimal transmission power control and remote estimation over an infinite horizon is studied. A sensor observes a dynamic process and sends its observations to a remote estimator over a wireless fading channel characterized by a…

系统与控制 · 计算机科学 2016-05-02 Xiaoqiang Ren , Junfeng Wu , Karl H. Johansson , Guodong Shi , Ling Shi

In this work we present a method for learning a reactive policy for a simple dynamic locomotion task involving hard impact and switching contacts where we assume the contact location and contact timing to be unknown. To learn such a policy,…

机器人学 · 计算机科学 2018-08-07 Julian Viereck , Jules Kozolinsky , Alexander Herzog , Ludovic Righetti