中文
相关论文

相关论文: Two Timescale Convergent Q-learning for Sleep--Sch…

200 篇论文

Preventing and detecting intrusions and attacks on wireless networks has become an important and serious challenge. On the other hand, due to the limited resources of wireless nodes, the use of monitoring nodes for permanent monitoring in…

密码学与安全 · 计算机科学 2022-01-04 Amir Mojtahedi , Farid Sorouri , Alireza Najafi Souha , Aidin Molazadeh , Saeedeh Shafaei Mehr

Consider a Markov decision process (MDP) that admits a set of state-action features, which can linearly express the process's probabilistic transition model. We propose a parametric Q-learning algorithm that finds an approximate-optimal…

机器学习 · 计算机科学 2019-06-07 Lin F. Yang , Mengdi Wang

We present dual-attention neural biasing, an architecture designed to boost Wake Words (WW) recognition and improve inference time latency on speech recognition tasks. This architecture enables a dynamic switch for its runtime compute paths…

In real-world reinforcement learning (RL) scenarios, agents often encounter partial observability, where incomplete or noisy information obscures the true state of the environment. Partially Observable Markov Decision Processes (POMDPs) are…

机器学习 · 计算机科学 2025-05-19 Ashok Arora , Neetesh Kumar

Usage of mobile sink(s) for data gathering in wireless sensor networks(WSNs) improves the performance of WSNs in many respects such as power consumption, lifetime, etc. In some applications, the mobile sink $MS$ travels along a predefined…

网络与互联网体系结构 · 计算机科学 2022-03-22 Dinesh Dash

Penetration testing, the simulation of cyberattacks to identify security vulnerabilities, presents a sequential decision-making problem well-suited for reinforcement learning (RL) automation. Like many applications of RL to real-world…

机器学习 · 计算机科学 2025-09-25 Raphael Simon , Pieter Libin , Wim Mees

We are proposing fully parallel and maximally distributed hardware realization of a generic neuro-computing system. More specifically, the proposal relates to the wireless sensor networks technology to serve as a massively parallel and…

神经与进化计算 · 计算机科学 2025-10-31 Gursel Serpen

Max weighted queue (MWQ) control policy is a widely used cross-layer control policy that achieves queue stability and a reasonable delay performance. In most of the existing literature, it is assumed that optimal MWQ policy can be obtained…

系统与控制 · 计算机科学 2013-11-20 Junting Chen , Vincent K. N. Lau

We consider a wireless sensor network whose main function is to detect certain infrequent alarm events, and to forward alarm packets to a base station, using geographical forwarding. The nodes know their locations, and they sleep-wake…

网络与互联网体系结构 · 计算机科学 2009-12-21 K. P. Naveen , A. Kumar

Signal classification problems arise in a wide variety of applications, and their demand is only expected to grow. In this paper, we focus on the wireless sensor network signal classification setting, where each sensor forwards quantized…

信号处理 · 电气工程与系统科学 2022-06-29 Jing Guo , Raghu G. Raj , David J. Love , Christopher G. Brinton

Passive monitoring utilizing distributed wireless sniffers is an effective technique to monitor activities in wireless infrastructure networks for fault diagnosis, resource management and critical path analysis. In this paper, we introduce…

网络与互联网体系结构 · 计算机科学 2013-02-01 Huy Nguyen , Gabriel Scalosub , Rong Zheng

Next generation communications demand for better spectrum management, lower latency, and guaranteed quality-of-service (QoS). Recently, Artificial intelligence (AI) has been widely introduced to advance these aspects in next generation…

网络与互联网体系结构 · 计算机科学 2024-11-07 Hanwen Zhang , Mingzhe Chen , Alireza Vahid , Feng Ye , Haijian Sun

We investigate an energy-harvesting wireless sensor transmitting latency-sensitive data over a fading channel. The sensor injects captured data packets into its transmission queue and relies on ambient energy harvested from the environment…

网络与互联网体系结构 · 计算机科学 2019-05-07 Nikhilesh Sharma , Nicholas Mastronarde , Jacob Chakareski

Time-inhomogeneous finite-horizon Markov decision processes (MDP) are frequently employed to model decision-making in dynamic treatment regimes and other statistical reinforcement learning (RL) scenarios. These fields, especially healthcare…

机器学习 · 计算机科学 2025-10-21 Elynn Chen , Sai Li , Michael I. Jordan

Sleep disorders are very widespread in the world population and suffer from a generalized underdiagnosis, given the complexity of their diagnostic methods. Therefore, there is an increasing interest in developing simpler screening methods.…

信号处理 · 电气工程与系统科学 2021-02-08 Ramiro Casal , Leandro E. Di Persia , Gastón Schlotthauer

Accurate wireless timing synchronization has been an extremely important topic in wireless sensor networks, required in applications ranging from distributed beam forming to precision localization and navigation. However, it is very…

网络与互联网体系结构 · 计算机科学 2014-09-29 Marcelo Segura , S. Niranjayan , Hossein Hashemi , Andreas F. Molisch

In this paper, a new cooperation structure for spectrum sensing in cognitive radio networks is proposed which outperforms the existing commonly-used ones in terms of energy efficiency. The efficiency is achieved in the proposed design by…

信息论 · 计算机科学 2015-09-11 Younes Abdi , Tapani Ristaniemi

In this paper, we investigate the problem of controlling probabilistic Boolean control networks (PBCNs) to achieve reachability with maximum probability in the finite time horizon. We address three questions: 1) finding control policies…

系统与控制 · 电气工程与系统科学 2023-12-13 Hongyue Fan , Jingjie Ni , Fangfei Li

This paper proposes a multi-agent reinforcement learning based medium access framework for wireless networks. The access problem is formulated as a Markov Decision Process (MDP), and solved using reinforcement learning with every network…

机器学习 · 计算机科学 2021-04-30 Hrishikesh Dutta , Subir Biswas

Soft Q-learning is a variation of Q-learning designed to solve entropy regularized Markov decision problems where an agent aims to maximize the entropy regularized value function. Despite its empirical success, there have been limited…

机器学习 · 计算机科学 2024-09-06 Narim Jeong , Donghwan Lee