中文
相关论文

相关论文: Distributed Feedback-Feedforward Algorithms for Ti…

200 篇论文

This paper considers distributed resource allocation problems (DRAPs) with a coupled constraint for real-time systems. Based on primal-dual methods, we adopt a control perspective for optimization algorithm design by synthesizing a safe…

最优化与控制 · 数学 2025-08-05 Wenwen Wu , Shanying Zhu , Cailian Chen , Xinping Guan

Integrating time-frequency resource conversion (TFRC), a new network resource allocation strategy, with call admission control can not only increase the cell capacity but also reduce network congestion effectively. However, the optimal…

网络与互联网体系结构 · 计算机科学 2016-05-24 Hangguan Shan , Yani Zhang , Weihua Zhuang , Aiping Huang , Zhaoyang Zhang

Load frequency control (LFC) is a key factor to maintain the stable frequency in multi-area power systems. As the modern power systems evolve from centralized to distributed paradigm, LFC needs to consider the peer-to-peer (P2P) based…

最优化与控制 · 数学 2022-09-27 Kyung-bin Kwon , Sayak Mukherjee , Hao Zhu , Thanh Long Vu

Distributional reinforcement learning (DRL) enhances the understanding of the effects of the randomness in the environment by letting agents learn the distribution of a random return, rather than its expected value as in standard RL. At the…

最优化与控制 · 数学 2023-03-27 Zifan Wang , Yulong Gao , Siyi Wang , Michael M. Zavlanos , Alessandro Abate , Karl H. Johansson

In most of the transfer learning approaches to reinforcement learning (RL) the distribution over the tasks is assumed to be stationary. Therefore, the target and source tasks are i.i.d. samples of the same distribution. In the context of…

机器学习 · 计算机科学 2020-06-19 Giuseppe Canonaco , Andrea Soprani , Manuel Roveri , Marcello Restelli

Diffusion policies are competitive for offline reinforcement learning (RL) but are typically guided at sampling time by heuristics that lack a statistical notion of risk. We introduce LRT-Diffusion, a risk-aware sampling rule that treats…

机器学习 · 计算机科学 2026-02-20 Ximan Sun , Xiang Cheng

This paper presents a model reference adaptive control (MRAC) framework for uncertain linear time-invariant (LTI) systems subject to user-defined, time-varying state and input constraints. The proposed design seamlessly integrates a…

系统与控制 · 电气工程与系统科学 2025-09-01 Poulomee Ghosh , Shubhendu Bhasin

Distributional reinforcement learning (DRL) enhances the understanding of the effects of the randomness in the environment by letting agents learn the distribution of a random return, rather than its expected value as in standard…

最优化与控制 · 数学 2024-03-26 Zifan Wang , Yulong Gao , Siyi Wang , Michael M. Zavlanos , Alessandro Abate , Karl H. Johansson

This paper proposes three novel resource and user scheduling algorithms with contiguous frequency-domain resource allocation (FDRA) for wireless communications systems. The first proposed algorithm jointly schedules users and resources…

网络与互联网体系结构 · 计算机科学 2020-10-07 Shu Sun , Sungho Moon

Meta-reinforcement learning algorithms provide a data-driven way to acquire policies that quickly adapt to many tasks with varying rewards or dynamics functions. However, learned meta-policies are often effective only on the exact task…

机器学习 · 计算机科学 2023-07-13 Anurag Ajay , Abhishek Gupta , Dibya Ghosh , Sergey Levine , Pulkit Agrawal

Resource allocation plays a central role in many networked systems such as smart grids, communication networks and urban transportation systems. In these systems, many constraints have physical meaning and having feasible allocation is…

最优化与控制 · 数学 2022-07-14 Xuyang Wu , Sindri Magnusson , Mikael Johansson

This paper is devoted to the distributed continuous-time optimization problem with time-varying objective functions and time-varying nonlinear inequality constraints. Different from most studied distributed optimization problems with…

最优化与控制 · 数学 2020-09-08 Shan Sun , Wei Ren

This paper develops an algorithmic framework for tracking fixed points of time-varying contraction mappings. Analytical results for the tracking error are established for the cases where: (i) the underlying contraction self-map changes at…

最优化与控制 · 数学 2018-09-14 Andrey Bernstein , Emiliano Dall'Anese

We investigate energy-efficiency issues and resource allocation policies for time division multi-access (TDMA) over fading channels in the power-limited regime. Supposing that the channels are frequency-flat block-fading and transmitters…

信息论 · 计算机科学 2007-07-13 Xin Wang , Georgios B. Giannakis

Scheduling plays a pivotal role in multi-user wireless communications, since the quality of service of various users largely depends upon the allocated radio resources. In this paper, we propose a novel scheduling algorithm with contiguous…

网络与互联网体系结构 · 计算机科学 2020-11-30 Shu Sun , Xiaofeng Li

This paper focuses on finite-time (FT) convergent distributed algorithms for solving time-varying (TV) distributed optimization (TVDO). The objective is to minimize the sum of local TV cost functions subject to the possible TV constraints…

最优化与控制 · 数学 2023-09-04 Xinli Shi , Guanghui Wen , Xinghuo Yu

In this paper, we address distributed convergence to fair allocations of CPU resources for time-sensitive applications. We propose a novel resource management framework where a centralized objective for fair allocations is decomposed into a…

最优化与控制 · 数学 2015-08-20 Georgios C. Chasparis , Martina Maggio , Enrico Bini , Karl-Eric Årzén

Model-free or learning-based control, in particular, reinforcement learning (RL), is expected to be applied for complex robotic tasks. Traditional RL requires a policy to be optimized is state-dependent, that means, the policy is a kind of…

机器学习 · 计算机科学 2022-08-09 Taisuke Kobayashi , Kenta Yoshizawa

We study the policy evaluation problem in multi-agent reinforcement learning, modeled by a Markov decision process. In this problem, the agents operate in a common environment under a fixed control policy, working together to discover the…

最优化与控制 · 数学 2020-01-13 Thinh T. Doan , Siva Theja Maguluri , Justin Romberg

This paper considers the problem of steady-state real-time optimization (RTO) of interconnected systems with a common constraint that couples several units, for example, a shared resource. Such problems are often studied under the context…

‹ 上一页 1 2 3 10 下一页 ›