中文
相关论文

相关论文: Reducing Learning Difficulties: One-Step Two-Criti…

200 篇论文

Complex mechanical systems such as vehicle powertrains are inherently subject to multiple nonlinearities and uncertainties arising from parametric variations. Modeling errors are therefore unavoidable, making the transfer of control systems…

系统与控制 · 电气工程与系统科学 2026-02-13 Heisei Yonezawa , Ansei Yonezawa , Itsuro Kajiwara

High-precision control tasks present substantial challenges for reinforcement learning (RL) algorithms, frequently resulting in suboptimal performance attributed to network approximation inaccuracies and inadequate sample quality.These…

机器学习 · 计算机科学 2025-02-05 Donghe Chen , Yubin Peng , Tengjie Zheng , Han Wang , Chaoran Qu , Lin Cheng

Modern distribution grids are currently being challenged by frequent and sizable voltage fluctuations, due mainly to the increasing deployment of electric vehicles and renewable generators. Existing approaches to maintaining bus voltage…

系统与控制 · 计算机科学 2019-12-05 Qiuling Yang , Gang Wang , Alireza Sadeghi , Georgios B. Giannakis , Jian Sun

We study off-dynamics Reinforcement Learning (RL), where the policy is trained on a source domain and deployed to a distinct target domain. We aim to solve this problem via online distributionally robust Markov decision processes (DRMDPs),…

机器学习 · 计算机科学 2024-02-26 Zhishuai Liu , Pan Xu

Reducing operation and maintenance costs is a key objective for advanced reactors in general and microreactors in particular. To achieve this reduction, developing robust autonomous control algorithms is essential to ensure safe and…

系统与控制 · 电气工程与系统科学 2024-06-25 Majdi I. Radaideh , Leo Tunkle , Dean Price , Kamal Abdulraheem , Linyu Lin , Moutaz Elias

This paper proposes a data-driven affinely adjustable robust Volt/VAr control (AARVVC) scheme, which modulates the smart inverter reactive power in an affine function of its active power, based on the voltage sensitivities with respect to…

系统与控制 · 电气工程与系统科学 2022-09-13 Naihao Shi , Rui Cheng , Liming Liu , Zhaoyu Wang , Qianzhi Zhang

Recently, distributed controller architectures have been quickly gaining popularity in Software-Defined Networking (SDN). However, the use of distributed controllers introduces a new and important Request Dispatching (RD) problem with the…

网络与互联网体系结构 · 计算机科学 2023-05-19 Victoria Huang , Gang Chen , Qiang Fu

This paper considers an incremental Volt/Var control scheme for distribution systems with high integration of inverter-interfaced distributed generation (such as photovoltaic systems). The incremental Volt/Var controller is implemented with…

系统与控制 · 电气工程与系统科学 2024-12-16 Antonin Colot , Elisabetta Perotti , Mevludin Glavic , Emiliano Dall'Anese

Transmission switching is a well-established approach primarily applied to minimize operational costs through strategic network reconfiguration. However, exclusive focus on cost reduction can compromise system reliability. While…

系统与控制 · 电气工程与系统科学 2025-07-17 Ding Lin , Jianhui Wang , Tianqiao Zhao , Meng Yue

Deploying controllers trained with Reinforcement Learning (RL) on real robots can be challenging: RL relies on agents' policies being modeled as Markov Decision Processes (MDPs), which assume an inherently discrete passage of time. The use…

机器人学 · 计算机科学 2024-04-03 Dong Wang , Giovanni Beltrame

This research gauges the ability of deep reinforcement learning (DRL) techniques to assist the optimization and control of fluid mechanical systems. It combines a novel, "degenerate" version of the proximal policy optimization (PPO)…

最优化与控制 · 数学 2021-05-19 H. Ghraieb , J. Viquerat , A. Larcher , P. Meliga , E. Hachem

Integrated sensing and communication (ISAC) is emerging as a key enabler for vehicle-to-everything (V2X) systems. However, designing efficient beamforming schemes for ISAC signals to achieve accurate sensing and enhance communication…

网络与互联网体系结构 · 计算机科学 2025-08-20 Chen Shang , Jiadong Yu , Dinh Thai Hoang

Post-click conversion, as a strong signal indicating the user preference, is salutary for building recommender systems. However, accurately estimating the post-click conversion rate (CVR) is challenging due to the selection bias, i.e., the…

机器学习 · 计算机科学 2022-01-11 Siyuan Guo , Lixin Zou , Yiding Liu , Wenwen Ye , Suqi Cheng , Shuaiqiang Wang , Hechang Chen , Dawei Yin , Yi Chang

In this paper, the real-time deployment of unmanned aerial vehicles (UAVs) as flying base stations (BSs) for optimizing the throughput of mobile users is investigated for UAV networks. This problem is formulated as a time-varying…

网络与互联网体系结构 · 计算机科学 2020-02-04 Zhiwei Chen , Yi Zhong , Xiaohu Ge , Yi Ma

Robust control of mechanical systems with multiple uncertainties remains a fundamental challenge, particularly when nonlinear dynamics and operating-condition variations are intricately intertwined. Although deep reinforcement learning…

机器学习 · 计算机科学 2026-03-11 Heisei Yonezawa , Ansei Yonezawa , Itsuro Kajiwara

A major challenge in todays power grid is to manage the increasing load from electric vehicle (EV) charging. Demand response (DR) solutions aim to exploit flexibility therein, i.e., the ability to shift EV charging in time and thus avoid…

人工智能 · 计算机科学 2022-03-29 Manu Lahariya , Nasrin Sadeghianpourhamami , Chris Develder

Deep Reinforcement Learning (DRL) methods often rely on the meticulous tuning of hyperparameters to successfully resolve problems. One of the most influential parameters in optimization procedures based on stochastic gradient descent (SGD)…

机器学习 · 计算机科学 2020-08-05 Ralf Gulde , Marc Tuscher , Akos Csiszar , Oliver Riedel , Alexander Verl

The open radio access network (O-RAN) architecture supports intelligent network control algorithms as one of its core capabilities. Data-driven applications incorporate such algorithms to optimize radio access network (RAN) functions via…

网络与互联网体系结构 · 计算机科学 2023-09-20 Ahmad M. Nagib , Hatem Abou-Zeid , Hossam S. Hassanein

Deep Reinforcement Learning (DRL) has been applied successfully to many robotic applications. However, the large number of trials needed for training is a key issue. Most of existing techniques developed to improve training efficiency (e.g.…

机器人学 · 计算机科学 2018-12-13 Linhai Xie , Sen Wang , Stefano Rosa , Andrew Markham , Niki Trigoni

Adversarial training is a defense method that trains machine learning models on intentionally perturbed attack inputs, so they learn to be robust against adversarial examples. This paper develops a robust voltage control framework for…

系统与控制 · 电气工程与系统科学 2026-03-26 Sungjoo Chung , Ying Zhang