English
Related papers

Related papers: Reinforcement Learning Optimizes Power Dispatch in…

200 papers

The problem of regulating voltages within the required limits is complicated by the fact that power system supplies power to a vast number of loads and is fed from many generating units. As loads vary, reactive power requirements of the…

Systems and Control · Computer Science 2018-08-08 Talha Iqbal , Ali Dehghan Banadaki , Ali Feliachi

This paper introduces Group Sequence Policy Optimization (GSPO), our stable, efficient, and performant reinforcement learning algorithm for training large language models. Unlike previous algorithms that adopt token-level importance ratios,…

Machine Learning · Computer Science 2025-07-29 Chujie Zheng , Shixuan Liu , Mingze Li , Xiong-Hui Chen , Bowen Yu , Chang Gao , Kai Dang , Yuqiong Liu , Rui Men , An Yang , Jingren Zhou , Junyang Lin

We consider the problem of controlling the voltage of a distribution feeder using the reactive power capabilities of inverters. On a real distribution grid, we compare the local Volt/VAr droop control recommended in recent grid codes, a…

Systems and Control · Electrical Eng. & Systems 2022-07-22 Lukas Ortmann , Adrian Hauswirth , Ivo Caduff , Florian Dörfler , Saverio Bolognani

Time-varying renewable energy generation can result in serious under-/over-voltage conditions in future distribution grids. Augmenting conventional utility-owned voltage regulating equipment with the reactive power capabilities of…

Systems and Control · Computer Science 2015-08-27 Vassilis Kekatos , Liang Zhang , Georgios B. Giannakis , Ross Baldick

Designing robust algorithms for the optimal power flow (OPF) problem is critical for the control of large-scale power systems under uncertainty. The chance-constrained OPF (CCOPF) problem provides a natural formulation of the trade-off…

Optimization and Control · Mathematics 2025-01-23 Eli Brock , Haixiang Zhang , Javad Lavaei , Somayeh Sojoudi

Reinforcement learning is employed to optimize the periodic forcing signal of a pulsed blowing system that controls flow separation in a fully-turbulent $Re_\theta = 1000$ diffuser flow. Based on the state of the wind tunnel experiment that…

Fluid Dynamics · Physics 2024-12-11 Alexandra Müller , Tobias Schesny , Ben Steinfurth , Julien Weiss

Topology control for power grid operation is a challenging sequential decision making problem because the action space grows combinatorially with the size of the grid and action evaluation through simulation is computationally expensive. We…

Machine Learning · Computer Science 2026-04-03 Pantelis Dogoulis , Maxime Cordy

Power flow (PF) calculations are fundamental to power system analysis to ensure stable and reliable grid operation. The Newton-Raphson (NR) method is commonly used for PF analysis due to its rapid convergence when initialized properly.…

With the uptake of intelligent data-driven applications, edge computing infrastructures necessitate a new generation of admission control algorithms to maximize system performance under limited and highly heterogeneous resources. In this…

Networking and Internet Architecture · Computer Science 2024-07-01 A. Fox , F. De Pellegrini , F. Faticanti , E. Altman , F. Bronzino

In this paper, we consider the problem of distributed consensus optimization over multi-agent networks with directed network topology. Assuming each agent has a local cost function that is smooth and strongly convex, the global objective is…

Optimization and Control · Mathematics 2020-08-21 Shi Pu

Due to the rapid growth of data transmissions in internet of vehicles (IoV), finding schemes that can effectively alleviate access congestion has become an important issue. Recently, many traffic control schemes have been studied.…

Networking and Internet Architecture · Computer Science 2022-04-13 Haijun Zhang , Minghui Jiang , Xiangnan Liu , Keping Long , Victor C. M. Leung

Reinforcement learning (RL) has re-emerged as a natural approach for training interactive LLM agents in real-world environments. However, directly applying the widely used Group Relative Policy Optimization (GRPO) algorithm to multi-turn…

Machine Learning · Computer Science 2026-01-27 Junbo Li , Peng Zhou , Rui Meng , Meet P. Vadera , Lihong Li , Yang Li

Proximal Policy Optimization (PPO) is among the most widely used deep reinforcement learning algorithms, yet its theoretical foundations remain incomplete. Most importantly, convergence and understanding of fundamental PPO advantages remain…

Machine Learning · Computer Science 2026-02-04 Leif Doering , Daniel Schmidt , Moritz Melcher , Sebastian Kassing , Benedikt Wille , Tilman Aach , Simon Weissmann

Direct Preference Optimization (DPO) has been proposed as a promising alternative to Proximal Policy Optimization (PPO) based Reinforcement Learning with Human Feedback (RLHF). However, empirical evaluations consistently reveal suboptimal…

Machine Learning · Computer Science 2025-03-03 Qinwei Ma , Jingzhe Shi , Can Jin , Jenq-Neng Hwang , Serge Belongie , Lei Li

Transmission grid congestion increases as the electrification of various sectors requires transmitting more power. Topology control, through substation reconfiguration, can reduce congestion but its potential remains under-exploited in…

Machine Learning · Computer Science 2025-05-02 Thomas Lautenbacher , Ali Rajaei , Davide Barbieri , Jan Viebahn , Jochen L. Cremer

Since DeepSeek-R1 popularized, Group Relative Policy Optimization (GRPO) has become the core part of training Reasoning LLMs. However, we find some deficiency that influences RL stability and inference efficiency, like zero-variance in…

Computation and Language · Computer Science 2025-09-30 Chen Li , Nazhou Liu , Kai Yang

We propose a computationally efficient approach to safe reinforcement learning (RL) for frequency regulation in power systems with high levels of variable renewable energy resources. The approach draws on set-theoretic control techniques to…

Systems and Control · Electrical Eng. & Systems 2022-03-24 Daniel Tabas , Baosen Zhang

With the proliferation of distributed generation into distribution networks, the need to consider fault currents in the dispatch problem becomes increasingly relevant. This paper introduces a method for adding fault current constraints into…

Systems and Control · Electrical Eng. & Systems 2023-01-31 Jose E. Tabarez , Arthur K. Barnes , Adam Mate , Russell W. Bent

Machine learning assisted optimal power flow (OPF) aims to reduce the computational complexity of these non-linear and non-convex constrained optimization problems by consigning expensive (online) optimization to offline training. The…

Machine Learning · Computer Science 2022-04-28 Thomas Falconer , Letif Mones

Stock portfolio optimization is the process of constant re-distribution of money to a pool of various stocks. In this paper, we will formulate the problem such that we can apply Reinforcement Learning for the task properly. To maintain a…

Machine Learning · Computer Science 2020-12-14 Le Trung Hieu
‹ Prev 1 8 9 10 Next ›