中文
相关论文

相关论文: Reinforcement-learning-based actuator selection me…

200 篇论文

Deep reinforcement learning (RL) has gained widespread adoption in recent years but faces significant challenges, particularly in unknown and complex environments. Among these, high-dimensional action selection stands out as a critical…

机器学习 · 统计学 2025-07-08 Wenbo Zhang , Hengrui Cai

We propose an active set selection framework for Gaussian process classification for cases when the dataset is large enough to render its inference prohibitive. Our scheme consists of a two step alternating procedure of active set update…

机器学习 · 统计学 2011-06-24 Ricardo Henao , Ole Winther

To address computational challenges associated with power flow nonconvexities, significant research efforts over the last decade have developed convex relaxations and approximations of optimal power flow (OPF) problems. However, benefits…

系统与控制 · 电气工程与系统科学 2023-02-24 Babak Taheri , Daniel K. Molzahn

Training expressive flow-based policies with off-policy reinforcement learning is notoriously unstable due to gradient pathologies in the multi-step action sampling process. We trace this instability to a fundamental connection: the flow…

机器人学 · 计算机科学 2026-01-15 Yixian Zhang , Shu'ang Yu , Tonghe Zhang , Mo Guang , Haojia Hui , Kaiwen Long , Yu Wang , Chao Yu , Wenbo Ding

Traffic flow prediction is an important part of smart transportation. The goal is to predict future traffic conditions based on historical data recorded by sensors and the traffic network. As the city continues to build, parts of the…

机器学习 · 统计学 2022-12-27 Yanan Xiao , Minyu Liu , Zichen Zhang , Lu Jiang , Minghao Yin , Jianan Wang

With the increasing proportion of renewable energy in the generation side, it becomes more difficult to accurately predict the power generation and adapt to the large deviations between the optimal dispatch scheme and the day-ahead…

系统与控制 · 电气工程与系统科学 2023-03-07 Xinyue Wang , Haiwang Zhong , Guanglun Zhang , Guangchun Ruan , Yiliu He , Zekuan Yu

Maneuver decision-making is the core of unmanned combat aerial vehicle for autonomous air combat. To solve this problem, we propose an automatic curriculum reinforcement learning method, which enables agents to learn effective decisions in…

人工智能 · 计算机科学 2023-07-13 Zhang Hong-Peng

High-fidelity simulations are performed to study active flow control techniques for alleviating deep dynamic stall of a SD7003 airfoil in plunging motion. The flow Reynolds number is $Re=60{,}000$ and the freestream Mach number is $M=0.1$.…

流体动力学 · 物理学 2019-07-15 Brener D'Lélis Oliveira Ramos , William Roberto Wolf , Chi-An Yeh , Kunihiko Taira

Order Picker Routing is a critical issue in Warehouse Operations Management. Due to the complexity of the problem and the need for quick solutions, suboptimal algorithms are frequently employed in practice. However, Reinforcement Learning…

机器学习 · 计算机科学 2024-02-07 George Dunn , Hadi Charkhgard , Ali Eshragh , Sasan Mahmoudinazlou , Elizabeth Stojanovski

A novel deep multi-agent reinforcement learning framework is proposed to identify and resolve conflicts among a variable number of aircraft in a high-density, stochastic, and dynamic sector. Currently the sector capacity is constrained by…

机器学习 · 计算机科学 2020-08-28 Marc Brittain , Xuxi Yang , Peng Wei

The real power of artificial intelligence appears in reinforcement learning, which is computationally and physically more sophisticated due to its dynamic nature. Rotation and injection are some of the proven ways in active flow control for…

流体动力学 · 物理学 2024-01-02 Kamyar Dobakhti , Jafar Ghazanfarian

Likelihood-based policy gradient methods are the dominant approach for training robot control policies from rewards. These methods rely on differentiable action likelihoods, which constrain policy outputs to simple distributions like…

In this paper, we investigate the problem of actuator selection for linear dynamical systems. We develop a framework to design a sparse actuator schedule for a given large-scale linear system with guaranteed performance bounds using…

系统与控制 · 计算机科学 2020-06-04 Milad Siami , Alex Olshevsky , Ali Jadbabaie

In this paper, we propose the first fully push-forward-based distributional reinforcement learning algorithm, named PACER, which consists of a distributional critic, a stochastic actor and a sample-based encourager. Specifically, the…

机器学习 · 计算机科学 2024-10-10 Wensong Bai , Chao Zhang , Yichao Fu , Peilin Zhao , Hui Qian , Bin Dai

We focus on a simulation-based optimization problem of choosing the best design from the feasible space. Although the simulation model can be queried with finite samples, its internal processing rule cannot be utilized in the optimization…

机器学习 · 计算机科学 2021-11-02 Kuo Li , Qing-Shan Jia , Jiaqi Yan

While reinforcement learning algorithms provide automated acquisition of optimal policies, practical application of such methods requires a number of design decisions, such as manually designing reward functions that not only define the…

机器学习 · 计算机科学 2022-12-29 Tim G. J. Rudner , Vitchyr H. Pong , Rowan McAllister , Yarin Gal , Sergey Levine

We present a new optimization-based method for aggregating preferences in settings where each voter expresses preferences over pairs of alternatives. Our approach to identifying a consensus partial order is motivated by the observation that…

社会与信息网络 · 计算机科学 2025-01-28 Nathan Atkinson , Scott C. Ganz , Dorit S. Hochbaum , James B. Orlin

This paper proposes a Separable Projective Approximation Routine-Optimal Power Flow (SPAR-OPF) framework for solving two-stage stochastic optimization problems in power systems. The framework utilizes a separable piecewise linear…

系统与控制 · 电气工程与系统科学 2025-09-25 Shishir Lamichhane , Abodh Poudyal , Nicholas R. Jones , Bala Krishnamoorthy , Anamika Dubey

As trajectories sampled by policies used by reinforcement learning (RL) and generative flow networks (GFlowNets) grow longer, credit assignment and exploration become more challenging, and the long planning horizon hinders mode discovery…

In many engineered systems, optimization is used for decision making at time-scales ranging from real-time operation to long-term planning. This process often involves solving similar optimization problems over and over again with slightly…

最优化与控制 · 数学 2019-01-18 Sidhant Misra , Line Roald , Yeesian Ng