中文
相关论文

相关论文: Real-time Mode-Aware Dataflow: A Dataflow Model to…

200 篇论文

In many real-world reinforcement learning (RL) problems, besides optimizing the main objective function, an agent must concurrently avoid violating a number of constraints. In particular, besides optimizing performance it is crucial to…

机器学习 · 计算机科学 2018-05-22 Yinlam Chow , Ofir Nachum , Edgar Duenez-Guzman , Mohammad Ghavamzadeh

Perimeter control maintains high traffic efficiency within protected regions by controlling transfer flows among regions to ensure that their traffic densities are below critical values. Existing approaches can be categorized as either…

机器学习 · 计算机科学 2023-06-01 Xiaocan Li , Ray Coden Mercurius , Ayal Taitler , Xiaoyu Wang , Mohammad Noaeen , Scott Sanner , Baher Abdulhai

Optical microscopy provides rich spatio-temporal information characterizing in vivo molecular motion. However, effective forces and other parameters used to summarize molecular motion change over time in live cells due to latent state…

定量方法 · 定量生物学 2015-11-06 Christopher P. Calderon , Kerry S. Bloom

In this ongoing work, we are interested in multiprocessor energy efficient systems, where task durations are not known in advance, but are know stochastically. More precisely, we consider global scheduling algorithms for frame-based…

操作系统 · 计算机科学 2008-09-25 Vandy Berten , Joël Goossens

In this work, we focus on the problem of safe policy transfer in reinforcement learning: we seek to leverage existing policies when learning a new task with specified constraints. This problem is important for safety-critical applications…

机器学习 · 计算机科学 2022-11-11 Zeyu Feng , Bowen Zhang , Jianxin Bi , Harold Soh

Reinforcement learning (RL) has shown a promising performance in learning optimal policies for a variety of sequential decision-making tasks. However, in many real-world RL problems, besides optimizing the main objectives, the agent is…

机器学习 · 计算机科学 2021-07-30 Ashkan B. Jeddi , Nariman L. Dehghani , Abdollah Shafieezadeh

Safety and robustness are two desired properties for any reinforcement learning algorithm. CMDPs can handle additional safety constraints and RMDPs can perform well under model uncertainties. In this paper, we propose to unite these two…

机器学习 · 计算机科学 2021-08-21 Reazul Hasan Russel , Mouhacine Benosman , Jeroen Van Baar , Radu Corcodel

Imitation learning has emerged as an effective approach for bootstrapping sequential decision-making in robotics, achieving strong performance even in high-dimensional dexterous manipulation tasks. Recent behavior cloning methods further…

机器人学 · 计算机科学 2026-02-06 Entong Su , Tyler Westenbroek , Anusha Nagabandi , Abhishek Gupta

Motion prediction systems aim to capture the future behavior of traffic scenarios enabling autonomous vehicles to perform safe and efficient planning. The evolution of these scenarios is highly uncertain and depends on the interactions of…

Many applications -- including power systems, robotics, and economics -- involve a dynamical system interacting with a stochastic and hard-to-model environment. We adopt a reinforcement learning approach to control such systems.…

最优化与控制 · 数学 2025-08-26 Abed AlRahman Al Makdah , Oliver Kosut , Lalitha Sankar , Shaofeng Zou

Multi-fidelity modelling arises in many situations in computational science and engineering world. It enables accurate inference even when only a small set of accurate data is available. Those data often come from a high-fidelity model,…

机器学习 · 统计学 2022-04-12 Jiahao Zhang , Shiqi Zhang , Guang Lin

While Deep Reinforcement Learning (DRL) has emerged as a promising solution for intricate control tasks, the lack of explainability of the learned policies impedes its uptake in safety-critical applications, such as automated driving…

机器学习 · 计算机科学 2024-04-30 Amir Samadi , Konstantinos Koufos , Kurt Debattista , Mehrdad Dianati

With an increasing high penetration of solar photovoltaic generation in electric power grids, voltage phasors and branch power flows experience more severe fluctuations. In this context, probabilistic power flow (PPF) study aims at…

系统与控制 · 电气工程与系统科学 2022-05-03 Kejun Chen , Yu Zhang

Model-based approaches to the verification of non-terminating Cyber-Physical Systems (CPSs) usually rely on numerical simulation of the System Under Verification (SUV) model under input scenarios of possibly varying duration, chosen among…

计算机科学中的逻辑 · 计算机科学 2021-09-09 Toni Mancini , Igor Melatti , Enrico Tronci

Multi-agent pathfinding (MAPF) is a critical field in many large-scale robotic applications, often being the fundamental step in multi-agent systems. The increasing complexity of MAPF in complex and crowded environments, however, critically…

人工智能 · 计算机科学 2024-02-09 Jaehoon Chung , Jamil Fayyad , Younes Al Younes , Homayoun Najjaran

Reinforcement learning (RL) in partially observable, fully cooperative multi-agent settings (Dec-POMDPs) can in principle be used to address many real-world challenges such as controlling a swarm of rescue robots or a team of quadcopters.…

人工智能 · 计算机科学 2022-02-08 Qizhen Zhang , Chris Lu , Animesh Garg , Jakob Foerster

Soft matter materials and polymers are widely used in the controlled delivery of drugs. Simulation and modeling provide insight at the atomic scale enabling a level of control unavailable to experiments. We present a workflow protocol for…

软凝聚态物质 · 物理学 2022-03-08 James P. Andrews , Estela Blaisten-Barojas

Deep Learning Recommendation Models (DLRMs) have become increasingly popular and prevalent in today's datacenters, consuming most of the AI inference cycles. The performance of DLRMs is heavily influenced by available bandwidth due to their…

Cyber-physical systems (CPSs) are man-made complex systems coupled with natural processes that, as a whole, should be described by distributed parameter systems (DPSs) in general forms. This paper presents three such general models for…

经典分析与常微分方程 · 数学 2015-09-16 Fudong Ge , YangQuan Chen , Chunhai Kou

We propose a generalization of the neurotransmitter release model proposed in \emph{Rodrigues et al. (PNAS, 2016)}. We increase the complexity of the underlying slow-fast system by considering a degree-four polynomial as parametrization of…

动力系统 · 数学 2024-05-10 Mattia Sensi , Mathieu Desroches , Serafim Rodrigues
‹ 上一页 1 8 9 10 下一页 ›