中文
相关论文

相关论文: Bayesian Critique-Tune-Based Reinforcement Learnin…

200 篇论文

Safe reinforcement learning (RL) is a standard paradigm for safety-critical decision making. However, real-world safety constraints can be complex, subjective, and even hard to explicitly specify. Existing works on constraint inference rely…

机器学习 · 计算机科学 2026-05-25 Chenglin Li , Grant Ruan , Hua Geng

Air traffic control (ATC) is a safety-critical service system that demands constant attention from ground air traffic controllers (ATCos) to maintain daily aviation operations. The workload of the ATCos can have negative effects on…

机器学习 · 计算机科学 2023-07-25 Yutian Pang , Jueming Hu , Christopher S. Lieber , Nancy J. Cooke , Yongming Liu

Reinforcement learning (RL) is gaining popularity as an effective approach for traffic signal control (TSC) and is increasingly applied in this domain. However, most existing RL methodologies are confined to a single-stage TSC framework,…

机器学习 · 计算机科学 2024-05-03 Liang Zhang , Yutong Zhang , Shubin Xie , Jianming Deng , Chen Li

Autonomous driving at intersections is one of the most complicated and accident-prone traffic scenarios, especially with mixed traffic participants such as vehicles, bicycles and pedestrians. The driving policy should make safe decisions to…

机器学习 · 计算机科学 2022-04-27 Jianhua Jiang , Yangang Ren , Yang Guan , Shengbo Eben Li , Yuming Yin , Xiaoping Jin

End-to-end autonomous driving, where the entire driving pipeline is replaced with a single neural network, has recently gained research attention because of its simpler structure and faster inference time. Despite this appealing approach…

计算机视觉与模式识别 · 计算机科学 2024-09-13 Hongkuan Zhou , Wei Cao , Aifen Sui , Zhenshan Bing

In this position paper, we address the problems of automated road congestion detection and alerting systems and their security properties. We review different theoretical adaptive road traffic control approaches, and three widely deployed…

密码学与安全 · 计算机科学 2016-06-06 Vinh Thong Ta

Achieving both realism and controllability in closed-loop traffic simulation remains a key challenge in autonomous driving. Dataset-based methods reproduce realistic trajectories but suffer from covariate shift in closed-loop deployment,…

机器人学 · 计算机科学 2025-09-23 Keyu Chen , Wenchao Sun , Hao Cheng , Sifa Zheng

Appropriate traffic state representation is crucial for learning traffic signal control policies. However, most of the current traffic state representations are heuristically designed, with insufficient theoretical support. In this paper,…

机器学习 · 计算机科学 2025-03-27 Xiao-Cheng Liao , Yi Mei , Mengjie Zhang , Xiang-Ling Chen

Reinforcement learning (RL) for traffic signal control (TSC) has shown better performance in simulation for controlling the traffic flow of intersections than conventional approaches. However, due to several challenges, no RL-based TSC has…

机器学习 · 计算机科学 2022-06-22 Arthur Müller , Matthia Sabatelli

The emergence of reinforcement learning (RL) methods in traffic signal control tasks has achieved better performance than conventional rule-based approaches. Most RL approaches require the observation of the environment for the agent to…

机器学习 · 计算机科学 2023-11-16 Hao Mei , Junxian Li , Bin Shi , Hua Wei

Bayesian inference in high-dimensional discrete-input additive noise models is a fundamental challenge in communication systems, as the support of the required joint a posteriori probability (APP) mass function grows exponentially with the…

信息论 · 计算机科学 2026-04-08 Luca Schmid , Dominik Sulz , Shrinivas Chimmalgi , Laurent Schmalen

Large pre-trained Vision-Language Models (VLMs) such as CLIP have demonstrated excellent zero-shot generalizability across various downstream tasks. However, recent studies have shown that the inference performance of CLIP can be greatly…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Xin Wang , Kai Chen , Jiaming Zhang , Jingjing Chen , Xingjun Ma

This work introduces a new multi-task, parameter-efficient language model (LM) tuning method that learns to transfer knowledge across different tasks via a mixture of soft prompts-small prefix embedding vectors pre-trained for different…

计算与语言 · 计算机科学 2022-12-02 Akari Asai , Mohammadreza Salehi , Matthew E. Peters , Hannaneh Hajishirzi

The Congestion Control (CC) module plays a critical role in the Transmission Control Protocol (TCP), ensuring the stability and efficiency of network data transmission. The CC approaches that are commonly used these days employ…

网络与互联网体系结构 · 计算机科学 2025-09-15 Jinming Xing , Muhammad Shahzad

Gradient-based approaches in reinforcement learning (RL) have achieved tremendous success in learning policies for autonomous vehicles. While the performance of these approaches warrants real-world adoption, these policies lack…

机器学习 · 计算机科学 2023-08-01 Rohan Paleja , Yaru Niu , Andrew Silva , Chace Ritchie , Sugju Choi , Matthew Gombolay

Reinforcement learning-based traffic signal control (RL-TSC) has emerged as a promising approach for improving urban mobility. However, its robustness under real-world disruptions such as traffic incidents remains largely underexplored. In…

机器学习 · 计算机科学 2025-06-18 Dang Viet Anh Nguyen , Carlos Lima Azevedo , Tomer Toledo , Filipe Rodrigues

Recently, safe reinforcement learning (RL) with the actor-critic structure for continuous control tasks has received increasing attention. It is still challenging to learn a near-optimal control policy with safety and convergence…

机器学习 · 计算机科学 2024-02-06 Xinglong Zhang , Yaoqian Peng , Biao Luo , Wei Pan , Xin Xu , Haibin Xie

As travel demand increases and urban traffic condition becomes more complicated, applying multi-agent deep reinforcement learning (MARL) to traffic signal control becomes one of the hot topics. The rise of Reinforcement Learning (RL) has…

人工智能 · 计算机科学 2023-06-06 Shijie Wang , Shangbo Wang

Coordinating intersections in arterial networks is critical to the performance of urban transportation systems. Deep reinforcement learning (RL) has gained traction in traffic control research along with data-driven approaches for traffic…

系统与控制 · 电气工程与系统科学 2022-08-30 Keith Anshilo Diaz , Damian Dailisan , Umang Sharaf , Carissa Santos , Qijian Gan , Francis Aldrine Uy , May T. Lim , Alexandre M. Bayen

Transient stability and critical clearing time (CCT) are important concepts in power system protection and control. This paper explores and compares various learning-based methods for predicting CCT under uncertainties arising from…

系统与控制 · 电气工程与系统科学 2024-09-05 Xingjian Wu , Xiaoting Wang , Xiaozhe Wang , Peter E. Caines , Jingyu Liu