English
Related papers

Related papers: ReasonLight: A Multimodal Foundation Model-Enhance…

200 papers

Inspired by the remarkable reasoning capabilities of Deepseek-R1 in complex textual tasks, many works attempt to incentivize similar capabilities in Multimodal Large Language Models (MLLMs) by directly applying reinforcement learning (RL).…

Machine Learning · Computer Science 2026-01-29 Shuang Chen , Yue Guo , Zhaochen Su , Yafu Li , Yulun Wu , Jiacheng Chen , Jiayu Chen , Weijie Wang , Xiaoye Qu , Yu Cheng

DeepSeek-R1-Zero has successfully demonstrated the emergence of reasoning capabilities in LLMs purely through Reinforcement Learning (RL). Inspired by this breakthrough, we explore how RL can be utilized to enhance the reasoning capability…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Wenxuan Huang , Bohan Jia , Zijie Zhai , Shaosheng Cao , Zheyu Ye , Fei Zhao , Zhe Xu , Xu Tang , Yao Hu , Shaohui Lin

Currently, traffic signal control (TSC) methods based on reinforcement learning (RL) have proven superior to traditional methods. However, most RL methods face difficulties when applied in the real world due to three factors: input, output,…

Multiagent Systems · Computer Science 2024-07-16 Haoyuan Jiang , Xuantang Xiong , Ziyue Li , Hangyu Mao , Guanghu Sui , Jingqing Ruan , Yuheng Cheng , Hua Wei , Wolfgang Ketter , Rui Zhao

Reinforcement Learning (RL) has been shown to be effective in many scenarios. However, it typically requires the exploration of a sufficiently large number of state-action pairs, some of which may be unsafe. Consequently, its application to…

Systems and Control · Electrical Eng. & Systems 2022-06-24 Yousef Emam , Gennaro Notomista , Paul Glotfelter , Zsolt Kira , Magnus Egerstedt

Reinforcement learning (RL) offers a compelling data-driven paradigm for synthesizing controllers for complex systems when accurate physical models are unavailable; however, most existing control-oriented RL methods assume stationarity and,…

Machine Learning · Computer Science 2026-04-22 Austin Coursey , Abel Diaz-Gonzalez , Marcos Quinones-Grueiro , Gautam Biswas

Effective traffic signal control (TSC) is crucial in mitigating urban congestion and reducing emissions. Recently, reinforcement learning (RL) has been the research trend for TSC. However, existing RL algorithms face several real-world…

Systems and Control · Electrical Eng. & Systems 2025-09-30 Mingyuan Li , Jiahao Wang , Bo Du , Jun Shen , Qiang Wu

End-to-end autonomous driving frameworks face persistent challenges in generalization, training efficiency, and interpretability. While recent methods leverage Vision-Language Models (VLMs) through supervised learning on large-scale…

Robotics · Computer Science 2025-12-11 Lin Li , Yuxin Cai , Jianwu Fang , Jianru Xue , Chen Lv

Zero-shot reasoning on text-rich networks (TRNs) remains a challenging frontier, as models must integrate textual semantics with relational structure without task-specific supervision. While graph neural networks rely on fixed label spaces…

Computation and Language · Computer Science 2026-04-22 Yilun Liu , Ruihong Qiu , Zi Huang

Traffic signal control has the potential to reduce congestion in dynamic networks. Recent studies show that traffic signal control with reinforcement learning (RL) methods can significantly reduce the average waiting time. However, a…

Systems and Control · Electrical Eng. & Systems 2024-06-13 Maonan Wang , Yutong Xu , Xi Xiong , Yuheng Kan , Chengcheng Xu , Man-On Pun

Zero-shot reinforcement learning (RL) promises to provide agents that can perform any task in an environment after an offline, reward-free pre-training phase. Methods leveraging successor measures and successor features have shown strong…

Machine Learning · Computer Science 2024-10-31 Scott Jeen , Tom Bewley , Jonathan M. Cullen

Reinforcement learning (RL) holds significant promise for adaptive traffic signal control. While existing RL-based methods demonstrate effectiveness in reducing vehicular congestion, their predominant focus on vehicle-centric optimization…

Machine Learning · Computer Science 2025-07-24 Bibek Poudel , Xuan Wang , Weizi Li , Lei Zhu , Kevin Heaslip

The application of reinforcement learning in traffic signal control (TSC) has been extensively researched and yielded notable achievements. However, most existing works for TSC assume that traffic data from all surrounding intersections is…

Systems and Control · Electrical Eng. & Systems 2024-11-01 Hanyang Chen , Yang Jiang , Shengnan Guo , Xiaowei Mao , Youfang Lin , Huaiyu Wan

Reinforcement learning (RL) has emerged as a promising approach for eliciting reasoning chains before generating final answers. However, multimodal large language models (MLLMs) generate reasoning that lacks integration of visual…

Computer Vision and Pattern Recognition · Computer Science 2026-01-05 Omar Sharif , Eftekhar Hossain , Patrick Ng

Multi-agent Deep Reinforcement Learning (MADRL) based traffic signal control becomes a popular research topic in recent years. To alleviate the scalability issue of completely centralized RL techniques and the non-stationarity issue of…

Artificial Intelligence · Computer Science 2023-09-08 Hankang Gu , Shangbo Wang , Xiaoguang Ma , Dongyao Jia , Guoqiang Mao , Eng Gee Lim , Cheuk Pong Ryan Wong

Single occupancy vehicles are the most attractive transportation alternative for many commuters, leading to increased traffic congestion and air pollution. Advancements in information technologies create opportunities for smart solutions…

Machine Learning · Computer Science 2023-04-10 Dimitris M. Vlachogiannis , Hua Wei , Scott Moura , Jane Macfarlane

Urban congestion remains a critical challenge, with traffic signal control (TSC) emerging as a potent solution. TSC is often modeled as a Markov Decision Process problem and then solved using reinforcement learning (RL), which has proven…

Artificial Intelligence · Computer Science 2024-07-09 Aoyu Pang , Maonan Wang , Man-On Pun , Chung Shue Chen , Xi Xiong

Recent advancements in Chain of Thought (COT) generation have significantly improved the reasoning capabilities of Large Language Models (LLMs), with reinforcement learning (RL) emerging as an effective post-training approach. Multimodal…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Yi Chen , Yuying Ge , Rui Wang , Yixiao Ge , Lu Qiu , Ying Shan , Xihui Liu

The growing demand for road use in urban areas has led to significant traffic congestion, posing challenges that are costly to mitigate through infrastructure expansion alone. As an alternative, optimizing existing traffic management…

Artificial Intelligence · Computer Science 2024-09-04 Muhammad Tahir Rafique , Ahmed Mustafa , Hasan Sajid

Traffic signal control (TSC) is a high-stakes domain that is growing in importance as traffic volume grows globally. An increasing number of works are applying reinforcement learning (RL) to TSC; RL can draw on an abundance of traffic data…

Artificial Intelligence · Computer Science 2022-10-05 Rex Chen , Fei Fang , Norman Sadeh

Traffic congestion remains a major challenge for urban transportation, leading to significant economic and environmental impacts. Traffic Signal Control (TSC) is one of the key measures to mitigate congestion, and recent studies have…

Multiagent Systems · Computer Science 2026-01-27 Hsiao-Chuan Chang , Sheng-You Huang , Yen-Chi Chen , I-Chen Wu