English
Related papers

Related papers: Learning Rollout from Sampling:An R1-Style Tokeniz…

200 papers

Decoding strategies play a central role in shaping the reasoning ability of large language models (LLMs). Traditional methods such as greedy decoding and beam search often suffer from error propagation, while sampling-based approaches…

Navigating human-populated environments without causing discomfort is a critical capability for socially-aware agents. While rule-based approaches offer interpretability through predefined psychological principles, they often lack…

Artificial Intelligence · Computer Science 2025-11-17 Yitian Kou , Yihe Gu , Chen Zhou , DanDan Zhu , Shuguang Kuai

Traffic simulation is important for transportation optimization and policy making. While existing simulators such as SUMO and MATSim offer fully-featured platforms and utilities, users without too much knowledge about these platforms often…

Artificial Intelligence · Computer Science 2025-12-25 Yuwei Du , Jun Zhang , Jie Feng , Zhicheng Liu , Jian Yuan , Yong Li

This study introduces a novel approach to autonomous motion planning, informing an analytical algorithm with a reinforcement learning (RL) agent within a Frenet coordinate system. The combination directly addresses the challenges of…

Robotics · Computer Science 2024-07-31 Rainer Trauth , Alexander Hobmeier , Johannes Betz

Training intelligent agents that can drive autonomously in various urban and highway scenarios has been a hot topic in the robotics society within the last decades. However, the diversity of driving environments in terms of road topology…

Robotics · Computer Science 2022-04-06 Behrad Toghi , Rodolfo Valiente , Ramtin Pedarsani , Yaser P. Fallah

Forecasting the future trajectories of surrounding agents is crucial for autonomous vehicles to ensure safe, efficient, and comfortable route planning. While model ensembling has improved prediction accuracy in various fields, its…

Machine Learning · Computer Science 2024-09-23 Aron Distelzweig , Eitan Kosman , Andreas Look , Faris Janjoš , Denesh K. Manivannan , Abhinav Valada

Traffic signal control is an emerging application scenario for reinforcement learning. Besides being as an important problem that affects people's daily life in commuting, traffic signal control poses its unique challenges for reinforcement…

Multiagent Systems · Computer Science 2019-05-15 Huichu Zhang , Siyuan Feng , Chang Liu , Yaoyao Ding , Yichen Zhu , Zihan Zhou , Weinan Zhang , Yong Yu , Haiming Jin , Zhenhui Li

Sampling-based model predictive control methods like MPPI and CEM are essential for real-time control of nonlinear robotic systems, particularly where discontinuous dynamics preclude gradient-based optimization. However, these methods…

Robotics · Computer Science 2026-05-05 Vincent Pacelli , Akash Ratheesh , Evangelos A. Theodorou

Safe and feasible trajectory planning is critical for real-world autonomous driving systems. However, existing learning-based planners rely heavily on expert demonstrations, which not only lack explicit safety awareness but also risk…

Robotics · Computer Science 2025-09-29 Xiaolong Tang , Meina Kan , Shiguang Shan , Xilin Chen

Expert human drivers perform actions relying on traffic laws and their previous experience. While traffic laws are easily embedded into an artificial brain, modeling human complex behaviors which come from past experience is a more…

Multiagent Systems · Computer Science 2019-03-05 Giulio Bacchiani , Daniele Molinari , Marco Patander

A driving algorithm that aligns with good human driving practices, or at the very least collaborates effectively with human drivers, is crucial for developing safe and efficient autonomous vehicles. In practice, two main approaches are…

Multiagent Systems · Computer Science 2026-02-10 Zhihao Zhang , Keith Redmill , Chengyang Peng , Bowen Weng

Mobile robots are often tasked with repeatedly navigating through an environment whose traversability changes over time. These changes may exhibit some hidden structure, which can be learned. Many studies consider reactive algorithms for…

Robotics · Computer Science 2020-12-07 Florence Tsang , Tristan Walker , Ryan A. MacDonald , Armin Sadeghi , Stephen L. Smith

Recently, safe reinforcement learning (RL) with the actor-critic structure for continuous control tasks has received increasing attention. It is still challenging to learn a near-optimal control policy with safety and convergence…

Machine Learning · Computer Science 2024-02-06 Xinglong Zhang , Yaoqian Peng , Biao Luo , Wei Pan , Xin Xu , Haibin Xie

In this paper, we present the Role Playing Learning (RPL) scheme for a mobile robot to navigate socially with its human companion in populated environments. Neural networks (NN) are constructed to parameterize a stochastic policy that…

Robotics · Computer Science 2017-05-30 Mingming Li , Rui Jiang , Shuzhi Sam Ge , Tong Heng Lee

Identifying uncertainty and taking mitigating actions is crucial for safe and trustworthy reinforcement learning agents, especially when deployed in high-risk environments. In this paper, risk sensitivity is promoted in a model-based…

Machine Learning · Computer Science 2021-11-10 Stefan Radic Webster , Peter Flach

Multi-turn tool calling is challenging for Large Language Models (LLMs) because rewards are sparse and exploration is expensive. A common recipe, SFT followed by GRPO, can stall when within-group reward variation is low (e.g., more rollouts…

Artificial Intelligence · Computer Science 2026-02-04 Haitian Zhong , Jixiu Zhai , Lei Song , Jiang Bian , Qiang Liu , Tieniu Tan

In multi-agent based traffic simulation, agents are always supposed to move following existing instructions, and mechanically and unnaturally imitate human behavior. The human drivers perform acceleration or deceleration irregularly all the…

Multiagent Systems · Computer Science 2021-01-26 Junjie Zhong , Hiromitsu Hattori

The growing demand for road use in urban areas has led to significant traffic congestion, posing challenges that are costly to mitigate through infrastructure expansion alone. As an alternative, optimizing existing traffic management…

Artificial Intelligence · Computer Science 2024-09-04 Muhammad Tahir Rafique , Ahmed Mustafa , Hasan Sajid

The primary goal of motion planning is to generate safe and efficient trajectories for vehicles. Traditionally, motion planning models are trained using imitation learning to mimic the behavior of human experts. However, these models often…

Traffic simulators act as an essential component in the operating and planning of transportation systems. Conventional traffic simulators usually employ a calibrated physical car-following model to describe vehicles' behaviors and their…

Artificial Intelligence · Computer Science 2022-07-12 Guanjie Zheng , Hanyang Liu , Kai Xu , Zhenhui Li