中文
相关论文

相关论文: ViP3D: End-to-end Visual Trajectory Prediction via…

200 篇论文

Autonomous vehicles (AVs) rely on sensors and deep neural networks (DNNs) to perceive their surrounding environment and make maneuver decisions in real time. However, achieving real-time DNN inference in the AV's perception pipeline is…

计算机视觉与模式识别 · 计算机科学 2026-02-12 Liangkai Liu , Kang G. Shin , Jinkyu Lee , Chengmo Yang , Weisong Shi

When a vision model performs image recognition, which visual attributes drive its predictions? Detecting unintended reliance on specific visual features is critical for ensuring model robustness, preventing overfitting, and avoiding…

计算机视觉与模式识别 · 计算机科学 2025-11-20 Christy Li , Josep Lopez Camuñas , Jake Thomas Touchet , Jacob Andreas , Agata Lapedriza , Antonio Torralba , Tamar Rott Shaham

Vehicle 3D extents and trajectories are critical cues for predicting the future location of vehicles and planning future agent ego-motion based on those predictions. In this paper, we propose a novel online framework for 3D vehicle…

计算机视觉与模式识别 · 计算机科学 2019-09-13 Hou-Ning Hu , Qi-Zhi Cai , Dequan Wang , Ji Lin , Min Sun , Philipp Krähenbühl , Trevor Darrell , Fisher Yu

The comprehension of environmental traffic situation largely ensures the driving safety of autonomous vehicles. Recently, the mission has been investigated by plenty of researches, while it is hard to be well addressed due to the limitation…

计算机视觉与模式识别 · 计算机科学 2020-01-09 Yanliang Zhu , Deheng Qian , Dongchun Ren , Huaxia Xia

Predicting the future motion of dynamic agents is of paramount importance to ensuring safety and assessing risks in motion planning for autonomous robots. In this study, we propose a two-stage motion prediction method, called R-Pred,…

计算机视觉与模式识别 · 计算机科学 2023-07-17 Sehwan Choi , Jungho Kim , Junyong Yun , Jun Won Choi

Object detection is the central issue of intelligent traffic systems, and recent advancements in single-vehicle lidar-based 3D detection indicate that it can provide accurate position information for intelligent agents to make decisions and…

人工智能 · 计算机科学 2023-10-11 Caizhen He , Hai Wang , Long Chen , Tong Luo , Yingfeng Cai

Devising intelligent agents able to live in an environment and learn by observing the surroundings is a longstanding goal of Artificial Intelligence. From a bare Machine Learning perspective, challenges arise when the agent is prevented…

计算机视觉与模式识别 · 计算机科学 2022-04-27 Matteo Tiezzi , Simone Marullo , Lapo Faggi , Enrico Meloni , Alessandro Betti , Stefano Melacci

A complex visual navigation task puts an agent in different situations which call for a diverse range of visual perception abilities. For example, to "go to the nearest chair", the agent might need to identify a chair in a living room using…

计算机视觉与模式识别 · 计算机科学 2021-08-05 Bokui Shen , Danfei Xu , Yuke Zhu , Leonidas J. Guibas , Li Fei-Fei , Silvio Savarese

We present an end-to-end method for object detection and trajectory prediction utilizing multi-view representations of LiDAR returns and camera images. In this work, we recognize the strengths and weaknesses of different view…

计算机视觉与模式识别 · 计算机科学 2021-10-20 Sudeep Fadadu , Shreyash Pandey , Darshan Hegde , Yi Shi , Fang-Chieh Chou , Nemanja Djuric , Carlos Vallespi-Gonzalez

Object tracking is a key challenge of computer vision with various applications that all require different architectures. Most tracking systems have limitations such as constraining all movement to a 2D plane and they often track only one…

计算机视觉与模式识别 · 计算机科学 2025-03-07 Lars Bredereke , Yale Hartmann , Tanja Schultz

We propose a robotic learning system for autonomous exploration and navigation in unexplored environments. We are motivated by the idea that even an unseen environment may be familiar from previous experiences in similar environments. The…

机器人学 · 计算机科学 2022-11-24 Huangying Zhan , Hamid Rezatofighi , Ian Reid

Object detection and forecasting are fundamental components of embodied perception. These two problems, however, are largely studied in isolation by the community. In this paper, we propose an end-to-end approach for detection and motion…

计算机视觉与模式识别 · 计算机科学 2022-04-01 Neehar Peri , Jonathon Luiten , Mengtian Li , Aljoša Ošep , Laura Leal-Taixé , Deva Ramanan

Occlusion is a major challenge for LiDAR-based object detection methods. This challenge becomes safety-critical in urban traffic where the ego vehicle must have reliable object detection to avoid collision while its field of view is…

机器人学 · 计算机科学 2023-09-20 Minh-Quan Dao , Julie Stephany Berrio , Vincent Frémont , Mao Shan , Elwan Héry , Stewart Worrall

Temporal information is crucial for detecting occluded instances. Existing temporal representations have progressed from BEV or PV features to more compact query features. Compared to these aforementioned features, predictions offer the…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Nan Peng , Xun Zhou , Mingming Wang , Xiaojun Yang , Songming Chen , Guisong Chen

Automating end-to-end data science pipeline with AI agents still stalls on two gaps: generating insightful, diverse visual evidence and assembling it into a coherent, professional report. We present A2P-Vis, a two-part, multi-agent pipeline…

机器学习 · 计算机科学 2025-12-29 Shuyu Gan , Renxiang Wang , James Mooney , Dongyeop Kang

Accurate prediction of driving scene is a challenging task due to uncertainty in sensor data, the complex behaviors of agents, and the possibility of multiple feasible futures. Existing prediction methods using occupancy grid maps primarily…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Rabbia Asghar , Lukas Rummelhard , Wenqian Liu , Anne Spalanzani , Christian Laugier

A unified neural network structure is presented for joint 3D object detection and point cloud segmentation in this paper. We leverage rich supervision from both detection and segmentation labels rather than using just one of them. In…

计算机视觉与模式识别 · 计算机科学 2021-11-16 Yuanxin Zhong , Minghan Zhu , Huei Peng

Vision-language models (VLMs) are increasingly being adopted for end-to-end autonomous driving systems due to their exceptional performance in handling long-tail scenarios. However, current VLM-based approaches suffer from two major…

机器人学 · 计算机科学 2026-03-31 Yuqi Ye , Zijian Zhang , Junhong Lin , Shangkun Sun , Changhao Peng , Wei Gao

Autonomous vehicle perception typically relies on modular pipelines that decompose the task into detection, tracking, and prediction. While interpretable, these pipelines suffer from error accumulation and limited inter-task synergy.…

计算机视觉与模式识别 · 计算机科学 2025-08-29 Loïc Stratil , Felix Fent , Esteban Rivera , Markus Lienkamp

Autonomous driving requires 3D perception of vehicles and other objects in the in environment. Much of the current methods support 2D vehicle detection. This paper proposes a flexible pipeline to adopt any 2D detection network and fuse it…

计算机视觉与模式识别 · 计算机科学 2018-03-02 Xinxin Du , Marcelo H. Ang , Sertac Karaman , Daniela Rus