中文
相关论文

相关论文: MapNeXt: Revisiting Training and Scaling Practices…

200 篇论文

Vision Transformers have shown great potential in computer vision tasks. Most recent works have focused on elaborating the spatial token mixer for performance gains. However, we observe that a well-designed general architecture can…

计算机视觉与模式识别 · 计算机科学 2023-12-25 Fangjian Lin , Jianlong Yuan , Sitong Wu , Fan Wang , Zhibin Wang

Accurately and efficiently extracting building footprints from a wide range of remote sensed imagery remains a challenge due to their complex structure, variety of scales and diverse appearances. Existing convolutional neural network…

计算机视觉与模式识别 · 计算机科学 2020-10-01 Qing Zhu , Cheng Liao , Han Hu , Xiaoming Mei , Haifeng Li

We present a framework for dynamic trajectory generation for autonomous navigation, which does not rely on HD maps as the underlying representation. High Definition (HD) maps have become a key component in most autonomous driving…

机器人学 · 计算机科学 2022-03-29 David Paz , Hao Xiang , Andrew Liang , Henrik I. Christensen

Robust high-definition (HD) map construction is vital for autonomous driving, yet existing methods often struggle with incomplete multi-view camera data. This paper presents SafeMap, a novel framework specifically designed to secure…

计算机视觉与模式识别 · 计算机科学 2025-07-02 Xiaoshuai Hao , Lingdong Kong , Rong Yin , Pengwei Wang , Jing Zhang , Yunfeng Diao , Shu Zhao

Vectorized High-Definition (HD) map construction requires predictions of the category and point coordinates of map elements (e.g. road boundary, lane divider, pedestrian crossing, etc.). State-of-the-art methods are mainly based on…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Yi Zhou , Hui Zhang , Jiaqian Yu , Yifan Yang , Sangil Jung , Seung-In Park , ByungIn Yoo

Maps provide robots with crucial environmental knowledge, thereby enabling them to perform interactive tasks effectively. Easily accessing accurate abstract-to-detailed geometric and semantic concepts from maps is crucial for robots to make…

机器人学 · 计算机科学 2024-03-26 Tianshuai Hu , Jianhao Jiao , Yucheng Xu , Hongji Liu , Sheng Wang , Ming Liu

Next-generation multimodal foundation models capable of any-to-any cross-modal generation and multi-turn interaction will serve as core components of artificial general intelligence systems, playing a pivotal role in human-machine…

计算与语言 · 计算机科学 2025-10-17 Run Luo , Xiaobo Xia , Lu Wang , Longze Chen , Renke Shan , Jing Luo , Min Yang , Tat-Seng Chua

Safe and efficient path planning is crucial for autonomous mobile robots. A prerequisite for path planning is to have a comprehensive understanding of the 3D structure of the robot's environment. On MAVs this is commonly achieved using…

机器人学 · 计算机科学 2019-11-15 Marius Fehr , Tim Taubner , Yang Liu , Roland Siegwart , Cesar Cadena

The ever-growing demand and complexity of machine learning are putting pressure on hyper-parameter tuning systems: while the evaluation cost of models continues to increase, the scalability of state-of-the-arts starts to become a crucial…

机器学习 · 计算机科学 2022-01-19 Yang Li , Yu Shen , Huaijun Jiang , Wentao Zhang , Jixiang Li , Ji Liu , Ce Zhang , Bin Cui

Implicit Neural Representations (INRs) have emerged as a powerful paradigm for representing signals such as images, 3D shapes, signed distance fields, and radiance fields. While significant progress has been made in architecture design…

人工智能 · 计算机科学 2026-04-10 Plein Versace

Tactile sensing is critical in advanced interactive systems by emulating the human sense of touch to detect stimuli. Vision-based tactile sensors are promising for providing multimodal capabilities and high robustness, yet existing…

机器人学 · 计算机科学 2025-04-07 Mayue Shi , Yongqi Zhang , Xiaotong Guo , Eric M. Yeatman

Recent advancements in vision transformers (ViTs) have demonstrated that larger models often achieve superior performance. However, training these models remains computationally intensive and costly. To address this challenge, we introduce…

计算机视觉与模式识别 · 计算机科学 2025-10-23 Zhiwei Hao , Jianyuan Guo , Li Shen , Kai Han , Yehui Tang , Han Hu , Yunhe Wang

Autonomous driving needs various line-of-sight sensors to perceive surroundings that could be impaired under diverse environment uncertainties such as visual occlusion and extreme weather. To improve driving safety, we explore to wirelessly…

网络与互联网体系结构 · 计算机科学 2020-12-21 Qiang Liu , Tao Han , Jiang , Xie , BaekGyu Kim

Over the last few years, neural image compression has gained wide attention from research and industry, yielding promising end-to-end deep neural codecs outperforming their conventional counterparts in rate-distortion performance. Despite…

计算机视觉与模式识别 · 计算机科学 2023-07-14 Ahmed Ghorbel , Wassim Hamidouche , Luce Morin

Video saliency prediction has recently attracted attention of the research community, as it is an upstream task for several practical applications. However, current solutions are particularly computationally demanding, especially due to the…

计算机视觉与模式识别 · 计算机科学 2023-01-12 Feiyan Hu , Simone Palazzo , Federica Proietto Salanitri , Giovanni Bellitto , Morteza Moradi , Concetto Spampinato , Kevin McGuinness

The ability to autonomously navigate in unknown environments is important for mobile robots. The map is the core component to achieve this. Most map representations rely on drift-free state estimation and provide a global metric map to…

机器人学 · 计算机科学 2021-09-21 Xuecheng Xu , Cheng Wang , Yue Wang , Rong Xiong

Can an autonomous agent navigate in a new environment without building an explicit map? For the task of PointGoal navigation ('Go to $\Delta x$, $\Delta y$') under idealized settings (no RGB-D and actuation noise, perfect GPS+Compass), the…

计算机视觉与模式识别 · 计算机科学 2022-06-08 Ruslan Partsey , Erik Wijmans , Naoki Yokoyama , Oles Dobosevych , Dhruv Batra , Oleksandr Maksymets

The construction of Vectorized High-Definition (HD) map typically requires capturing both category and geometry information of map elements. Current state-of-the-art methods often adopt solely either point-level or instance-level…

计算机视觉与模式识别 · 计算机科学 2025-06-13 Jing Yang , Minyue Jiang , Sen Yang , Xiao Tan , Yingying Li , Errui Ding , Hanli Wang , Jingdong Wang

UNet and its latest extensions like TransUNet have been the leading medical image segmentation methods in recent years. However, these networks cannot be effectively adopted for rapid image segmentation in point-of-care applications as they…

图像与视频处理 · 电气工程与系统科学 2022-03-11 Jeya Maria Jose Valanarasu , Vishal M. Patel

Existing AGR navigation systems have advanced in lightly occluded scenarios (e.g., buildings) by employing 3D semantic scene completion networks for voxel occupancy prediction and constructing Euclidean Signed Distance Field (ESDF) maps for…