中文
相关论文

相关论文: BiTAA: A Bi-Task Adversarial Attack for Object Det…

200 篇论文

Generalizable perception is one of the pillars of high-level autonomy in space robotics. Estimating the structure and motion of unknown objects in dynamic environments is fundamental for such autonomous systems. Traditionally, the solutions…

机器人学 · 计算机科学 2024-11-26 Kuldeep R Barad , Antoine Richard , Jan Dentler , Miguel Olivares-Mendez , Carol Martinez

We present VicaSplat, a novel framework for joint 3D Gaussians reconstruction and camera pose estimation from a sequence of unposed video frames, which is a critical yet underexplored task in real-world 3D applications. The core of our…

计算机视觉与模式识别 · 计算机科学 2025-03-14 Zhiqi Li , Chengrui Dong , Yiming Chen , Zhangchi Huang , Peidong Liu

3D dense captioning is a task involving the localization of objects and the generation of descriptions for each object in a 3D scene. Recent approaches have attempted to incorporate contextual information by modeling relationships with…

计算机视觉与模式识别 · 计算机科学 2024-08-14 Minjung Kim , Hyung Suk Lim , Soonyoung Lee , Bumsoo Kim , Gunhee Kim

In recent years, many deep learning models have been adopted in autonomous driving. At the same time, these models introduce new vulnerabilities that may compromise the safety of autonomous vehicles. Specifically, recent studies have…

计算机视觉与模式识别 · 计算机科学 2021-10-22 Jindi Zhang , Yang Lou , Jianping Wang , Kui Wu , Kejie Lu , Xiaohua Jia

Deep neural networks are found to be prone to adversarial examples which could deliberately fool the model to make mistakes. Recently, a few of works expand this task from 2D image to 3D point cloud by using global point cloud optimization.…

计算机视觉与模式识别 · 计算机科学 2021-10-12 Yiming Sun , Feng Chen , Zhiyu Chen , Mingjie Wang

This paper tackles the problem of generalizable 3D-aware generation from monocular datasets, e.g., ImageNet. The key challenge of this task is learning a robust 3D-aware representation without multi-view or dynamic data, while ensuring…

计算机视觉与模式识别 · 计算机科学 2025-03-12 Yuxin Wang , Qianyi Wu , Dan Xu

Compared with real-time multi-object tracking (MOT), offline multi-object tracking (OMOT) has the advantages to perform 2D-3D detection fusion, erroneous link correction, and full track optimization but has to deal with the challenges from…

计算机视觉与模式识别 · 计算机科学 2025-03-19 Kemiao Huang , Yinqi Chen , Meiying Zhang , Qi Hao

Self-supervised learning of depth has been a highly studied topic of research as it alleviates the requirement of having ground truth annotations for predicting depth. Depth is learnt as an intermediate solution to the task of view…

计算机视觉与模式识别 · 计算机科学 2021-03-02 Vinay Kaushik , Kartik Jindgar , Brejesh Lall

3D Gaussian Splatting (3DGS) has emerged as a powerful paradigm for real-time and high-fidelity 3D reconstruction from posed images. However, recent studies reveal its vulnerability to adversarial corruptions in input views, where…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Yiran Qiao , Yiren Lu , Yunlai Zhou , Rui Yang , Linlin Hou , Yu Yin , Jing Ma

Guided image synthesis methods, like SDEdit based on the diffusion model, excel at creating realistic images from user inputs such as stroke paintings. However, existing efforts mainly focus on image quality, often overlooking a key point:…

计算机视觉与模式识别 · 计算机科学 2024-02-07 Qi Zhou , Dongxia Wang , Tianlin Li , Zhihong Xu , Yang Liu , Kui Ren , Wenhai Wang , Qing Guo

In this work, we propose a novel method to supervise 3D Gaussian Splatting (3DGS) scenes using optical tactile sensors. Optical tactile sensors have become widespread in their use in robotics for manipulation and object representation;…

机器人学 · 计算机科学 2024-08-19 Aiden Swann , Matthew Strong , Won Kyung Do , Gadiel Sznaier Camps , Mac Schwager , Monroe Kennedy

This paper proposes an online multi-camera multi-object tracker that only requires monocular detector training, independent of the multi-camera configurations, allowing seamless extension/deletion of cameras without retraining effort. The…

计算机视觉与模式识别 · 计算机科学 2020-10-28 Jonah Ong , Ba Tuong Vo , Ba Ngu Vo , Du Yong Kim , Sven Nordholm

Autonomous vehicles rely on LiDAR sensors to detect obstacles such as pedestrians, other vehicles, and fixed infrastructures. LiDAR spoofing attacks have been demonstrated that either create erroneous obstacles or prevent detection of real…

系统与控制 · 电气工程与系统科学 2023-02-16 Hongchao Zhang , Zhouchi Li , Shiyu Cheng , Andrew Clark

3D semantic occupancy prediction is a pivotal task in autonomous driving, providing a dense and fine-grained understanding of the surrounding environment, yet single-modality methods face trade-offs between camera semantics and LiDAR…

计算机视觉与模式识别 · 计算机科学 2026-02-02 A. Enes Doruk , Hasan F. Ates

Implicit neural representations and 3D Gaussian splatting (3DGS) have shown great potential for scene reconstruction. Recent studies have expanded their applications in autonomous reconstruction through task assignment methods. However,…

机器人学 · 计算机科学 2024-12-04 Jing Zeng , Qi Ye , Tianle Liu , Yang Xu , Jin Li , Jinming Xu , Liang Li , Jiming Chen

Autonomous driving systems (ADS) increasingly rely on deep learning-based perception models, which remain vulnerable to adversarial attacks. In this paper, we revisit adversarial attacks and defense methods, focusing on road sign…

机器人学 · 计算机科学 2025-05-26 Cheng Chen , Yuhong Wang , Nafis S Munir , Xiangwei Zhou , Xugui Zhou

Diffusion models (DMs) embark a new era of generative modeling and offer more opportunities for efficient generating high-quality and realistic data samples. However, their widespread use has also brought forth new challenges in model…

计算机视觉与模式识别 · 计算机科学 2024-10-14 Jingyao Xu , Yuetong Lu , Yandong Li , Siyang Lu , Dongdong Wang , Xiang Wei

Real-time processing is crucial in autonomous driving systems due to the imperative of instantaneous decision-making and rapid response. In real-world scenarios, autonomous vehicles are continuously tasked with interpreting their…

计算机视觉与模式识别 · 计算机科学 2024-03-07 Wonhyeok Choi , Mingyu Shin , Hyukzae Lee , Jaehoon Cho , Jaehyeon Park , Sunghoon Im

We present VISAT, a novel open dataset and benchmarking suite for evaluating model robustness in the task of traffic sign recognition with the presence of visual attributes. Built upon the Mapillary Traffic Sign Dataset (MTSD), our dataset…

密码学与安全 · 计算机科学 2025-11-03 Simon Yu , Peilin Yu , Hongbo Zheng , Huajie Shao , Han Zhao , Lui Sha

Due to the sophisticated imaging process, an identical scene captured by different cameras could exhibit distinct imaging patterns, introducing distinct proficiency among the super-resolution (SR) models trained on images from different…

图像与视频处理 · 电气工程与系统科学 2022-05-10 Xiaoqian Xu , Pengxu Wei , Weikai Chen , Mingzhi Mao , Liang Lin , Guanbin Li