中文
相关论文

相关论文: Enhancing Multi-Camera Gymnast Tracking Through Do…

200 篇论文

Achieving robust and precise pose estimation in dynamic scenes is a significant research challenge in Visual Simultaneous Localization and Mapping (SLAM). Recent advancements integrating Gaussian Splatting into SLAM systems have proven…

机器人学 · 计算机科学 2024-11-14 Yueming Xu , Haochen Jiang , Zhongyang Xiao , Jianfeng Feng , Li Zhang

The filming of sporting events projects and flattens the movement of athletes in the world onto a 2D broadcast image. The pixel locations of joints in these images can be detected with high validity. Recovering the actual 3D movement of the…

计算机视觉与模式识别 · 计算机科学 2023-04-11 Tobias Baumgartner , Stefanie Klatt

Simultaneous localization and mapping (SLAM) has achieved impressive performance in static environments. However, SLAM in dynamic environments remains an open question. Many methods directly filter out dynamic objects, resulting in…

机器人学 · 计算机科学 2024-11-26 Haoang Li , Xiangqi Meng , Xingxing Zuo , Zhe Liu , Hesheng Wang , Daniel Cremers

This paper addresses the problem of 3D pose estimation for multiple people in a few calibrated camera views. The main challenge of this problem is to find the cross-view correspondences among noisy and incomplete 2D pose predictions. Most…

计算机视觉与模式识别 · 计算机科学 2019-01-15 Junting Dong , Wen Jiang , Qixing Huang , Hujun Bao , Xiaowei Zhou

We present TrackGS, a novel method to integrate global feature tracks with 3D Gaussian Splatting (3DGS) for COLMAP-free novel view synthesis. While 3DGS delivers impressive rendering quality, its reliance on accurate precomputed camera…

计算机视觉与模式识别 · 计算机科学 2025-11-24 Dongbo Shi , Shen Cao , Lubin Fan , Bojian Wu , Jinhui Guo , Ligang Liu , Renjie Chen

Multi-camera 3D object detection (MC3D) has attracted increasing attention with the growing deployment of multi-sensor physical agents, such as robots and autonomous vehicles. However, MC3D models still struggle to generalize to unseen…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Zhaonian Kuang , Rui Ding , Haotian Wang , Xinhu Zheng , Meng Yang , Gang Hua

We present the first application of 3D Gaussian Splatting in monocular SLAM, the most fundamental but the hardest setup for Visual SLAM. Our method, which runs live at 3fps, utilises Gaussians as the only 3D representation, unifying the…

计算机视觉与模式识别 · 计算机科学 2024-04-16 Hidenobu Matsuki , Riku Murai , Paul H. J. Kelly , Andrew J. Davison

We present a novel method to estimate the motion matrix between overlapping pairs of 3D views in the context of indoor scenes. We use the Manhattan world assumption to introduce lightweight geometric constraints under the form of planes…

计算机视觉与模式识别 · 计算机科学 2020-01-22 Adrien Kaiser , José Alonso Ybanez Zepeda , Tamy Boubekeur

In this paper, we propose a novel approach to address the problem of camera and radar sensor fusion for 3D object detection in autonomous vehicle perception systems. Our approach builds on recent advances in deep learning and leverages the…

计算机视觉与模式识别 · 计算机科学 2024-04-26 Daniel Dworak , Mateusz Komorkiewicz , Paweł Skruch , Jerzy Baranowski

This paper addresses the gaze target detection problem in single images captured from the third-person perspective. We present a multimodal deep architecture to infer where a person in a scene is looking. This spatial model is trained on…

计算机视觉与模式识别 · 计算机科学 2022-08-24 Francesco Tonini , Cigdem Beyan , Elisa Ricci

Detecting and localizing glass in 3D environments poses significant challenges for visual perception systems, as the optical properties of glass often hinder conventional sensors from accurately distinguishing glass surfaces. The lack of…

机器人学 · 计算机科学 2025-09-09 Kai Zhang , Guoyang Zhao , Jianxing Shi , Bonan Liu , Weiqing Qi , Jun Ma

Medical image synthesis generates additional imaging modalities that are costly, invasive or harmful to acquire, which helps to facilitate the clinical workflow. When training pairs are substantially misaligned (e.g., lung MRI-CT pairs with…

图像与视频处理 · 电气工程与系统科学 2024-08-20 Bowen Xin , Tony Young , Claire E Wainwright , Tamara Blake , Leo Lebrat , Thomas Gaass , Thomas Benkert , Alto Stemmer , David Coman , Jason Dowling

The 3D Gaussian Splatting (3DGS)-based SLAM system has garnered widespread attention due to its excellent performance in real-time high-fidelity rendering. However, in real-world environments with dynamic objects, existing 3DGS-based SLAM…

机器人学 · 计算机科学 2025-02-19 Mingrui Li , Weijian Chen , Na Cheng , Jingyuan Xu , Dong Li , Hongyu Wang

Tracking in gigapixel scenarios holds numerous potential applications in video surveillance and pedestrian analysis. Existing algorithms attempt to perform tracking in crowded scenes by utilizing multiple cameras or group relationships.…

计算机视觉与模式识别 · 计算机科学 2024-07-29 Yunqi Zhao , Yuchen Guo , Zheng Cao , Kai Ni , Ruqi Huang , Lu Fang

Gait recognition is emerging as a promising and innovative area within the field of computer vision, widely applied to remote person identification. Although existing gait recognition methods have achieved substantial success in controlled…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Zhengxian Wu , Chuanrui Zhang , Hangrui Xu , Peng Jiao , Haoqian Wang

Visual SLAM systems targeting static scenes have been developed with satisfactory accuracy and robustness. Dynamic 3D object tracking has then become a significant capability in visual SLAM with the requirement of understanding dynamic…

计算机视觉与模式识别 · 计算机科学 2022-10-06 Hanwei Zhang , Hideaki Uchiyama , Shintaro Ono , Hiroshi Kawasaki

In this paper, based on the assumption that the object boundaries (e.g., buildings) from the over-view data should coincide with footprints of fa\c{c}ade 3D points generated from street-view photogrammetric images, we aim to address this…

计算机视觉与模式识别 · 计算机科学 2022-02-15 Xiao Ling , Rongjun Qin

Multi-camera multiple people tracking has become an increasingly important area of research due to the growing demand for accurate and efficient indoor people tracking systems, particularly in settings such as retail, healthcare centers,…

3D-aware image synthesis aims to generate images of objects from multiple views by learning a 3D representation. However, one key challenge remains: existing approaches lack geometry constraints, hence usually fail to generate multi-view…

计算机视觉与模式识别 · 计算机科学 2022-04-14 Xuanmeng Zhang , Zhedong Zheng , Daiheng Gao , Bang Zhang , Pan Pan , Yi Yang

Modern deep learning developments create new opportunities for 3D mapping technology, scene reconstruction pipelines, and virtual reality development. Despite advances in 3D deep learning technology, direct training of deep learning models…

计算机视觉与模式识别 · 计算机科学 2025-09-03 Xueyang Kang