中文
相关论文

相关论文: SENDD: Sparse Efficient Neural Depth and Deformati…

200 篇论文

Bridging the gap between data-rich training regimes and observation-sparse deployment conditions remains a central challenge in spatiotemporal field reconstruction, particularly when target domains exhibit distributional shifts,…

机器学习 · 计算机科学 2026-01-30 Xingyue Zhang , Yuxuan Bao , Mars Liyao Gao , J. Nathan Kutz

Objective: The computation of anatomical information and laparoscope position is a fundamental block of surgical navigation in Minimally Invasive Surgery (MIS). Recovering a dense 3D structure of surgical scene using visual cues remains a…

计算机视觉与模式识别 · 计算机科学 2022-11-29 Ruofeng Wei , Bin Li , Hangjie Mo , Bo Lu , Yonghao Long , Bohan Yang , Qi Dou , Yunhui Liu , Dong Sun

The rapid increment of morbidity of brain stroke in the last few years have been a driving force towards fast and accurate segmentation of stroke lesions from brain MRI images. With the recent development of deep-learning, computer-aided…

图像与视频处理 · 电气工程与系统科学 2021-10-25 Hritam Basak , Rukhshanda Hussain , Ajay Rana

Robotic automation in surgery requires precise tracking of surgical tools and mapping of deformable tissue. Previous works on surgical perception frameworks require significant effort in developing features for surgical tool and tissue…

机器人学 · 计算机科学 2021-03-26 Jingpei Lu , Ambareesh Jayakumari , Florian Richter , Yang Li , Michael C. Yip

Real-time reconstruction of deformable surgical scenes is vital for advancing robotic surgery, improving surgeon guidance, and enabling automation. Recent methods achieve dense reconstructions from da Vinci robotic surgery videos, with…

计算机视觉与模式识别 · 计算机科学 2026-02-23 Tianyi Song , Danail Stoyanov , Evangelos Mazomenos , Francisco Vasconcelos

Neural networks that map 3D coordinates to signed distance function (SDF) or occupancy values have enabled high-fidelity implicit representations of object shape. This paper develops a new shape model that allows synthesizing novel distance…

计算机视觉与模式识别 · 计算机科学 2021-12-07 Ehsan Zobeidi , Nikolay Atanasov

Depth sensing is a critical function for robotic tasks such as localization, mapping and obstacle detection. There has been a significant and growing interest in depth estimation from a single RGB image, due to the relatively low cost and…

计算机视觉与模式识别 · 计算机科学 2019-03-11 Diana Wofk , Fangchang Ma , Tien-Ju Yang , Sertac Karaman , Vivienne Sze

3D object detection using point cloud (PC) data is essential for perception pipelines of autonomous driving, where efficient encoding is key to meeting stringent resource and latency requirements. PointPillars, a widely adopted bird's-eye…

硬件体系结构 · 计算机科学 2024-01-17 Minjae Lee , Seongmin Park , Hyungmin Kim , Minyong Yoon , Janghwan Lee , Jun Won Choi , Nam Sung Kim , Mingu Kang , Jungwook Choi

Tissue deformation poses a key challenge for accurate surgical scene reconstruction. Despite yielding high reconstruction quality, existing methods suffer from slow rendering speeds and long training times, limiting their intraoperative…

计算机视觉与模式识别 · 计算机科学 2024-05-31 Shuojue Yang , Qian Li , Daiyun Shen , Bingchen Gong , Qi Dou , Yueming Jin

The last several years have seen significant progress in using depth cameras for tracking articulated objects such as human bodies, hands, and robotic manipulators. Most approaches focus on tracking skeletal parameters of a fixed shape…

计算机视觉与模式识别 · 计算机科学 2017-11-23 Aaron Walsman , Weilin Wan , Tanner Schmidt , Dieter Fox

In minimal invasive surgery, it is important to rebuild and visualize the latest deformed shape of soft-tissue surfaces to mitigate tissue damages. This paper proposes an innovative Simultaneous Localization and Mapping (SLAM) algorithm for…

计算机视觉与模式识别 · 计算机科学 2020-03-25 Jingwei Song , Jun Wang , Liang Zhao , Shoudong Huang , Gamini Dissanayake

Precise camera tracking, high-fidelity 3D tissue reconstruction, and real-time online visualization are critical for intrabody medical imaging devices such as endoscopes and capsule robots. However, existing SLAM (Simultaneous Localization…

计算机视觉与模式识别 · 计算机科学 2024-03-25 Kailing Wang , Chen Yang , Yuehao Wang , Sikuang Li , Yan Wang , Qi Dou , Xiaokang Yang , Wei Shen

3D reconstruction of highly deformable surfaces (e.g. cloths) from monocular RGB videos is a challenging problem, and no solution provides a consistent and accurate recovery of fine-grained surface details. To account for the ill-posed…

图形学 · 计算机科学 2025-03-27 Navami Kairanda , Marc Habermann , Shanthika Naik , Christian Theobalt , Vladislav Golyanik

Low signal-to-noise ratio videos -- such as those from underwater sonar, ultrasound, and microscopy -- pose significant challenges for computer vision models, particularly when paired clean imagery is unavailable. We present Spatiotemporal…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Suzanne Stathatos , Michael Hobley , Pietro Perona , Markus Marks

Blood vessel segmentation is crucial for many diagnostic and research applications. In recent years, CNN-based models have leaded to breakthroughs in the task of segmentation, however, such methods usually lose high-frequency information…

图像与视频处理 · 电气工程与系统科学 2021-04-09 Mo Zhang , Fei Yu , Jie Zhao , Li Zhang , Quanzheng Li

Pillar-based 3D object detection has gained traction in self-driving technology due to its speed and accuracy facilitated by the artificial densification of pillars for GPU-friendly processing. However, dense pillar processing fundamentally…

计算机视觉与模式识别 · 计算机科学 2024-08-27 Seongmin Park , Minjae Lee , Junwon Choi , Jungwook Choi

Multi-view stereo (MVS) is the golden mean between the accuracy of active depth sensing and the practicality of monocular depth estimation. Cost volume based approaches employing 3D convolutional neural networks (CNNs) have considerably…

计算机视觉与模式识别 · 计算机科学 2020-08-26 Ayan Sinha , Zak Murez , James Bartolozzi , Vijay Badrinarayanan , Andrew Rabinovich

Monocular depth estimation (MDE) plays a pivotal role in various computer vision applications, such as robotics, augmented reality, and autonomous driving. Despite recent advancements, existing methods often fail to meet key requirements…

计算机视觉与模式识别 · 计算机科学 2025-09-29 Andrii Litvynchuk , Ivan Livinsky , Anand Ravi , Nima Kalantari , Andrii Tsarov

Convolutional networks are the de-facto standard for analyzing spatio-temporal data such as images, videos, and 3D shapes. Whilst some of this data is naturally dense (e.g., photos), many other data sources are inherently sparse. Examples…

计算机视觉与模式识别 · 计算机科学 2017-11-29 Benjamin Graham , Martin Engelcke , Laurens van der Maaten

The prediction of upcoming events in industrial processes has been a long-standing research goal since it enables optimization of manufacturing parameters, planning of equipment maintenance and more importantly prediction and eventually…