English
Related papers

Related papers: Geo-RepNet: Geometry-Aware Representation Learning…

200 papers

Recently, stereo vision based on lightweight RGBD cameras has been widely used in various fields. However, limited by the imaging principles, the commonly used RGB-D cameras based on TOF, structured light, or binocular vision acquire some…

Computer Vision and Pattern Recognition · Computer Science 2023-09-06 Dongyue Chen , Tingxuan Huang , Zhimin Song , Shizhuo Deng , Tong Jia

We present a novel end-to-end framework named as GSNet (Geometric and Scene-aware Network), which jointly estimates 6DoF poses and reconstructs detailed 3D car shapes from single urban street view. GSNet utilizes a unique four-way feature…

Computer Vision and Pattern Recognition · Computer Science 2020-07-28 Lei Ke , Shichao Li , Yanan Sun , Yu-Wing Tai , Chi-Keung Tang

Segmentation has been a major task in neuroimaging. A large number of automated methods have been developed for segmenting healthy and diseased brain tissues. In recent years, deep learning techniques have attracted a lot of attention as a…

Image and Video Processing · Electrical Eng. & Systems 2019-07-05 Jimit Doshi , Guray Erus , Mohamad Habes , Christos Davatzikos

Endoscopic Submucosal Dissection (ESD) is a minimally invasive procedure initially developed for early gastric cancer treatment and has expanded to address diverse gastrointestinal lesions. While computer-assisted surgery (CAS) systems…

Computer Vision and Pattern Recognition · Computer Science 2025-08-12 Xiangning Zhang , Qingwei Zhang , Jinnan Chen , Chengfeng Zhou , Yaqi Wang , Zhengjie Zhang , Xiaobo Li , Dahong Qian

Purpose: To propose a novel deep learning-based method called RG-Net (reconstruction and generation network) for highly accelerated MR parametric mapping by undersampling k-space and reducing the acquired contrast number simultaneously.…

Image and Video Processing · Electrical Eng. & Systems 2021-12-03 Yanjie Zhu , Haoxiang Li , Yuanyuan Liu , Muzi Guo , Guanxun Cheng , Gang Yang , Haifeng Wang , Dong Liang

Automatic recognition and segmentation methods now become the essential requirement in identifying co-seismic landslides, which are fundamental for disaster assessment and mitigation in large-scale earthquakes. This approach used to be…

Computer Vision and Pattern Recognition · Computer Science 2023-10-27 Qingsong Xu , Chaojun Ouyang , Tianhai Jiang , Xuanmei Fan , Duoxiang Cheng

Robust gait recognition requires highly discriminative representations, which are closely tied to input modalities. While binary silhouettes and skeletons have dominated recent literature, these 2D representations fall short of capturing…

Computer Vision and Pattern Recognition · Computer Science 2025-08-06 Xinzhu Li , Juepeng Zheng , Yikun Chen , Xudong Mao , Guanghui Yue , Wei Zhou , Chenlei Lv , Ruomei Wang , Fan Zhou , Baoquan Zhao

Accurate and robust tracking and reconstruction of the surgical scene is a critical enabling technology toward autonomous robotic surgery. Existing algorithms for 3D perception in surgery mainly rely on geometric information, while we…

Image and Video Processing · Electrical Eng. & Systems 2023-02-21 Shan Lin , Albert J. Miao , Jingpei Lu , Shunkai Yu , Zih-Yun Chiu , Florian Richter , Michael C. Yip

Multimodal perception systems for robotics and embodied AI often assume reliable RGB-D sensing, but in practice, depth is frequently missing, noisy, or corrupted. We thus present GeomPrompt, a lightweight cross-modal adaptation module that…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Krishna Jaganathan , Patricio Vela

Conventional deep learning-based image reconstruction methods require a large amount of training data which can be hard to obtain in practice. Untrained deep learning methods overcome this limitation by training a network to invert a…

Image and Video Processing · Electrical Eng. & Systems 2024-07-09 Carlos Osorio Quero , Daniel Leykam , Irving Rondon Ojeda

Depth estimation from monocular endoscopic images presents significant challenges due to the complexity of endoscopic surgery, such as irregular shapes of human soft tissues, as well as variations in lighting conditions. Existing methods…

Image and Video Processing · Electrical Eng. & Systems 2025-02-07 Dawei Lu , Deqiang Xiao , Danni Ai , Jingfan Fan , Tianyu Fu , Yucong Lin , Hong Song , Xujiong Ye , Lei Zhang , Jian Yang

Segmentation is a fundamental task in medical image analysis. However, most existing methods focus on primary region extraction and ignore edge information, which is useful for obtaining accurate segmentation. In this paper, we propose a…

Computer Vision and Pattern Recognition · Computer Science 2019-07-26 Zhijie Zhang , Huazhu Fu , Hang Dai , Jianbing Shen , Yanwei Pang , Ling Shao

Purpose: To systematically investigate the influence of various data consistency layers, (semi-)supervised learning and ensembling strategies, defined in a $\Sigma$-net, for accelerated parallel MR image reconstruction using deep learning.…

Image and Video Processing · Electrical Eng. & Systems 2019-12-20 Kerstin Hammernik , Jo Schlemper , Chen Qin , Jinming Duan , Ronald M. Summers , Daniel Rueckert

Machine Learning surrogates for Computational Fluid Dynamics (CFD), particularly Graph Neural Networks (GNNs) and Transformers, have become a new important approach for accelerating physics simulations. However, we identify a critical…

Machine Learning · Computer Science 2026-05-05 Paul Garnier , Vincent Lannelongue , Elie Hachem

Multi-organ segmentation is a critical task in computer-aided diagnosis. While recent deep learning methods have achieved remarkable success in image segmentation, huge variations in organ size and shape challenge their effectiveness in…

Image and Video Processing · Electrical Eng. & Systems 2025-10-31 Xizhi Tian , Changjun Zhou , Yulin. Yang

Occlusion and the scarcity of labeled surgical data are significant challenges in disparity estimation for stereo laparoscopic images. To address these issues, this study proposes a Depth Guided Occlusion-Aware Disparity Refinement Network…

Computer Vision and Pattern Recognition · Computer Science 2025-05-14 Ziteng Liu , Dongdong He , Chenghong Zhang , Wenpeng Gao , Yili Fu

Vision foundation models (VFMs) have emerged as powerful tools for surgical scene understanding. However, current approaches predominantly rely on unimodal RGB pre-training, overlooking the complex 3D geometry inherent to surgical…

Computer Vision and Pattern Recognition · Computer Science 2026-01-28 John J. Han , Adam Schmidt , Muhammad Abdullah Jamal , Chinedu Nwoye , Anita Rau , Jie Ying Wu , Omid Mohareri

SEGSRNet addresses the challenge of precisely identifying surgical instruments in low-resolution stereo endoscopic images, a common issue in medical imaging and robotic surgery. Our innovative framework enhances image clarity and…

Image and Video Processing · Electrical Eng. & Systems 2024-04-29 Mansoor Hayat , Supavadee Aramvith , Titipat Achakulvisut

Deep learning approaches have achieved highly accurate face recognition by training the models with very large face image datasets. Unlike the availability of large 2D face image datasets, there is a lack of large 3D face datasets available…

Computer Vision and Pattern Recognition · Computer Science 2021-12-23 Meng-Tzu Chiu , Hsun-Ying Cheng , Chien-Yi Wang , Shang-Hong Lai

Exploiting internal spatial geometric constraints of sparse LiDARs is beneficial to depth completion, however, has been not explored well. This paper proposes an efficient method to learn geometry-aware embedding, which encodes the local…

Computer Vision and Pattern Recognition · Computer Science 2022-06-02 Wenchao Du , Hu Chen , Hongyu Yang , Yi Zhang