中文
相关论文

相关论文: Spatiotemporal motion prediction in free-breathing…

200 篇论文

This paper introduces a learning-based modeling framework for a magnetically steerable soft suction device designed for endoscopic endonasal brain tumor resection. The device is miniaturized (4 mm outer diameter, 2 mm inner diameter, 40 mm…

机器人学 · 计算机科学 2026-02-06 Majid Roshanfar , Alex Zhang , Changyan He , Amir Hooshiar , Dale J. Podolsky , Thomas Looi , Eric Diller

Estimating the shape and motion state of the myocardium is essential in diagnosing cardiovascular diseases.However, cine magnetic resonance (CMR) imaging is dominated by 2D slices, whose large slice spacing challenges inter-slice shape…

计算机视觉与模式识别 · 计算机科学 2023-08-29 Xiaohan Yuan , Cong Liu , Yangang Wang

Temporal volume images with 3D+t (4D) information are often used in medical imaging to statistically analyze temporal dynamics or capture disease progression. Although deep-learning-based generative models for natural images have been…

图像与视频处理 · 电气工程与系统科学 2022-06-28 Boah Kim , Jong Chul Ye

The prevailing deep learning-based methods of predicting cardiac segmentation involve reconstructed magnetic resonance (MR) images. The heavy dependency of segmentation approaches on image quality significantly limits the acceleration rate…

图像与视频处理 · 电气工程与系统科学 2025-07-03 Yundi Zhang , Nil Stolt-Ansó , Jiazhen Pan , Wenqi Huang , Kerstin Hammernik , Daniel Rueckert

Estimating the camera's pose given images from a single camera is a traditional task in mobile robots and autonomous vehicles. This problem is called monocular visual odometry and often relies on geometric approaches that require…

计算机视觉与模式识别 · 计算机科学 2025-01-22 André O. Françani , Marcos R. O. A. Maximo

Medical image registration is a challenging task involving the estimation of spatial transformations to establish anatomical correspondence between pairs or groups of images. Recently, deep learning-based image registration methods have…

计算机视觉与模式识别 · 计算机科学 2022-11-28 Xiang Chen , Yan Xia , Nishant Ravikumar , Alejandro F Frangi

Encoder-decoder transformer models have achieved great success on various vision-language (VL) tasks, but they suffer from high inference latency. Typically, the decoder takes up most of the latency because of the auto-regressive decoding.…

计算机视觉与模式识别 · 计算机科学 2023-11-16 Peng Tang , Pengkai Zhu , Tian Li , Srikar Appalaraju , Vijay Mahadevan , R. Manmatha

This study proposed a deep learning-based tracking method for ultrasound (US) image-guided radiation therapy. The proposed cascade deep learning model is composed of an attention network, a mask region-based convolutional neural network…

计算机视觉与模式识别 · 计算机科学 2023-02-15 Yupei Zhang , Xianjin Dai , Zhen Tian , Yang Lei , Jacob F. Wynne , Pretesh Patel , Yue Chen , Tian Liu , Xiaofeng Yang

Safe motion planning in robotics requires planning into space which has been verified to be free of obstacles. However, obtaining such environment representations using lidars is challenging by virtue of the sparsity of their depth…

机器人学 · 计算机科学 2022-07-27 Yifu Tao , Marija Popović , Yiduo Wang , Sundara Tejaswi Digumarti , Nived Chebrolu , Maurice Fallon

Monocular depth estimation is a fundamental task in computer vision and has drawn increasing attention. Recently, some methods reformulate it as a classification-regression task to boost the model performance, where continuous depth is…

计算机视觉与模式识别 · 计算机科学 2022-04-05 Zhenyu Li , Xuyang Wang , Xianming Liu , Junjun Jiang

We propose an architecture and training scheme to predict video frames by explicitly modeling dis-occlusions and capturing the evolution of semantically consistent regions in the video. The scene layout (semantic map) and motion (optical…

计算机视觉与模式识别 · 计算机科学 2021-04-21 Xinzhu Bei , Yanchao Yang , Stefano Soatto

Labeled sequence transduction is a task of transforming one sequence into another sequence that satisfies desiderata specified by a set of labels. In this paper we propose multi-space variational encoder-decoders, a new model for labeled…

计算与语言 · 计算机科学 2019-10-08 Chunting Zhou , Graham Neubig

Deformable image registration estimates voxel-wise correspondences between images through spatial transformations, and plays a key role in medical imaging. While deep learning methods have significantly reduced runtime, efficiently handling…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Tianran Li , Marius Staring , Yuchuan Qiao

In laparoscopic liver surgery, augmented reality technology enhances intraoperative anatomical guidance by overlaying 3D liver models from preoperative CT/MRI onto laparoscopic 2D views. However, existing registration methods lack explicit…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Ruize Cui , Jialun Pei , Haiqiao Wang , Jun Zhou , Jeremy Yuen-Chun Teoh , Pheng-Ann Heng , Jing Qin

In this paper, we propose a novel medical image segmentation using iterative deep learning framework. We have combined an iterative learning approach and an encoder-decoder network to improve segmentation results, which enables to precisely…

计算机视觉与模式识别 · 计算机科学 2017-08-14 Jung Uk Kim , Hak Gu Kim , Yong Man Ro

Flexible sensors are increasingly employed in soft robotics and wearable devices to provide proprioception of freeform deformations.Although supervised learning can train shape predictors from sensor signals, prediction accuracy strongly…

机器人学 · 计算机科学 2026-03-12 Yingjun Tian , Guoxin Fang , Aoran Lyu , Xilong Wang , Zikang Shi , Yuhu Guo , Weiming Wang , Charlie C. L. Wang

Automatic segmentation of abdominal organs in computed tomography (CT) images can support radiation therapy and image-guided surgery workflows. Developing of such automatic solutions remains challenging mainly owing to complex organ…

图像与视频处理 · 电气工程与系统科学 2023-05-22 Zefan Yang , Di Lin , Dong Ni , Yi Wang

Pre-training strategies based on self-supervised learning (SSL) have proven to be effective pretext tasks for many downstream tasks in computer vision. Due to the significant disparity between medical and natural images, the application of…

In this paper, a self-supervised model that simultaneously predicts a sequence of future frames from video-input with a novel spatial-temporal attention (ST) network is proposed. The ST transformer network allows constraining both temporal…

计算机视觉与模式识别 · 计算机科学 2023-03-03 Houssem Boulahbal , Adrian Voicila , Andrew Comport

One of the most elusive goals in myographic prosthesis control is the ability to reliably decode continuous positions simultaneously across multiple degrees-of-freedom. Goal: To demonstrate dexterous, natural, biomimetic finger and wrist…