English
Related papers

Related papers: Refining 3D Medical Segmentation with Verbal Instr…

200 papers

Evaluating conversational systems in multi-turn settings remains a fundamental challenge. Conventional pipelines typically rely on manually defined rubrics and fixed conversational context$-$a static approach that limits coverage and fails…

Computation and Language · Computer Science 2026-01-21 Yunzhe Li , Richie Yueqi Feng , Tianxin Wei , Chin-Chia Hsu

Incorporation of prior knowledge about organ shape and location is key to improve performance of image analysis approaches. In particular, priors can be useful in cases where images are corrupted and contain artefacts due to limitations in…

Diagnosis based on medical images, such as X-ray images, often involves manual annotation of anatomical keypoints. However, this process involves significant human efforts and can thus be a bottleneck in the diagnostic process. To fully…

Computer Vision and Pattern Recognition · Computer Science 2024-05-07 Jinhee Kim , Taesung Kim , Taewoo Kim , Jaegul Choo , Dong-Wook Kim , Byungduk Ahn , In-Seok Song , Yoon-Ji Kim

Medical image segmentation is a critical task in medical image analysis. In recent years, deep learning based approaches have shown exceptional performance when trained on a fully-annotated dataset. However, data annotation is often a…

Computer Vision and Pattern Recognition · Computer Science 2023-07-25 Han Liu , Hao Li , Xing Yao , Yubo Fan , Dewei Hu , Benoit Dawant , Vishwesh Nath , Zhoubing Xu , Ipek Oguz

Annotation of medical images has been a major bottleneck for the development of accurate and robust machine learning models. Annotation is costly and time-consuming and typically requires expert knowledge, especially in the medical domain.…

Computer Vision and Pattern Recognition · Computer Science 2020-01-08 Holger Roth , Ling Zhang , Dong Yang , Fausto Milletari , Ziyue Xu , Xiaosong Wang , Daguang Xu

Designing mechanical linkages involves combinatorial topology selection and continuous parameter fitting. We show that language models can systematically improve linkage designs through symbolic representations. Language model agents…

Artificial Intelligence · Computer Science 2026-05-01 João Pedro Gandarela , Thiago Rios , Stefan Menzel , André Freitas

X-ray angiography is widely used in cardiac interventions to visualize coronary vessels, assess integrity, detect stenoses and guide treatment. We propose a framework for reconstructing 3D vessel trees from biplanar X-ray images which are…

Image and Video Processing · Electrical Eng. & Systems 2025-09-18 Ethan Koland , Lin Xi , Nadeev Wijesuriya , YingLiang Ma

Many applications in 3D shape design and augmentation require the ability to make specific edits to an object's semantic parameters (e.g., the pose of a person's arm or the length of an airplane's wing) while preserving as much existing…

Computer Vision and Pattern Recognition · Computer Science 2020-11-11 Fangyin Wei , Elena Sizikova , Avneesh Sud , Szymon Rusinkiewicz , Thomas Funkhouser

Medical image annotation is a major hurdle for developing precise and robust machine learning models. Annotation is expensive, time-consuming, and often requires expert knowledge, particularly in the medical field. Here, we suggest using…

Computer Vision and Pattern Recognition · Computer Science 2020-09-28 Holger R Roth , Dong Yang , Ziyue Xu , Xiaosong Wang , Daguang Xu

Although the preservation of shape continuity and physiological anatomy is a natural assumption in the segmentation of medical images, it is often neglected by deep learning methods that mostly aim for the statistical modeling of input data…

Computer Vision and Pattern Recognition · Computer Science 2023-05-01 Yousef Yeganeh , Azade Farshad , Goktug Guevercin , Amr Abu-zer , Rui Xiao , Yongjian Tang , Ehsan Adeli , Nassir Navab

Medical vision-language models (Med-VLMs) have shown impressive results in tasks such as report generation and visual question answering, but they still face several limitations. Most notably, they underutilize patient metadata and lack…

Computer Vision and Pattern Recognition · Computer Science 2025-10-16 Fangqi Cheng , Surajit Ray , Xiaochen Yang

Piece-wise planar 3D reconstruction simultaneously segments plane instances and recovers their 3D plane parameters from an image, which is particularly useful for indoor or man-made environments. Efficient reconstruction of 3D planes…

Computer Vision and Pattern Recognition · Computer Science 2023-12-22 Luan Wei , Anna Hilsmann , Peter Eisert

Many successful methods developed for medical image analysis that are based on machine learning use supervised learning approaches, which often require large datasets annotated by experts to achieve high accuracy. However, medical data…

Computer Vision and Pattern Recognition · Computer Science 2022-07-25 Banafshe Felfeliyan , Abhilash Hareendranathan , Gregor Kuntze , David Cornell , Nils D. Forkert , Jacob L. Jaremko , Janet L. Ronsky

We propose a novel approach that uses sparse annotations from clinical studies to train a 3D segmentation of the carotid artery wall. We use a centerline annotation to sample perpendicular cross-sections of the carotid artery and use an…

Computer Vision and Pattern Recognition · Computer Science 2025-04-10 Hinrich Rahlfs , Markus Hüllebrand , Sebastian Schmitter , Christoph Strecker , Andreas Harloff , Anja Hennemuth

This paper aims to build a model that can Segment Anything in 3D medical images, driven by medical terminologies as Text prompts, termed as SAT. Our main contributions are three-fold: (i) We construct the first multimodal knowledge tree on…

Image and Video Processing · Electrical Eng. & Systems 2025-07-21 Ziheng Zhao , Yao Zhang , Chaoyi Wu , Xiaoman Zhang , Xiao Zhou , Ya Zhang , Yanfeng Wang , Weidi Xie

3D shapes captured by scanning devices are often incomplete due to occlusion. 3D shape completion methods have been explored to tackle this limitation. However, most of these methods are only trained and tested on a subset of categories,…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Lintai Wu , Junhui Hou , Linqi Song , Yong Xu

Speech-driven 3D facial animation is a challenging cross-modal task that has attracted growing research interest. During speaking activities, the mouth displays strong motions, while the other facial regions typically demonstrate…

Computer Vision and Pattern Recognition · Computer Science 2023-10-18 Zhaojie Chu , Kailing Guo , Xiaofen Xing , Yilin Lan , Bolun Cai , Xiangmin Xu

Three-dimensional (3D) images, such as CT, MRI, and PET, are common in medical imaging applications and important in clinical diagnosis. Semantic ambiguity is a typical feature of many medical image labels. It can be caused by many factors,…

Image and Video Processing · Electrical Eng. & Systems 2022-09-19 Lin Wang , Xiufen Ye , Donghao Zhang , Wanji He , Lie Ju , Xin Wang , Wei Feng , Kaimin Song , Xin Zhao , Zongyuan Ge

Automatic segmentation of organs-at-risk (OARs) in CT scans using convolutional neural networks (CNNs) is being introduced into the radiotherapy workflow. However, these segmentations still require manual editing and approval by clinicians…

Computer Vision and Pattern Recognition · Computer Science 2022-06-28 Edward G. A. Henderson , Andrew F. Green , Marcel van Herk , Eliana M. Vasquez Osorio

We propose a novel technique for adding geometric details to an input coarse 3D mesh guided by a text prompt. Our method is composed of three stages. First, we generate a single-view RGB image conditioned on the input coarse geometry and…

Computer Vision and Pattern Recognition · Computer Science 2024-09-12 Yun-Chun Chen , Selena Ling , Zhiqin Chen , Vladimir G. Kim , Matheus Gadelha , Alec Jacobson