English
Related papers

Related papers: TiAVox: Time-aware Attenuation Voxels for Sparse-v…

200 papers

This paper proposes VARA-TTS, a non-autoregressive (non-AR) text-to-speech (TTS) model using a very deep Variational Autoencoder (VDVAE) with Residual Attention mechanism, which refines the textual-to-acoustic alignment layer-wisely.…

Sound · Computer Science 2021-02-15 Peng Liu , Yuewen Cao , Songxiang Liu , Na Hu , Guangzhi Li , Chao Weng , Dan Su

Speed-of-sound has been shown as a potential biomarker for breast cancer imaging, successfully differentiating malignant tumors from benign ones. Speed-of-sound images can be reconstructed from time-of-flight measurements from ultrasound…

Image and Video Processing · Electrical Eng. & Systems 2020-07-23 Melanie Bernhardt , Valery Vishnevskiy , Richard Rau , Orcun Goksel

We present DiffVox, a self-supervised framework for Cone-Beam Computed Tomography (CBCT) reconstruction by directly optimizing a voxelgrid representation using physics-based differentiable X-ray rendering. Further, we investigate how the…

Image and Video Processing · Electrical Eng. & Systems 2025-03-21 Mohammadhossein Momeni , Vivek Gopalakrishnan , Neel Dey , Polina Golland , Sarah Frisken

In Angiographic Parametric Imaging (API), accurate estimation of parameters from Time Density Curves (TDC) is crucial. However, these estimations are often marred by errors arising from factors such as patient motion, procedural…

Cerebrovascular pathology significantly contributes to cognitive decline and neurological disorders, underscoring the need for advanced tools to assess vascular integrity. Three-dimensional Time-of-Flight Magnetic Resonance Angiography (3D…

Image and Video Processing · Electrical Eng. & Systems 2025-07-11 Abrar Faiyaz , Nhat Hoang , Giovanni Schifitto , Md Nasir Uddin

In intracranial aneurysm (IA) treatment, digital subtraction angiography (DSA) monitors device-induced hemodynamic changes. Quantitative angiography (QA) provides more precise assessments but is limited by hand-injection variability. This…

This work introduces an unsupervised Divergence and Aliasing-Free neural network (DAF-FlowNet) for 4D Flow Magnetic Resonance Imaging (4D Flow MRI) that jointly enhances noisy velocity fields and corrects phase wrapping artifacts.…

This study leverages convolutional neural networks to enhance the temporal resolution of 3D angiography in intracranial aneurysms focusing on the reconstruction of volumetric contrast data from sparse and limited projections. Three…

Purpose: To develop an algorithm for real-time volumetric image reconstruction and 3D tumor localization based on a single x-ray projection image for lung cancer radiotherapy. Methods: Given a set of volumetric images of a patient at N…

Medical Physics · Physics 2015-05-18 Ruijiang Li , Xun Jia , John H. Lewis , Xuejun Gu , Michael Folkerts , Chunhua Men , Steve B. Jiang

Accurate segmentation of 3D medical scans is crucial for clinical diagnostics and treatment planning, yet existing methods often fail to achieve both high accuracy and computational efficiency across diverse anatomies and imaging…

Computer Vision and Pattern Recognition · Computer Science 2026-04-08 Chenxin Yuan , Shoupeng Chen , Haojiang Ye , Yiming Miao , Limei Peng , Pin-Han Ho

Three-dimensional (3D) and dynamic 3D+time (4D) reconstruction of coronary arteries from X-ray coronary angiography (CA) has the potential to improve clinical procedures. However, there are multiple challenges to be addressed, most notably,…

Image and Video Processing · Electrical Eng. & Systems 2025-07-28 Kirsten W. H. Maas , Danny Ruijters , Nicola Pezzotti , Anna Vilanova

Speed-of-sound is a biomechanical property for quantitative tissue differentiation, with great potential as a new ultrasound-based image modality. A conventional ultrasound array transducer can be used together with an acoustic mirror, or…

Computer Vision and Pattern Recognition · Computer Science 2018-07-20 Valery Vishnevskiy , Sergio J Sanabria , Orcun Goksel

Recovering the 3D representation of an object from single-view or multi-view RGB images by deep neural networks has attracted increasing attention in the past few years. Several mainstream works (e.g., 3D-R2N2) use recurrent neural networks…

Computer Vision and Pattern Recognition · Computer Science 2020-06-02 Haozhe Xie , Hongxun Yao , Xiaoshuai Sun , Shangchen Zhou , Shengping Zhang

The unprecedented X-ray flux density provided by modern X-ray sources offers new spatiotemporal possibilities for X-ray imaging of fast dynamic processes. Approaches to exploit such possibilities often result in either i) a limited number…

Image and Video Processing · Electrical Eng. & Systems 2025-04-07 Zisheng Yao , Yuhe Zhang , Zhe Hu , Robert Klöfkorn , Tobias Ritschel , Pablo Villanueva-Perez

Phase-contrast magnetic resonance imaging (MRI) provides time-resolved quantification of blood flow dynamics that can aid clinical diagnosis. Long in vivo scan times due to repeated three-dimensional (3D) volume sampling over cardiac phases…

Image and Video Processing · Electrical Eng. & Systems 2020-04-22 Valery Vishnevskiy , Jonas Walheim , Sebastian Kozerke

Accelerating compute intensive non-real-time beam-forming algorithms in ultrasound imaging using deep learning architectures has been gaining momentum in the recent past. Nonetheless, the complexity of the state-of-the-art deep learning…

Image and Video Processing · Electrical Eng. & Systems 2024-01-17 Abdul Rahoof , Vivek Chaturvedi , Mahesh Raveendranatha Panicker , Muhammad Shafique

Quantitative angiography (QA) in two dimensions has been instrumental in assessing neurovascular contrast flow patterns, aiding disease severity and treatment outcome evaluations. However, QA requires high spatio-temporal resolution,…

This work presents TV-LoRA, a novel method for low-dose sparse-view CT reconstruction that combines a diffusion generative prior (NCSN++ with SDE modeling) and multi-regularization constraints, including anisotropic TV and nuclear norm…

Computer Vision and Pattern Recognition · Computer Science 2025-10-07 Zongyin Deng , Qing Zhou , Yuhao Fang , Zijian Wang , Yao Lu , Ye Zhang , Chun Li

Reconstructing 3D objects from extremely sparse views is a long-standing and challenging problem. While recent techniques employ image diffusion models for generating plausible images at novel viewpoints or for distilling pre-trained…

Computer Vision and Pattern Recognition · Computer Science 2023-12-21 Zi-Xin Zou , Weihao Cheng , Yan-Pei Cao , Shi-Sheng Huang , Ying Shan , Song-Hai Zhang

Diffusion models have recently set new benchmarks in Speech Enhancement (SE). However, most existing score-based models treat speech spectrograms merely as generic 2D images, applying uniform processing that ignores the intrinsic structural…

Sound · Computer Science 2026-02-03 Ke Xue , Rongfei Fan , Kai Li , Shanping Yu , Puning Zhao , Jianping An