中文
相关论文

相关论文: TiAVox: Time-aware Attenuation Voxels for Sparse-v…

200 篇论文

Vehicle-to-Everything (V2X) collaborative perception has recently gained significant attention due to its capability to enhance scene understanding by integrating information from various agents, e.g., vehicles, and infrastructure. However,…

计算机视觉与模式识别 · 计算机科学 2023-12-27 Li Xiang , Junbo Yin , Wei Li , Cheng-Zhong Xu , Ruigang Yang , Jianbing Shen

Tomographic image reconstruction is relevant for many medical imaging modalities including X-ray, ultrasound (US) computed tomography (CT) and photoacoustics, for which the access to full angular range tomographic projections might be not…

图像与视频处理 · 电气工程与系统科学 2019-06-14 Valery Vishnevskiy , Richard Rau , Orcun Goksel

Dynamic Photoacoustic Computed Tomography (PACT) is an important imaging technique for monitoring physiological processes, capable of providing high-contrast images of optical absorption at much greater depths than traditional optical…

图像与视频处理 · 电气工程与系统科学 2025-06-05 Youshen Xiao , Yiling Shi , Ruixi Sun , Hongjiang Wei , Fei Gao , Yuyao Zhang

Diffusion Weighted Imaging (DWI) is an advanced imaging technique commonly used in neuroscience and neurological clinical research through a Diffusion Tensor Imaging (DTI) model. Volumetric scalar metrics including fractional anisotropy,…

图像与视频处理 · 电气工程与系统科学 2022-11-01 Zihao Tang , Xinyi Wang , Lihaowen Zhu , Mariano Cabezas , Dongnan Liu , Michael Barnett , Weidong Cai , Chengyu Wang

As PET imaging is accompanied by substantial radiation exposure and cancer risk, reducing radiation dose in PET scans is an important topic. Recently, diffusion models have emerged as the new state-of-the-art generative model to generate…

图像与视频处理 · 电气工程与系统科学 2023-11-30 Huidong Xie , Weijie Gan , Bo Zhou , Xiongchao Chen , Qiong Liu , Xueqi Guo , Liang Guo , Hongyu An , Ulugbek S. Kamilov , Ge Wang , Chi Liu

Video Analytics Software as a Service (VA SaaS) has been rapidly growing in recent years. VA SaaS is typically accessed by users using a lightweight client. Because the transmission bandwidth between the client and cloud is usually limited…

计算机视觉与模式识别 · 计算机科学 2018-08-16 Zhaoyang Zhang , Zhanghui Kuang , Ping Luo , Litong Feng , Wei Zhang

In the treatment of ovarian cancer, precise residual disease prediction is significant for clinical and surgical decision-making. However, traditional methods are either invasive (e.g., laparoscopy) or time-consuming (e.g., manual…

图像与视频处理 · 电气工程与系统科学 2023-06-27 Xiangneng Gao , Shulan Ruan , Jun Shi , Guoqing Hu , Wei Wei

Since current Vision-Language-Action (VLA) systems suffer from limited spatial perception and the absence of memory throughout manipulation, we investigate visual anchors as a means to enhance spatial and temporal reasoning within VLA…

机器人学 · 计算机科学 2026-03-16 Juan Zhu , Zhanying Shao , Xiaoqi Li , Ethan Morgan , Jiadong Xu , Hongwei Fan , Hao Dong

Neural representations (NRs), such as neural fields and 3D Gaussians, effectively model volumetric data in computed tomography (CT) but suffer from severe artifacts under sparse-view settings. To address this, we propose DiffNR, a novel…

图像与视频处理 · 电气工程与系统科学 2026-04-24 Shiyan Su , Ruyi Zha , Danli Shi , Hongdong Li , Xuelian Cheng

We present a novel approach to variational volume reconstruction from sparse, noisy slice data using the Deep Ritz method. Motivated by biomedical imaging applications such as MRI-based slice-to-volume reconstruction (SVR), our approach…

图像与视频处理 · 电气工程与系统科学 2025-08-13 Conor Rowan , Sumedh Soman , John A. Evans

Dynamic three-dimensional (4D) reconstruction from two-dimensional X-ray coronary angiography (CA) remains a significant clinical problem. Existing CA reconstruction methods often require extensive user interaction or large training…

图像与视频处理 · 电气工程与系统科学 2025-06-12 Kirsten W. H. Maas , Danny Ruijters , Anna Vilanova , Nicola Pezzotti

Diffusion Tensor Cardiac Magnetic Resonance (DT-CMR) is the only in vivo method to non-invasively examine the microstructure of the human heart. Current research in DT-CMR aims to improve the understanding of how the cardiac microstructure…

图像与视频处理 · 电气工程与系统科学 2026-05-19 Yinzhe Wu , Jiahao Huang , Fanwen Wang , Pedro Ferreira , Andrew Scott , Sonia Nielles-Vallespin , Guang Yang

Accurate and computationally efficient 3D medical image segmentation remains a critical challenge in clinical workflows. Transformer-based architectures often demonstrate superior global contextual modeling but at the expense of excessive…

图像与视频处理 · 电气工程与系统科学 2026-02-19 Kavyansh Tyagi , Vishwas Rathi , Puneet Goyal

Audio-visual speech recognition (AVSR) combines audio-visual modalities to improve speech recognition, especially in noisy environments. However, most existing methods deploy the unidirectional enhancement or symmetric fusion manner, which…

多媒体 · 计算机科学 2025-08-12 Junxiao Xue , Xiaozhen Liu , Xuecheng Wu , Xinyi Yin , Danlei Huang , Fei Yu

Typical optical coherence tomographic angiography (OCTA) acquisition areas on commercial devices are 3x3- or 6x6-mm. Compared to 3x3-mm angiograms with proper sampling density, 6x6-mm angiograms have significantly lower scan quality, with…

图像与视频处理 · 电气工程与系统科学 2020-06-11 Min Gao , Yukun Guo , Tristan T. Hormel , Jiande Sun , Thomas Hwang , Yali Jia

In image-guided radiotherapy (IGRT), four-dimensional cone-beam computed tomography (4D-CBCT) is critical for assessing tumor motion during a patients breathing cycle prior to beam delivery. However, generating 4D-CBCT images with…

医学物理 · 物理学 2025-01-09 Yabo Fu , Hao Zhang , Weixing Cai , Huiqiao Xie , Licheng Kuo , Laura Cervino , Jean Moran , Xiang Li , Tianfang Li

Continuous treatment effect estimation holds significant practical importance across various decision-making and assessment domains, such as healthcare and the military. However, current methods for estimating dose-response curves hinge on…

机器学习 · 计算机科学 2024-06-05 Ruijing Cui , Jianbin Sun , Bingyu He , Kewei Yang , Bingfeng Ge

Continuous speech representations based on Variational Autoencoders (VAEs) have emerged as a promising alternative to traditional spectrogram or discrete token based features for speech generation and reconstruction. Recent research has…

声音 · 计算机科学 2026-05-26 Changhao Cheng , Wei Wang , Wangyou Zhang , Dongya Jia , Jian Wu , Zhuo Chen , Yanmin Qian

Recent feed-forward 3D reconstruction methods, such as visual geometry transformers, have substantially advanced the traditional per-scene optimization paradigm by enabling effective multi-view reconstruction in a single forward pass.…

计算机视觉与模式识别 · 计算机科学 2026-05-15 David Huang , Guile Wu , Chengjie Huang , Bingbing Liu , Dongfeng Bai

Diffusion-based large multimodal models, such as LLaDA-V, have demonstrated impressive capabilities in vision-language understanding and generation. However, their bidirectional attention mechanism and diffusion-style iterative denoising…

计算机视觉与模式识别 · 计算机科学 2026-01-29 Zhewen Wan , Tianchen Song , Chen Lin , Zhiyong Zhao , Xianpeng Lang
‹ 上一页 1 8 9 10 下一页 ›