中文
相关论文

相关论文: ILV: Iterative Latent Volumes for Fast and Accurat…

200 篇论文

Recently, vision model pre-training has evolved from relying on manually annotated datasets to leveraging large-scale, web-crawled image-text data. Despite these advances, there is no pre-training method that effectively exploits the…

计算机视觉与模式识别 · 计算机科学 2024-12-23 Chenyu Yang , Xizhou Zhu , Jinguo Zhu , Weijie Su , Junjie Wang , Xuan Dong , Wenhai Wang , Lewei Lu , Bin Li , Jie Zhou , Yu Qiao , Jifeng Dai

Sparse-view 3D CT reconstruction aims to recover volumetric structures from a limited number of 2D X-ray projections. Existing feedforward methods are constrained by the scarcity of large-scale training datasets and the absence of direct…

图像与视频处理 · 电气工程与系统科学 2026-01-29 Guofeng Zhang , Ruyi Zha , Hao He , Yixun Liang , Alan Yuille , Hongdong Li , Yuanhao Cai

Recent advances in large generative models have shown that simple autoregressive formulations, when scaled appropriately, can exhibit strong zero-shot generalization across domains. Motivated by this trend, we investigate whether…

计算机视觉与模式识别 · 计算机科学 2026-04-27 Yuxiang Lai , Jike Zhong , Ming Li , Yuheng Li , Xiaofeng Yang

Chest computed tomography (CT) at inspiration is often complemented by an expiratory CT to identify peripheral airways disease. Additionally, co-registered inspiratory-expiratory volumes can be used to derive various markers of lung…

High-resolution medical images can provide more detailed information for better diagnosis. Conventional medical image super-resolution relies on a single task which first performs the extraction of the features and then upscaling based on…

图像与视频处理 · 电气工程与系统科学 2025-04-25 Xiaoyan Kui , Zexin Ji , Beiji Zou , Yang Li , Yulan Dai , Liming Chen , Pierre Vera , Su Ruan

3D Particle Imaging Velocimetry (3D-PIV) aim to recover the flow field in a volume of fluid, which has been seeded with tracer particles and observed from multiple camera viewpoints. The first step of 3D-PIV is to reconstruct the 3D…

计算机视觉与模式识别 · 计算机科学 2018-05-23 Katrin Lasinger , Christoph Vogel , Thomas Pock , Konrad Schindler

Purpose: To accelerate MRI acquisition by incorporating the previous scans of a subject during reconstruction. Although longitudinal imaging constitutes much of clinical MRI, leveraging previous scans is challenging due to the complex…

图像与视频处理 · 电气工程与系统科学 2025-10-21 Yonatan Urman , Zachary Shah , Ashwin Kumar , Bruno P. Soares , Kawin Setsompop

Vision-Language Models (VLMs) are increasingly tasked with ultra-long multimodal understanding. While linear architectures offer constant computation and memory footprints, they often struggle with high-frequency visual perception compared…

计算机视觉与模式识别 · 计算机科学 2026-04-01 Hongyuan Tao , Bencheng Liao , Shaoyu Chen , Haoran Yin , Qian Zhang , Wenyu Liu , Xinggang Wang

In-Context derived Vector (ICV) methods extract task-relevant representations from large language models (LLMs) and reinject them during inference, achieving comparable performance to few-shot In-Context Learning (ICL) without repeated…

计算与语言 · 计算机科学 2025-10-13 Wang Cai , Hsiu-Yuan Huang , Zhixiang Wang , Yunfang Wu

A major challenge in X-ray computed tomography (CT) is reducing radiation dose while maintaining high quality of reconstructed images. To reduce the radiation dose, one can reduce the number of projection views (sparse-view CT); however, it…

机器学习 · 统计学 2019-09-17 Xuehang Zheng , Il Yong Chun , Zhipeng Li , Yong Long , Jeffrey A. Fessler

Computed Tomography (CT) reconstruction is a fundamental component to a wide variety of applications ranging from security, to healthcare. The classical techniques require measuring projections, called sinograms, from a full 180$^\circ$…

计算机视觉与模式识别 · 计算机科学 2018-07-12 Rushil Anirudh , Hyojin Kim , Jayaraman J. Thiagarajan , K. Aditya Mohan , Kyle Champley , Timo Bremer

This paper introduces a novel approach to efficiently feeding knowledge to language models (LLMs) during prediction by integrating retrieval and generation processes within a unified framework. While the Retrieval-Augmented Generation (RAG)…

计算与语言 · 计算机科学 2025-02-11 S Santosh Kumar , Rishi Gottimukkala , Supriya Devidutta , Karthikeyan S

The dose of X-ray radiation and the scanning time are crucial factors in computed tomography (CT) for clinical applications. In this work, we introduce a multi-source static CT imaging system designed to rapidly acquire sparse view and…

医学物理 · 物理学 2025-01-03 Ziju Shen , Haimiao Zhang , Bin Dong , Jun Qiu , Yunxiang Li , Zhili Cui

Recent generative models have shown strong performance in generating diverse 3D assets from 2D images, a fundamental research topic in computer vision and graphics. However, these models still struggle to generate voluminous 3D assets when…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Hankyeol Lee , Wooyeol Baek , Seongdo Kim , Jongyoo Kim

Reconstructing dynamic assets from video data is central to many in computer vision and graphics tasks. Existing 4D reconstruction approaches are limited by category-specific models or slow optimization-based methods. Inspired by the recent…

计算机视觉与模式识别 · 计算机科学 2025-03-31 Remy Sabathier , Niloy J. Mitra , David Novotny

In the context of large-angle cone-beam tomography (CBCT), we present a practical iterative reconstruction (IR) scheme designed for rapid convergence as required for large datasets. The robustness of the reconstruction is provided by the…

Generative modelling of entire CT volumes conditioned on clinical reports has the potential to accelerate research through data augmentation, privacy-preserving synthesis and reducing regulator-constraints on patient data while preserving…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Jiayi Wang , Hadrien Reynaud , Franciskus Xaverius Erick , Bernhard Kainz

Limited view tomographic reconstruction aims to reconstruct a tomographic image from a limited number of sinogram or projection views arising from sparse view or limited angle acquisitions that reduce radiation dose or shorten scanning…

图像与视频处理 · 电气工程与系统科学 2020-09-04 Bo Zhou , S. Kevin Zhou , James S. Duncan , Chi Liu

We propose Inner Loop Feedback (ILF), a novel approach to accelerate diffusion models' inference. ILF trains a lightweight module to predict future features in the denoising process by leveraging the outputs from a chosen diffusion backbone…

计算机视觉与模式识别 · 计算机科学 2025-03-31 Matthew Gwilliam , Han Cai , Di Wu , Abhinav Shrivastava , Zhiyu Cheng

Cone beam computed tomography (CBCT) is an important imaging technology widely used in medical scenarios, such as diagnosis and preoperative planning. Using fewer projection views to reconstruct CT, also known as sparse-view reconstruction,…

图像与视频处理 · 电气工程与系统科学 2024-06-07 Yiqun Lin , Jiewen Yang , Hualiang Wang , Xinpeng Ding , Wei Zhao , Xiaomeng Li