English
Related papers

Related papers: Render-FM: A Foundation Model for Real-time Photor…

200 papers

We propose a Transformer architecture for volumetric segmentation, a challenging task that requires keeping a complex balance in encoding local and global spatial cues, and preserving information along all axes of the volume. Encoder of the…

Image and Video Processing · Electrical Eng. & Systems 2022-07-04 Himashi Peiris , Munawar Hayat , Zhaolin Chen , Gary Egan , Mehrtash Harandi

A long-standing goal in scene understanding is to obtain interpretable and editable representations that can be directly constructed from a raw monocular RGB-D video, without requiring specialized hardware setup or priors. The problem is…

Computer Vision and Pattern Recognition · Computer Science 2023-06-22 Yu-Shiang Wong , Niloy J. Mitra

Neural networks have shown great potential in compressing volume data for visualization. However, due to the high cost of training and inference, such volumetric neural representations have thus far only been applied to offline data…

Graphics · Computer Science 2023-07-03 Qi Wu , David Bauer , Michael J. Doyle , Kwan-Liu Ma

We propose a novel Neural Radiance Field (NeRF) representation for non-opaque scenes that enables fast inference by utilizing textured polygons. Despite the high-quality novel view rendering that NeRF provides, a critical limitation is that…

Graphics · Computer Science 2024-07-11 Gopal Sharma , Daniel Rebain , Kwang Moo Yi , Andrea Tagliasacchi

Deep learning-based segmentation of genito-pelvic structures in MRI and CT is crucial for applications such as radiation therapy, surgical planning, and disease diagnosis. However, existing segmentation models often struggle with…

Image and Video Processing · Electrical Eng. & Systems 2025-03-19 Yuheng Li , Mingzhe Hu , Richard L. J. Qiu , Maria Thor , Andre Williams , Deborah Marshall , Xiaofeng Yang

We introduce GRM, a large-scale reconstructor capable of recovering a 3D asset from sparse-view images in around 0.1s. GRM is a feed-forward transformer-based model that efficiently incorporates multi-view information to translate the input…

Computer Vision and Pattern Recognition · Computer Science 2024-03-22 Yinghao Xu , Zifan Shi , Wang Yifan , Hansheng Chen , Ceyuan Yang , Sida Peng , Yujun Shen , Gordon Wetzstein

We present a neural-field-based large-scale reconstruction system that fuses lidar and vision data to generate high-quality reconstructions that are geometrically accurate and capture photo-realistic textures. This system adapts the…

Robotics · Computer Science 2025-02-18 Yifu Tao , Yash Bhalgat , Lanke Frank Tarimo Fu , Matias Mattamala , Nived Chebrolu , Maurice Fallon

3D Gaussian splatting (3DGS) has recently emerged as an alternative representation that leverages a 3D Gaussian-based representation and introduces an approximated volumetric rendering, achieving very fast rendering speed and promising…

Computer Vision and Pattern Recognition · Computer Science 2024-08-08 Joo Chan Lee , Daniel Rho , Xiangyu Sun , Jong Hwan Ko , Eunbyung Park

Reconstruction of the soft tissues in robotic surgery from endoscopic stereo videos is important for many applications such as intra-operative navigation and image-guided robotic surgery automation. Previous works on this task mainly rely…

Computer Vision and Pattern Recognition · Computer Science 2022-07-01 Yuehao Wang , Yonghao Long , Siu Hin Fan , Qi Dou

Differentiable volumetric rendering-based methods made significant progress in novel view synthesis. On one hand, innovative methods have replaced the Neural Radiance Fields (NeRF) network with locally parameterized structures, enabling…

Computer Vision and Pattern Recognition · Computer Science 2025-02-06 Hugo Blanc , Jean-Emmanuel Deschaud , Alexis Paljic

Real-time visualization of large-scale volumetric data remains challenging, as direct volume rendering and voxel-based methods suffer from prohibitively high computational cost. We propose Variable Basis Mapping (VBM), a framework that…

Graphics · Computer Science 2026-01-15 Qibiao Li , Yuxuan Wang , Youcheng Cai , Huangsheng Du , Ligang Liu

This work aims to generate realistic anatomical deformations from static patient scans. Specifically, we present a method to generate these deformations/augmentations via deep learning driven respiratory motion simulation that provides the…

Computer Vision and Pattern Recognition · Computer Science 2023-01-30 Donghoon Lee , Ellen Yorke , Masoud Zarepisheh , Saad Nadeem , Yu-Chi Hu

Pretraining with large-scale 3D volumes has a potential for improving the segmentation performance on a target medical image dataset where the training images and annotations are limited. Due to the high cost of acquiring pixel-level…

Computer Vision and Pattern Recognition · Computer Science 2023-06-30 Guotai Wang , Jianghao Wu , Xiangde Luo , Xinglong Liu , Kang Li , Shaoting Zhang

3D scene reconstruction is fundamental for spatial intelligence applications such as AR, robotics, and digital twins. Traditional multi-view stereo struggles with sparse viewpoints or low-texture regions, while neural rendering approaches,…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Jiaqi Yao , Zhongmiao Yan , Jingyi Xu , Songpengcheng Xia , Yan Xiang , Ling Pei

Direct Volume Rendering (DVR) using Volumetric Path Tracing (VPT) is a scientific visualization technique that simulates light transport with objects' matter using physically-based lighting models. Monte Carlo (MC) path tracing is often…

Graphics · Computer Science 2021-06-16 Jose A. Iglesias-Guitian , Prajita Mane , Bochang Moon

Photo-realistic rendering and novel view synthesis play a crucial role in human-computer interaction tasks, from gaming to path planning. Neural Radiance Fields (NeRFs) model scenes as continuous volumetric functions and achieve remarkable…

Computer Vision and Pattern Recognition · Computer Science 2025-03-19 Iryna Repinetska , Anna Hilsmann , Peter Eisert

Magnetic Resonance Imaging (MRI) is a crucial diagnostic tool, but high-resolution scans are often slow and expensive due to extensive data acquisition requirements. Traditional MRI reconstruction methods aim to expedite this process by…

Image and Video Processing · Electrical Eng. & Systems 2025-11-11 Emmanuelle Bourigault , Abdullah Hamdi , Amir Jamaludin

Recent medical vision-language models (VLMs) have shown promise in 2D medical image interpretation. However extending them to 3D medical imaging has been challenging due to computational complexities and data scarcity. Although a few recent…

Image and Video Processing · Electrical Eng. & Systems 2024-12-19 Changsun Lee , Sangjoon Park , Cheong-Il Shin , Woo Hee Choi , Hyun Jeong Park , Jeong Eun Lee , Jong Chul Ye

The convolutional neural network (CNN) has become a powerful tool for various biomedical image analysis tasks, but there is a lack of visual explanation for the machinery of CNNs. In this paper, we present a novel algorithm,…

Computer Vision and Pattern Recognition · Computer Science 2018-06-08 Guannan Zhao , Bo Zhou , Kaiwen Wang , Rui Jiang , Min Xu

Diffusion models have demonstrated significant potential in producing high-quality images in medical image translation to aid disease diagnosis, localization, and treatment. Nevertheless, current diffusion models have limited success in…

Image and Video Processing · Electrical Eng. & Systems 2024-11-26 Yunxiang Li , Hua-Chieh Shao , Xiaoxue Qian , You Zhang