English
Related papers

Related papers: KaoLRM: Repurposing Pre-trained Large Reconstructi…

200 papers

We propose PRM, a novel photometric stereo based large reconstruction model to reconstruct high-quality meshes with fine-grained local details. Unlike previous large reconstruction models that prepare images under fixed and simple lighting…

Computer Vision and Pattern Recognition · Computer Science 2024-12-11 Wenhang Ge , Jiantao Lin , Guibao Shen , Jiawei Feng , Tao Hu , Xinli Xu , Ying-Cong Chen

In this study, we present a large-scale earth surface reconstruction pipeline for linear-array charge-coupled device (CCD) satellite imagery. While mainstream satellite image-based reconstruction approaches perform exceptionally well, the…

Computer Vision and Pattern Recognition · Computer Science 2023-11-01 Hong Danyang , Yu Anzhu , Ji Song , Cao Xuefeng , Quan Yujun , Guo Wenyue , Qiu Chunping

Multimodal Large Language Models (MLLMs) have recently demonstrated strong performance on a wide range of vision-language tasks, raising interest in their potential use for biometric applications. In this paper, we conduct a systematic…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Hatef Otroshi Shahreza , Anjith George , Sébastien Marcel

Monocular SLAM has received a lot of attention due to its simple RGB inputs and the lifting of complex sensor constraints. However, existing monocular SLAM systems are designed for bounded scenes, restricting the applicability of SLAM…

Computer Vision and Pattern Recognition · Computer Science 2024-03-11 Heng Zhou , Zhetao Guo , Shuhong Liu , Lechen Zhang , Qihao Wang , Yuxiang Ren , Mingrui Li

Recent large reconstruction models have made notable progress in generating high-quality 3D objects from single images. However, current reconstruction methods often rely on explicit camera pose estimation or fixed viewpoints, restricting…

Computer Vision and Pattern Recognition · Computer Science 2025-03-11 Hao He , Yixun Liang , Luozhou Wang , Yuanhao Cai , Xinli Xu , Hao-Xiang Guo , Xiang Wen , Yingcong Chen

Deep learning-based 3-dimensional (3D) shape reconstruction from 2-dimensional (2D) magnetic resonance imaging (MRI) has become increasingly important in medical disease diagnosis, treatment planning, and computational modeling. This review…

Machine Learning · Computer Science 2025-10-03 Emma McMillian , Abhirup Banerjee , Alfonso Bueno-Orovio

High-resolution (HR) MRI scans obtained from research-grade medical centers provide precise information about imaged tissues. However, routine clinical MRI scans are typically in low-resolution (LR) and vary greatly in contrast and spatial…

Image and Video Processing · Electrical Eng. & Systems 2023-08-25 Jueqi Wang , Jacob Levman , Walter Hugo Lopez Pinaya , Petru-Daniel Tudosiu , M. Jorge Cardoso , Razvan Marinescu

Existing works on motion deblurring either ignore the effects of depth-dependent blur or work with the assumption of a multi-layered scene wherein each layer is modeled in the form of fronto-parallel plane. In this work, we consider the…

Computer Vision and Pattern Recognition · Computer Science 2022-02-08 Kuldeep Purohit , Subeesh Vasu , M. Purnachandra Rao , A. N. Rajagopalan

Composed Image Retrieval (CIR) is a complex task that aims to retrieve images based on a multimodal query. Typical training data consists of triplets containing a reference image, a textual description of desired modifications, and the…

Computer Vision and Pattern Recognition · Computer Science 2025-03-26 Chuong Huynh , Jinyu Yang , Ashish Tawari , Mubarak Shah , Son Tran , Raffay Hamid , Trishul Chilimbi , Abhinav Shrivastava

mmWave radar has been shown as an effective sensing technique in low visibility, smoke, dusty, and dense fog environment. However tapping the potential of radar sensing to reconstruct 3D object shapes remains a great challenge, due to the…

Image and Video Processing · Electrical Eng. & Systems 2021-08-29 Yue Sun , Zhuoming Huang , Honggang Zhang , Zhi Cao , Deqiang Xu

We aim to address sparse-view reconstruction of a 3D scene by leveraging priors from large-scale vision models. While recent advancements such as 3D Gaussian Splatting (3DGS) have demonstrated remarkable successes in 3D reconstruction,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-29 Hanyang Yu , Xiaoxiao Long , Ping Tan

Generating editable, parametric CAD models from a single image holds great potential to lower the barriers of industrial concept design. However, current multi-modal large language models (MLLMs) still struggle with accurately inferring 3D…

Computer Vision and Pattern Recognition · Computer Science 2025-10-21 Yinghui Wang , Xinyu Zhang , Peng Du

We propose Long-LRM, a feed-forward 3D Gaussian reconstruction model for instant, high-resolution, 360{\deg} wide-coverage, scene-level reconstruction. Specifically, it takes in 32 input images at a resolution of 960x540 and produces the…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Chen Ziwen , Hao Tan , Kai Zhang , Sai Bi , Fujun Luan , Yicong Hong , Li Fuxin , Zexiang Xu

As a classic statistical model of 3D facial shape and albedo, 3D Morphable Model (3DMM) is widely used in facial analysis, e.g., model fitting, image synthesis. Conventional 3DMM is learned from a set of 3D face scans with associated…

Computer Vision and Pattern Recognition · Computer Science 2019-07-16 Luan Tran , Xiaoming Liu

Enabling Large Language Models (LLMs) to interact with 3D environments is challenging. Existing approaches extract point clouds either from ground truth (GT) geometry or 3D scenes reconstructed by auxiliary models. Text-image aligned 2D…

Computer Vision and Pattern Recognition · Computer Science 2024-04-22 Tao Chu , Pan Zhang , Xiaoyi Dong , Yuhang Zang , Qiong Liu , Jiaqi Wang

Renovating existing buildings is essential for climate impact. Early-phase renovation planning requires simulations based on thermal 3D models at Level of Detail (LoD) 3, which include features like windows. However, scalable and accurate…

Computer Vision and Pattern Recognition · Computer Science 2025-08-07 Yinan Yu , Alex Gonzalez-Caceres , Samuel Scheidegger , Sanjay Somanath , Alexander Hollberg

New era has unlocked exciting possibilities for extending Large Language Models (LLMs) to tackle 3D vision-language tasks. However, most existing 3D multimodal LLMs (MLLMs) rely on compressing holistic 3D scene information or segmenting…

Computer Vision and Pattern Recognition · Computer Science 2025-07-23 Xiaoyan Wang , Zeju Li , Yifan Xu , Jiaxing Qi , Zhifei Yang , Ruifei Ma , Xiangde Liu , Chao Zhang

Achieving high-fidelity 3D reconstruction from monocular video remains challenging due to the inherent limitations of traditional methods like Structure-from-Motion (SfM) and monocular SLAM in accurately capturing scene details. While…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Yue Hu , Rong Liu , Meida Chen , Peter Beerel , Andrew Feng

Recent methods for 3D reconstruction and rendering increasingly benefit from end-to-end optimization of the entire image formation process. However, this approach is currently limited: effects of the optical hardware stack and in particular…

Computer Vision and Pattern Recognition · Computer Science 2023-04-12 Wenqi Xian , Aljaž Božič , Noah Snavely , Christoph Lassner

Facial motion retargeting is an important problem in both computer graphics and vision, which involves capturing the performance of a human face and transferring it to another 3D character. Learning 3D morphable model (3DMM) parameters from…

Computer Vision and Pattern Recognition · Computer Science 2019-03-01 Bindita Chaudhuri , Noranart Vesdapunt , Baoyuan Wang