中文
相关论文

相关论文: DITTO: Dual and Integrated Latent Topologies for I…

200 篇论文

3D vision foundation models have shown strong generalization in reconstructing key 3D attributes from uncalibrated images through a single feed-forward pass. However, when deployed in online settings such as driving scenarios, predictions…

计算机视觉与模式识别 · 计算机科学 2026-03-19 Fengyi Zhang , Tianjun Zhang , Kasra Khosoussi , Zheng Zhang , Zi Huang , Yadan Luo

Relighting a person from a single photo is an attractive but ill-posed task, as a 2D image ambiguously entangles 3D geometry, intrinsic appearance, and illumination. Current methods either use sequential pipelines that suffer from error…

计算机视觉与模式识别 · 计算机科学 2026-04-23 Yuxuan Xue , Ruofan Liang , Egor Zakharov , Timur Bagautdinov , Chen Cao , Giljoo Nam , Shunsuke Saito , Gerard Pons-Moll , Javier Romero

Photo-realistic scene reconstruction from sparse-view, uncalibrated images is highly required in practice. Although some successes have been made, existing methods are either Sparse-View but require accurate camera parameters (i.e.,…

计算机视觉与模式识别 · 计算机科学 2024-12-30 Xudong Cai , Yongcai Wang , Zhaoxin Fan , Deng Haoran , Shuo Wang , Wanting Li , Deying Li , Lun Luo , Minhang Wang , Jintao Xu

Recent studies in extreme image compression have achieved remarkable performance by compressing the tokens from generative tokenizers. However, these methods often prioritize clustering common semantics within the dataset, while overlooking…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Naifu Xue , Zhaoyang Jia , Jiahao Li , Bin Li , Yuan Zhang , Yan Lu

Learning-based 3D reconstruction using implicit neural representations has shown promising progress not only at the object level but also in more complicated scenes. In this paper, we propose Dynamic Plane Convolutional Occupancy Networks,…

计算机视觉与模式识别 · 计算机科学 2020-11-12 Stefan Lionar , Daniil Emtsev , Dusan Svilarkovic , Songyou Peng

This work presents a diffusion transformer framework for data-driven structural topology optimization that combines the accuracy of physics-based methods with the efficiency of generative deep learning. Conventional approaches such as the…

计算工程、金融与科学 · 计算机科学 2026-05-05 Aaron Lutheran , Srijan Das , Alireza Tabarraei

Implicit representations are widely used for object reconstruction due to their efficiency and flexibility. In 2021, a novel structure named neural implicit map has been invented for incremental reconstruction. A neural implicit map…

计算机视觉与模式识别 · 计算机科学 2022-06-20 Yijun Yuan , Andreas Nuechter

Accurate modeling of 3D objects exhibiting transparency, reflections and thin structures is an extremely challenging problem. Inspired by billboards and geometric proxies used in computer graphics, this paper proposes Generative Latent…

计算机视觉与模式识别 · 计算机科学 2020-08-12 Ricardo Martin-Brualla , Rohit Pandey , Sofien Bouaziz , Matthew Brown , Dan B Goldman

3D reconstruction from single view images is an ill-posed problem. Inferring the hidden regions from self-occluded images is both challenging and ambiguous. We propose a two-pronged approach to address these issues. To better incorporate…

计算机视觉与模式识别 · 计算机科学 2019-03-27 Priyanka Mandikal , K L Navaneet , Mayank Agarwal , R. Venkatesh Babu

Fringe projection profilometry-based 3-D reconstruction of objects with high reflectivity and low surface roughness remains a significant challenge. When measuring such glossy surfaces, specular reflection and indirect illumination often…

计算机视觉与模式识别 · 计算机科学 2026-02-06 Sanghoon Jeon , Gihyun Jung , Suhyeon Ka , Jae-Sang Hyun

Scientific machine learning has enabled the extraction of physical insights and data-driven modeling of high-dimensional spatiotemporal data, yet achieving physically interpretable latent representations and computationally efficient…

机器学习 · 计算机科学 2026-05-04 Siva Viknesh , Amirhossein Arzani

This paper introduces a 3D point cloud sequence learning model based on inconsistent spatio-temporal propagation for LiDAR odometry, termed DSLO. It consists of a pyramid structure with a spatial information reuse strategy, a sequential…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Huixin Zhang , Guangming Wang , Xinrui Wu , Chenfeng Xu , Mingyu Ding , Masayoshi Tomizuka , Wei Zhan , Hesheng Wang

We present STITCH, a novel approach for neural implicit surface reconstruction of a sparse and irregularly spaced point cloud while enforcing topological constraints (such as having a single connected component). We develop a new…

计算机视觉与模式识别 · 计算机科学 2025-01-10 Anushrut Jignasu , Ethan Herron , Zhanhong Jiang , Soumik Sarkar , Chinmay Hegde , Baskar Ganapathysubramanian , Aditya Balu , Adarsh Krishnamurthy

We introduce Multiresolution Deep Implicit Functions (MDIF), a hierarchical representation that can recover fine geometry detail, while being able to perform global operations such as shape completion. Our model represents a complex 3D…

计算机视觉与模式识别 · 计算机科学 2021-09-17 Zhang Chen , Yinda Zhang , Kyle Genova , Sean Fanello , Sofien Bouaziz , Christian Haene , Ruofei Du , Cem Keskin , Thomas Funkhouser , Danhang Tang

3D LiDAR scene completion from point clouds is a fundamental component of perception systems in autonomous vehicles. Previous methods have predominantly employed diffusion models for high-fidelity reconstruction. However, their multi-step…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Wenzhe He , Xiaojun Chen , Ruiqi Wang , Ruihui Li , Huilong Pi , Jiapeng Zhang , Zhuo Tang , Kenli Li

We present CpT: Convolutional point Transformer - a novel deep learning architecture for dealing with the unstructured nature of 3D point cloud data. CpT is an improvement over existing attention-based Convolutions Neural Networks as well…

计算机视觉与模式识别 · 计算机科学 2021-11-23 Chaitanya Kaul , Joshua Mitton , Hang Dai , Roderick Murray-Smith

While Computed Tomography (CT) reconstruction from X-ray sinograms is necessary for clinical diagnosis, iodine radiation in the imaging process induces irreversible injury, thereby driving researchers to study sparse-view CT reconstruction,…

图像与视频处理 · 电气工程与系统科学 2021-11-29 Ce Wang , Kun Shang , Haimiao Zhang , Qian Li , Yuan Hui , S. Kevin Zhou

Mobile robots operating indoors must be prepared to navigate challenging scenes that contain transparent surfaces. This paper proposes a novel method for the fusion of acoustic and visual sensing modalities through implicit neural…

计算机视觉与模式识别 · 计算机科学 2024-11-08 Advaith V. Sethuraman , Onur Bagoren , Harikrishnan Seetharaman , Dalton Richardson , Joseph Taylor , Katherine A. Skinner

Computed Tomography (CT) is widely used in healthcare for detailed imaging. However, Low-dose CT, despite reducing radiation exposure, often results in images with compromised quality due to increased noise. Traditional methods, including…

图像与视频处理 · 电气工程与系统科学 2024-09-17 Herman Verinaz-Jadan , Su Yan

Recent advancements in Diffusion Transformer (DiT) models have significantly improved 3D point cloud generation. However, existing methods primarily focus on local feature extraction while overlooking global topological information, such as…

计算机视觉与模式识别 · 计算机科学 2025-05-15 Zechao Guan , Feng Yan , Shuai Du , Lin Ma , Qingshan Liu