中文
相关论文

相关论文: LAM3D: Large Image-Point-Cloud Alignment Model for…

200 篇论文

Point cloud surface reconstruction has improved in accuracy with advances in deep learning, enabling applications such as infrastructure inspection. Recent approaches that reconstruct from small local regions rather than entire point clouds…

计算机视觉与模式识别 · 计算机科学 2025-11-12 Eito Ogawa , Taiga Hayami , Hiroshi Watanabe

Recovering 3D face models from 2D in-the-wild images has gained considerable attention in the computer vision community due to its wide range of potential applications. However, the lack of ground-truth labeled datasets and the complexity…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Danling Cao

Recent advances in Large Multimodal Models (LMM) have made it possible for various applications in human-machine interactions. However, developing LMMs that can comprehend, reason, and plan in complex and diverse 3D environments remains a…

计算机视觉与模式识别 · 计算机科学 2023-12-01 Sijin Chen , Xin Chen , Chi Zhang , Mingsheng Li , Gang Yu , Hao Fei , Hongyuan Zhu , Jiayuan Fan , Tao Chen

3D modeling based on point clouds is an efficient way to reconstruct and create detailed 3D content. However, the geometric procedure may lose accuracy due to high redundancy and the absence of an explicit structure. In this work, we…

图形学 · 计算机科学 2022-01-28 Xusheng Du , Yi He , Xi Yang , Chia-Ming Chang , Haoran Xie

3D face reconstruction is a fundamental Computer Vision problem of extraordinary difficulty. Current systems often assume the availability of multiple facial images (sometimes from the same subject) as input, and must address a number of…

计算机视觉与模式识别 · 计算机科学 2017-09-11 Aaron S. Jackson , Adrian Bulat , Vasileios Argyriou , Georgios Tzimiropoulos

3D reconstruction from a single view image is a long-standing prob-lem in computer vision. Various methods based on different shape representations(such as point cloud or volumetric representations) have been proposed. However,the 3D shape…

图形学 · 计算机科学 2020-03-10 Aihua Mao , Canglan Dai , Lin Gao , Ying He , Yong-jin Liu

Generating a 3D point cloud from a single 2D image is of great importance for 3D scene understanding applications. To reconstruct the whole 3D shape of the object shown in the image, the existing deep learning based approaches use either…

计算机视觉与模式识别 · 计算机科学 2022-10-11 Yao Wei , George Vosselman , Michael Ying Yang

Reconstructing meshes from point clouds is a fundamental task in computer vision with applications spanning robotics, autonomous systems, and medical imaging. Selecting an appropriate learning-based method requires understanding trade-offs…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Fatima Zahra Iguenfer , Achraf Hsain , Hiba Amissa , Yousra Chtouki

Reconstructing 3D models from single-view images is a long-standing problem in computer vision. The latest advances for single-image 3D reconstruction extract a textual description from the input image and further utilize it to synthesize…

计算机视觉与模式识别 · 计算机科学 2024-11-20 Yu Liu , Ruowei Wang , Jiaqi Li , Zixiang Xu , Qijun Zhao

Shape instantiation which predicts the 3D shape of a dynamic target from one or more 2D images is important for real-time intra-operative navigation. Previously, a general shape instantiation framework was proposed with manual image…

计算机视觉与模式识别 · 计算机科学 2019-07-26 Xiao-Yun Zhou , Zhao-Yang Wang , Peichao Li , Jian-Qing Zheng , Guang-Zhong Yang

Constructing high-quality generative models for 3D shapes is a fundamental task in computer vision with diverse applications in geometry processing, engineering, and design. Despite the recent progress in deep generative modelling,…

计算机视觉与模式识别 · 计算机科学 2021-02-08 Vage Egiazarian , Savva Ignatyev , Alexey Artemov , Oleg Voynov , Andrey Kravchenko , Youyi Zheng , Luiz Velho , Evgeny Burnaev

Multi-beam LiDAR sensors, as used on autonomous vehicles and mobile robots, acquire sequences of 3D range scans ("frames"). Each frame covers the scene sparsely, due to limited angular scanning resolution and occlusion. The sparsity…

计算机视觉与模式识别 · 计算机科学 2022-07-26 Shengyu Huang , Zan Gojcic , Jiahui Huang , Andreas Wieser , Konrad Schindler

Recent open-world 3D representation learning methods using Vision-Language Models (VLMs) to align 3D point cloud with image-text information have shown superior 3D zero-shot performance. However, CAD-rendered images for this alignment often…

计算机视觉与模式识别 · 计算机科学 2024-10-01 Ye Mao , Junpeng Jing , Krystian Mikolajczyk

We present a novel approach for generating isotropic surface triangle meshes directly from unoriented 3D point clouds, with the mesh density adapting to the estimated local feature size (LFS). Popular reconstruction pipelines first…

图形学 · 计算机科学 2025-04-24 Rao Fu , Kai Hormann , Pierre Alliez

The paper presents a simple and effective learning-based method for computing a discriminative 3D point cloud descriptor for place recognition purposes. Recent state-of-the-art methods have relatively complex architectures such as…

计算机视觉与模式识别 · 计算机科学 2022-04-11 Jacek Komorowski

We study 3D shape modeling from a single image and make contributions to it in three aspects. First, we present Pix3D, a large-scale benchmark of diverse image-shape pairs with pixel-level 2D-3D alignment. Pix3D has wide applications in…

计算机视觉与模式识别 · 计算机科学 2018-04-13 Xingyuan Sun , Jiajun Wu , Xiuming Zhang , Zhoutong Zhang , Chengkai Zhang , Tianfan Xue , Joshua B. Tenenbaum , William T. Freeman

Image-to-point cloud registration methods typically follow a coarse-to-fine pipeline, extracting patch-level correspondences and refining them into dense pixel-to-point matches. However, in scenes with repetitive patterns, images often lack…

计算机视觉与模式识别 · 计算机科学 2026-03-30 Zhixin Cheng , Jiacheng Deng , Xinjun Li , Bohao Liao , Li Liu , Xiaotian Yin , Baoqun Yin , Tianzhu Zhang

Digital neuron reconstruction from 3D microscopy images is an essential technique for investigating brain connectomics and neuron morphology. Existing reconstruction frameworks use convolution-based segmentation networks to partition the…

计算机视觉与模式识别 · 计算机科学 2022-10-19 Runkai Zhao , Heng Wang , Chaoyi Zhang , Weidong Cai

Storing and transmitting LiDAR point cloud data is essential for many AV applications, such as training data collection, remote control, cloud services or SLAM. However, due to the sparsity and unordered structure of the data, it is…

计算机视觉与模式识别 · 计算机科学 2024-02-20 Till Beemelmanns , Yuchen Tao , Bastian Lampe , Lennart Reiher , Raphael van Kempen , Timo Woopen , Lutz Eckstein

Large-scale 3D point clouds (LS3DPC) obtained by LiDAR scanners require huge storage space and transmission bandwidth due to a large amount of data. The existing methods of LS3DPC compression separately perform rule-based point sampling and…

计算机视觉与模式识别 · 计算机科学 2024-12-11 Jae-Young Yim , Jae-Young Sim