English
Related papers

Related papers: Ref-SAM3D: Bridging SAM3D with Text for Reference …

200 papers

Event cameras are rapidly emerging as powerful vision sensors for 3D reconstruction, uniquely capable of asynchronously capturing per-pixel brightness changes. Compared to traditional frame-based cameras, event cameras produce sparse yet…

Computer Vision and Pattern Recognition · Computer Science 2025-12-23 Chuanzhi Xu , Haoxian Zhou , Langyi Chen , Haodong Chen , Zeke Zexi Hu , Zhicheng Lu , Ying Zhou , Vera Chung , Qiang Qu , Weidong Cai

Electron microscopy (EM) images exhibit anisotropic axial resolution due to the characteristics inherent to the imaging modality, presenting challenges in analysis and downstream tasks.In this paper, we propose a diffusion-model-based…

Computer Vision and Pattern Recognition · Computer Science 2023-08-04 Kyungryun Lee , Won-Ki Jeong

In this paper, we rethink the problem of scene reconstruction from an embodied agent's perspective: While the classic view focuses on the reconstruction accuracy, our new perspective emphasizes the underlying functions and constraints such…

Robotics · Computer Science 2021-03-31 Muzhi Han , Zeyu Zhang , Ziyuan Jiao , Xu Xie , Yixin Zhu , Song-Chun Zhu , Hangxin Liu

In this paper, we propose a novel Visual Reference Prompt (VRP) encoder that empowers the Segment Anything Model (SAM) to utilize annotated reference images as prompts for segmentation, creating the VRP-SAM model. In essence, VRP-SAM can…

Computer Vision and Pattern Recognition · Computer Science 2025-11-04 Yanpeng Sun , Jiahui Chen , Shan Zhang , Xinyu Zhang , Xiaofan Li , Qiang Chen , Gang Zhang , Errui Ding , Jingdong Wang , Zechao Li

This paper presents a semi-automated framework for transforming two-dimensional miniatures from medieval manuscripts into three-dimensional digital models suitable for extended reality (XR), tactile 3D~printing, and web-based visualization.…

Computer Vision and Pattern Recognition · Computer Science 2026-04-13 Riccardo Pallotto , Pierluigi Feliciati , Tiberio Uricchio

In this paper, we introduce SAM3-UNet, a simplified variant of Segment Anything Model 3 (SAM3), designed to adapt SAM3 for downstream tasks at a low cost. Our SAM3-UNet consists of three components: a SAM3 image encoder, a simple adapter…

Computer Vision and Pattern Recognition · Computer Science 2025-12-02 Xinyu Xiong , Zihuang Wu , Lei Lu , Yufa Xia

Open-vocabulary 3D Scene Graph (3DSG) can enhance various downstream tasks in robotics by leveraging structured semantic representations, yet current 3DSG construction methods suffer from semantic inconsistencies caused by noisy cross-image…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Yue Chang , Rufeng Chen , Zhaofan Zhang , Yi Chen , Yifan Tian , Sihong Xie

Symmetry is a ubiquitous and fundamental property in the visual world, serving as a critical cue for perception and structure interpretation. This paper investigates the detection of 3D reflection symmetry from a single RGB image, and…

Computer Vision and Pattern Recognition · Computer Science 2024-11-28 Xiang Li , Zixuan Huang , Anh Thai , James M. Rehg

Reconstructing and understanding 3D structures from a limited number of images is a well-established problem in computer vision. Traditional methods usually break this task into multiple subtasks, each requiring complex transformations…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Zhiwen Fan , Jian Zhang , Wenyan Cong , Peihao Wang , Renjie Li , Kairun Wen , Shijie Zhou , Achuta Kadambi , Zhangyang Wang , Danfei Xu , Boris Ivanovic , Marco Pavone , Yue Wang

Segmenting objects with complex shapes, such as wires, bicycles, or structural grids, remains a significant challenge for current segmentation models, including the Segment Anything Model (SAM) and its high-quality variant SAM-HQ. These…

Computer Vision and Pattern Recognition · Computer Science 2025-06-09 Luka Vetoshkin , Dmitry Yudin

Accurate segmentation of medical images is fundamental to tumor diagnosis and treatment planning. SAM-based interactive segmentation has gained attention for its strong generalization, but most methods follow a single-point-to-single-object…

Computer Vision and Pattern Recognition · Computer Science 2025-11-04 Jierui Qu , Jianchun Zhao

Reconstructing real-world objects and estimating their movable joint structures are pivotal technologies within the field of robotics. Previous research has predominantly focused on supervised approaches, relying on extensively annotated…

Computer Vision and Pattern Recognition · Computer Science 2024-01-18 Haowen Wang , Zhen Zhao , Zhao Jin , Zhengping Che , Liang Qiao , Yakun Huang , Zhipeng Fan , Xiuquan Qiao , Jian Tang

The release of SAM 3D Body is a recent development in human mesh recovery, demonstrating improved performance in producing clean, topologically coherent meshes from single images. By leveraging the Momentum Human Rig (MHR), it achieves…

Graphics · Computer Science 2026-05-05 Aizierjiang Aiersilan , Ruting Cheng , James Hahn

3D Gaussian Splatting (3D-GS) enables real-time 3D scene reconstruction but lacks robust segmentation for editing tasks such as object removal, extraction, and recoloring. Existing approaches that lift 2D segmentations to the 3D domain…

Computer Vision and Pattern Recognition · Computer Science 2026-05-18 Raushan Joshi , Jean-Yves Guillemaut

We introduce ReplaceAnything3D model (RAM3D), a novel text-guided 3D scene editing method that enables the replacement of specific objects within a scene. Given multi-view images of a scene, a text prompt describing the object to replace,…

Computer Vision and Pattern Recognition · Computer Science 2024-02-01 Edward Bartrum , Thu Nguyen-Phuoc , Chris Xie , Zhengqin Li , Numair Khan , Armen Avetisyan , Douglas Lanman , Lei Xiao

Reconstructing 3D shapes from a single image plays an important role in computer vision. Many methods have been proposed and achieve impressive performance. However, existing methods mainly focus on extracting semantic information from…

Computer Vision and Pattern Recognition · Computer Science 2025-03-03 Shaoming Li , Qing Cai , Songqi Kong , Runqing Tan , Heng Tong , Shiji Qiu , Yongguo Jiang , Zhi Liu

Single-view 3D reconstruction is currently approached from two dominant perspectives: reconstruction of scenes with limited diversity using 3D data supervision or reconstruction of diverse singular objects using large image priors. However,…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Andreea Ardelean , Mert Özer , Bernhard Egger

Recovering the 3D geometry of a purely texture-less object with generally unknown surface reflectance (e.g. non-Lambertian) is regarded as a challenging task in multi-view reconstruction. The major obstacle revolves around establishing…

Computer Vision and Pattern Recognition · Computer Science 2021-05-26 Ziang Cheng , Hongdong Li , Yuta Asano , Yinqiang Zheng , Imari Sato

Recently, methods like Zero-1-2-3 have focused on single-view based 3D reconstruction and have achieved remarkable success. However, their predictions for unseen areas heavily rely on the inductive bias of large-scale pretrained diffusion…

Computer Vision and Pattern Recognition · Computer Science 2024-09-13 Hao Chen , Jiafu Wu , Ying Jin , Jinlong Peng , Xiaofeng Mao , Mingmin Chi , Mufeng Yao , Bo Peng , Jian Li , Yun Cao

Texture editing is a crucial task in 3D modeling that allows users to automatically manipulate the surface materials of 3D models. However, the inherent complexity of 3D models and the ambiguous text description lead to the challenge in…

Computer Vision and Pattern Recognition · Computer Science 2024-09-26 Shengqi Liu , Zhuo Chen , Jingnan Gao , Yichao Yan , Wenhan Zhu , Jiangjing Lyu , Xiaokang Yang
‹ Prev 1 4 5 6 7 8 10 Next ›