English
Related papers

Related papers: ART3mis: Ray-Based Textual Annotation on 3D Cultur…

200 papers

During interactive segmentation, a model and a user work together to delineate objects of interest in a 3D point cloud. In an iterative process, the model assigns each data point to an object (or the background), while the user corrects…

Computer Vision and Pattern Recognition · Computer Science 2024-04-11 Yuanwen Yue , Sabarinath Mahadevan , Jonas Schult , Francis Engelmann , Bastian Leibe , Konrad Schindler , Theodora Kontogianni

Virtual content placement in physical scenes is a crucial aspect of augmented reality (AR). This task is particularly challenging when the virtual elements must adapt to multiple target physical environments that are unknown during…

Human-Computer Interaction · Computer Science 2024-11-11 Arthur Caetano , Misha Sra

We introduce Referring 3D Gaussian Splatting Segmentation (R3DGS), a new task that aims to segment target objects in a 3D Gaussian scene based on natural language descriptions, which often contain spatial relationships or object attributes.…

Computer Vision and Pattern Recognition · Computer Science 2025-08-12 Shuting He , Guangquan Jie , Changshuo Wang , Yun Zhou , Shuming Hu , Guanbin Li , Henghui Ding

Interactive 3D scenes are increasingly vital for embodied intelligence, yet existing datasets remain limited due to the labor-intensive process of annotating part segmentation, kinematic types, and motion trajectories. We present REACT3D, a…

Computer Vision and Pattern Recognition · Computer Science 2026-04-13 Zhao Huang , Boyang Sun , Alexandros Delitzas , Jiaqi Chen , Marc Pollefeys

Current approaches to the annotation process focus on annotation schemas, languages for annotation, or are very application driven. In this paper it is proposed that a more flexible architecture for annotation requires a knowledge component…

Digital Libraries · Computer Science 2007-05-23 Afzal Ballim , Nastaran Fatemi , Hatem Ghorbel , Vincenzo Pallotta

In this paper, we propose a self-supervised learningmethod for multi-object pose estimation. 3D object under-standing from 2D image is a challenging task that infers ad-ditional dimension from reduced-dimensional information.In particular,…

Computer Vision and Pattern Recognition · Computer Science 2021-04-16 Hyeonwoo Yu , Jean Oh

The rapid development of Large Multimodal Models (LMMs) for 2D images and videos has spurred efforts to adapt these models for interpreting 3D scenes. However, the absence of large-scale 3D vision-language datasets has posed a significant…

Computer Vision and Pattern Recognition · Computer Science 2025-04-03 Haochen Wang , Yucheng Zhao , Tiancai Wang , Haoqiang Fan , Xiangyu Zhang , Zhaoxiang Zhang

Relation context has been proved to be useful for many challenging vision tasks. In the field of 3D object detection, previous methods have been taking the advantage of context encoding, graph embedding, or explicit relation reasoning to…

Computer Vision and Pattern Recognition · Computer Science 2025-03-06 Yuqing Lan , Yao Duan , Chenyi Liu , Chenyang Zhu , Yueshan Xiong , Hui Huang , Kai Xu

We introduce ReplaceAnything3D model (RAM3D), a novel text-guided 3D scene editing method that enables the replacement of specific objects within a scene. Given multi-view images of a scene, a text prompt describing the object to replace,…

Computer Vision and Pattern Recognition · Computer Science 2024-02-01 Edward Bartrum , Thu Nguyen-Phuoc , Chris Xie , Zhengqin Li , Numair Khan , Armen Avetisyan , Douglas Lanman , Lei Xiao

Class-agnostic 3D instance segmentation tackles the challenging task of segmenting all object instances, including previously unseen ones, without semantic class reliance. Current methods struggle with generalization due to the scarce…

Computer Vision and Pattern Recognition · Computer Science 2025-12-11 Shengchao Zhou , Jiehong Lin , Jiahui Liu , Shizhen Zhao , Chirui Chang , Xiaojuan Qi

Tactile recognition of 3D objects remains a challenging task. Compared to 2D shapes, the complex geometry of 3D surfaces requires richer tactile signals, more dexterous actions, and more advanced encoding techniques. In this work, we…

Computer Vision and Pattern Recognition · Computer Science 2023-03-01 Jingxi Xu , Han Lin , Shuran Song , Matei Ciocarlie

Unsupervised Domain Adaptation (UDA) technique has been explored in 3D cross-domain tasks recently. Though preliminary progress has been made, the performance gap between the UDA-based 3D model and the supervised one trained with fully…

Computer Vision and Pattern Recognition · Computer Science 2023-03-13 Jiakang Yuan , Bo Zhang , Xiangchao Yan , Tao Chen , Botian Shi , Yikang Li , Yu Qiao

High-fidelity 3D scanning is essential for preserving cultural heritage artefacts, supporting documentation, analysis, and long-term conservation. However, conventional methods typically require specialized expertise and manual intervention…

Computer Vision and Pattern Recognition · Computer Science 2025-10-15 Javed Ahmad , Federico Dassiè , Selene Frascella , Gabriele Marchello , Ferdinando Cannella , Arianna Traviglia

It is human to want to touch artworks, to feel their surface curvature and texture, their shapes and structures, and to feel the hand of the artist. Museum guards need to be constantly vigilant to protect art objects from adoring and…

Human-Computer Interaction · Computer Science 2021-10-05 Bernice Rogowitz , Laura J. Perovich , Yuke Li , Bjorn Kierulf , Dietmar Offenhuber

In this paper, we explore the existing challenges in 3D artistic scene generation by introducing ART3D, a novel framework that combines diffusion models and 3D Gaussian splatting techniques. Our method effectively bridges the gap between…

Computer Vision and Pattern Recognition · Computer Science 2024-05-20 Pengzhi Li , Chengshuai Tang , Qinxuan Huang , Zhiheng Li

3D object editing is essential for interactive content creation in gaming, animation, and robotics, yet current approaches remain inefficient, inconsistent, and often fail to preserve unedited regions. Most methods rely on editing…

Computer Vision and Pattern Recognition · Computer Science 2025-10-20 Junliang Ye , Shenghao Xie , Ruowen Zhao , Zhengyi Wang , Hongyu Yan , Wenqiang Zu , Lei Ma , Jun Zhu

Despite a growing number of datasets being collected for training 3D object detection models, significant human effort is still required to annotate 3D boxes on LiDAR scans. To automate the annotation and facilitate the production of…

Computer Vision and Pattern Recognition · Computer Science 2022-07-21 Chang Liu , Xiaoyan Qian , Binxiao Huang , Xiaojuan Qi , Edmund Lam , Siew-Chong Tan , Ngai Wong

This paper presents Objaverse++, a curated subset of Objaverse enhanced with detailed attribute annotations by human experts. Recent advances in 3D content generation have been driven by large-scale datasets such as Objaverse, which…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Chendi Lin , Heshan Liu , Qunshu Lin , Zachary Bright , Shitao Tang , Yihui He , Minghao Liu , Ling Zhu , Cindy Le

Creating 3D semantic reconstructions of environments is fundamental to many applications, especially when related to autonomous agent operation (e.g., goal-oriented navigation or object interaction and manipulation). Commonly, 3D semantic…

Robotics · Computer Science 2024-06-11 Jianhao Zheng , Daniel Barath , Marc Pollefeys , Iro Armeni

3D reconstruction from a single-RGB image in unconstrained real-world scenarios presents numerous challenges due to the inherent diversity and complexity of objects and environments. In this paper, we introduce Anything-3D, a methodical…

Computer Vision and Pattern Recognition · Computer Science 2023-04-21 Qiuhong Shen , Xingyi Yang , Xinchao Wang