中文
相关论文

相关论文: ART3mis: Ray-Based Textual Annotation on 3D Cultur…

200 篇论文

The digitisation campaigns carried out by libraries and archives in recent years have facilitated access to documents in their collections. However, exploring and exploiting these documents remain difficult tasks due to the sheer quantity…

数字图书馆 · 计算机科学 2024-03-29 Nicolas Gutehrlé , Iana Atanassova

We present a dataset of 998 3D models of everyday tabletop objects along with their 847,000 real world RGB and depth images. Accurate annotations of camera poses and object poses for each image are performed in a semi-automated fashion to…

计算机视觉与模式识别 · 计算机科学 2022-08-10 Rakesh Shrestha , Siqi Hu , Minghao Gou , Ziyuan Liu , Ping Tan

Existing image editing tools, while powerful, typically disregard the underlying 3D geometry from which the image is projected. As a result, edits made using these tools may become detached from the geometry and lighting conditions that are…

计算机视觉与模式识别 · 计算机科学 2023-07-21 Oscar Michel , Anand Bhattad , Eli VanderBilt , Ranjay Krishna , Aniruddha Kembhavi , Tanmay Gupta

We present an automatic method for annotating images of indoor scenes with the CAD models of the objects by relying on RGB-D scans. Through a visual evaluation by 3D experts, we show that our method retrieves annotations that are at least…

计算机视觉与模式识别 · 计算机科学 2022-12-23 Stefan Ainetter , Sinisa Stekovic , Friedrich Fraundorfer , Vincent Lepetit

In the last decades the rapid development of technologies and methodologies in the field of digitization and 3D modelling has led to an increasing proliferation of 3D technologies in the Cultural Heritage domain. Despite the great potential…

数字图书馆 · 计算机科学 2021-06-15 Nicola Amico , Achille Felicetti

Annotating 3D data remains a costly bottleneck for 3D object detection, motivating the development of weakly supervised annotation methods that rely on more accessible 2D box annotations. However, relying solely on 2D boxes introduces…

计算机视觉与模式识别 · 计算机科学 2025-09-10 Saad Lahlali , Alexandre Fournier Montgieux , Nicolas Granger , Hervé Le Borgne , Quoc Cuong Pham

Distant viewing approaches have typically used image datasets close to the contemporary image data used to train machine learning models. To work with images from other historical periods requires expert annotated data, and the quality of…

计算机视觉与模式识别 · 计算机科学 2024-04-12 Christofer Meinecke , Estelle Guéville , David Joseph Wrisley , Stefan Jänicke

Videos carry rich visual information including object description, action, interaction, etc., but the existing multimodal large language models (MLLMs) fell short in referential understanding scenarios such as video-based referring. In this…

计算机视觉与模式识别 · 计算机科学 2024-06-04 Jihao Qiu , Yuan Zhang , Xi Tang , Lingxi Xie , Tianren Ma , Pengyu Yan , David Doermann , Qixiang Ye , Yunjie Tian

The increase in data collection has made data annotation an interesting and valuable task in the contemporary world. This paper presents a new methodology for quickly annotating data using click-supervision and hierarchical object…

机器学习 · 计算机科学 2018-10-02 Adithya Subramanian , Anbumani Subramanian

Multi-instrument music transcription aims to convert polyphonic music recordings into musical scores assigned to each instrument. This task is challenging for modeling as it requires simultaneously identifying multiple instruments and…

音频与语音处理 · 电气工程与系统科学 2024-08-02 Sungkyun Chang , Emmanouil Benetos , Holger Kirchhoff , Simon Dixon

We introduce a new dataset for graphical object detection in business documents, more specifically annual reports. This dataset, IIIT-AR-13k, is created by manually annotating the bounding boxes of graphical or page objects in publicly…

计算机视觉与模式识别 · 计算机科学 2020-08-07 Ajoy Mondal , Peter Lipps , C. V. Jawahar

We propose a training-free method, Articulate3D, to pose a 3D asset through language control. Despite advances in vision and language models, this task remains surprisingly challenging. To achieve this goal, we decompose the problem into…

计算机视觉与模式识别 · 计算机科学 2025-08-27 Oishi Deb , Anjun Hu , Ashkan Khakzar , Philip Torr , Christian Rupprecht

3D semantic scene understanding is essential for digital twins, autonomous driving, smart agriculture, and embodied perception, yet dense point-wise annotation for point clouds remains expensive and difficult to scale. Existing…

计算机视觉与模式识别 · 计算机科学 2026-05-14 Yijing Wang , Ruonan Li , Qilin Wang , Rongqiang Zhao , Jie Liu

The performance of information retrieval algorithms depends upon the availability of ground truth labels annotated by experts. This is an important prerequisite, and difficulties arise when the annotated ground truth labels are incorrect or…

信息检索 · 计算机科学 2018-02-22 Ekta Vats , Anders Hast

Generating high-fidelity 3D content from text prompts remains a significant challenge in computer vision due to the limited size, diversity, and annotation depth of the existing datasets. To address this, we introduce MARVEL-40M+, an…

计算机视觉与模式识别 · 计算机科学 2025-03-27 Sankalp Sinha , Mohammad Sadil Khan , Muhammad Usama , Shino Sam , Didier Stricker , Sk Aziz Ali , Muhammad Zeshan Afzal

Accurate video annotation plays a vital role in modern retail applications, including customer behavior analysis, product interaction detection, and in-store activity recognition. However, conventional annotation methods heavily rely on…

计算机视觉与模式识别 · 计算机科学 2025-06-23 Varun Mannam , Zhenyu Shi

Image collections, if critical aspects of image content are exposed, can spur research and practical applications in many domains. Supervised machine learning may be the only feasible way to annotate very large collections, but leading…

计算机视觉与模式识别 · 计算机科学 2019-03-01 Sara Mousavi , Ramin Nabati , Megan Kleeschulte , Audris Mockus

Rapid advancements in text-to-3D generation require robust and scalable evaluation metrics that align closely with human judgment, a need unmet by current metrics such as PSNR and CLIP, which require ground-truth data or focus only on…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Shalini Maiti , Lourdes Agapito , Filippos Kokkinos

Weakly supervised 3D object detection aims to learn a 3D detector with lower annotation cost, e.g., 2D labels. Unlike prior work which still relies on few accurate 3D annotations, we propose a framework to study how to leverage constraints…

计算机视觉与模式识别 · 计算机科学 2024-08-22 Kuan-Chih Huang , Yi-Hsuan Tsai , Ming-Hsuan Yang

Large-scale pre-trained image-to-3D generative models have exhibited remarkable capabilities in diverse shape generations. However, most of them struggle to synthesize plausible 3D assets when the reference image is flat-colored like hand…

计算机视觉与模式识别 · 计算机科学 2026-04-20 Xiaoyan Cong , Jiayi Shen , Zekun Li , Rao Fu , Tao Lu , Srinath Sridhar
‹ 上一页 1 8 9 10 下一页 ›