中文
相关论文

相关论文: BUDA.ART: A Multimodal Content-Based Analysis and …

200 篇论文

The explosive growth of digital images in video surveillance and social media has led to the significant need for efficient search of persons of interest in law enforcement and forensic applications. Despite tremendous progress in primary…

计算机视觉与模式识别 · 计算机科学 2018-11-02 Hu Han , Jie Li , Anil K. Jain , Shiguang Shan , Xilin Chen

Image set recognition has been widely applied in many practical problems like real-time video retrieval and image caption tasks. Due to its superior performance, it has grown into a significant topic in recent years. However, images with…

计算机视觉与模式识别 · 计算机科学 2020-08-25 Chuan-Xian Ren , You-Wei Luo , Xiao-Lin Xu , Dao-Qing Dai , Hong Yan

The performance of CBIR algorithms is usually measured on an isolated workstation. In a real-world environment the algorithms would only constitute a minor component among the many interacting components. The Internet dramati-cally changes…

信息检索 · 计算机科学 2015-06-25 Neil J. Gunther , Giordano B. Beretta

Text-based visual question answering (VQA) requires to read and understand text in an image to correctly answer a given question. However, most current methods simply add optical character recognition (OCR) tokens extracted from the image…

计算机视觉与模式识别 · 计算机科学 2020-10-27 Zan-Xia Jin , Heran Wu , Chun Yang , Fang Zhou , Jingyan Qin , Lei Xiao , Xu-Cheng Yin

Incompatibility of image descriptor and ranking is always neglected in image retrieval. In this paper, manifold learning and Gestalt psychology theory are involved to solve the incompatibility problem. A new holistic descriptor called…

计算机视觉与模式识别 · 计算机科学 2016-09-27 Shenglan Liu , Jun Wu , Lin Feng , Yang Liu , Hong Qiao , Wenbo Luo Muxin Sun , Wei Wang

Visible-infrared person re-identification (VI-ReID) aims to match persons captured by visible and infrared cameras, allowing person retrieval and tracking in 24-hour surveillance systems. Previous methods focus on learning from…

计算机视觉与模式识别 · 计算机科学 2023-11-28 Yunhao Du , Cheng Lei , Zhicheng Zhao , Yuan Dong , Fei Su

Current methods for 3D generation still fall short in physically based rendering (PBR) texturing, primarily due to limited data and challenges in modeling multi-channel materials. In this work, we propose MuMA, a method for 3D PBR texturing…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Lingting Zhu , Jingrui Ye , Runze Zhang , Zeyu Hu , Yingda Yin , Lanjiong Li , Jinnan Chen , Shengju Qian , Xin Wang , Qingmin Liao , Lequan Yu

We present ALADIN (All Layer AdaIN); a novel architecture for searching images based on the similarity of their artistic style. Representation learning is critical to visual search, where distance in the learned search embedding reflects…

计算机视觉与模式识别 · 计算机科学 2021-03-18 Dan Ruta , Saeid Motiian , Baldo Faieta , Zhe Lin , Hailin Jin , Alex Filipkowski , Andrew Gilbert , John Collomosse

We present Holo-Artisan, a novel system architecture enabling immersive multi-user experiences in virtual museums through true holographic displays and personalized edge intelligence. In our design, local edge computing nodes process…

多媒体 · 计算机科学 2025-08-22 Nan-Hong Kuo , Hojjat Baghban

Blind facial image restoration is highly challenging due to unknown complex degradations and the sensitivity of humans to faces. Although existing methods introduce auxiliary information from generative priors or high-quality reference…

计算机视觉与模式识别 · 计算机科学 2025-07-15 Zhicun Yin , Junjie Chen , Ming Liu , Zhixin Wang , Fan Li , Renjing Pei , Xiaoming Li , Rynson W. H. Lau , Wangmeng Zuo

3D face reconstruction (3DFR) algorithms are based on specific assumptions tailored to distinct application scenarios. These assumptions limit their use when acquisition conditions, such as the subject's distance from the camera or the…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Simone Maurizio La Cava , Sara Concas , Ruben Tolosana , Roberto Casula , Giulia Orrù , Martin Drahansky , Julian Fierrez , Gian Luca Marcialis

We introduce MosAIc, an interactive web app that allows users to find pairs of semantically related artworks that span different cultures, media, and millennia. To create this application, we introduce Conditional Image Retrieval (CIR)…

Recent work has established the ecological importance of developing algorithms for identifying animals individually from images. Typically, a separate algorithm is trained for each species, a natural step but one that creates significant…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Lasha Otarashvili , Tamilselvan Subramanian , Jason Holmberg , J. J. Levenson , Charles V. Stewart

The IRMA project aims to design innovative methodologies for research in the field of historical and archaeological heritage based on a combination of medical imaging technologies and interactive 3D restitution modalities (virtual reality,…

图形学 · 计算机科学 2023-01-27 Théophane Nicolas , Ronan Gaugne , Bruno Arnaldi , Valérie Gouranton

Recent research has widely explored the problem of aesthetics assessment of images with generic content. However, few approaches have been specifically designed to predict the aesthetic quality of images containing human faces, which make…

计算机视觉与模式识别 · 计算机科学 2018-10-03 Simone Bianco , Luigi Celona , Raimondo Schettini

This paper introduces the large scale visual search algorithm and system infrastructure at Alibaba. The following challenges are discussed under the E-commercial circumstance at Alibaba (a) how to handle heterogeneous image data and bridge…

计算机视觉与模式识别 · 计算机科学 2021-02-10 Yanhao Zhang , Pan Pan , Yun Zheng , Kang Zhao , Yingya Zhang , Xiaofeng Ren , Rong Jin

Composed Image Retrieval (CIR) enables image retrieval by combining multiple query modalities, but existing benchmarks predominantly focus on general-domain imagery and rely on reference images with short textual modifications. As a result,…

信息检索 · 计算机科学 2026-04-21 Jinyu Xu , Yi Sun , Jiangling Zhang , Qing Xie , Daomin Ji , Zhifeng Bao , Jiachen Li , Yanchun Ma , Yongjian Liu

We present 3D Pick & Mix, a new 3D shape retrieval system that provides users with a new level of freedom to explore 3D shape and Internet image collections by introducing the ability to reason about objects at the level of their…

计算机视觉与模式识别 · 计算机科学 2018-11-06 Adrian Penate-Sanchez , Lourdes Agapito

Most existing robotic datasets capture static scene data and thus are limited in evaluating robots' dynamic performance. To address this, we present a mobile robot oriented large-scale indoor dataset, denoted as THUD (Tsinghua University…

机器人学 · 计算机科学 2024-07-02 Yifan Tang , Cong Tai , Fangxing Chen , Wanting Zhang , Tao Zhang , Xueping Liu , Yongjin Liu , Long Zeng

Biodiversity research requires complete and detailed information to study ecosystem dynamics at different scales. Employing data-driven methods like Machine Learning is getting traction in ecology and more specific biodiversity, offering…

定量方法 · 定量生物学 2025-10-27 Stylianos Stasinos , Martino Mensio , Elena Lazovik , Athanasios Trantas