中文
相关论文

相关论文: Privacy-Preserving Representations are not Enough …

200 篇论文

Recent years have witnessed remarkable progress in developing Vision-Language Models (VLMs) capable of processing both textual and visual inputs. These models have demonstrated impressive performance, leading to their widespread adoption in…

图像与视频处理 · 电气工程与系统科学 2025-07-15 Hanene F. Z. Brachemi Meftah , Wassim Hamidouche , Sid Ahmed Fezza , Olivier Déforges

Visual localization in large and complex indoor scenes, dominated by weakly textured rooms and repeating geometric patterns, is a challenging problem with high practical relevance for applications such as Augmented Reality and robotics. To…

计算机视觉与模式识别 · 计算机科学 2019-09-04 Hajime Taira , Ignacio Rocco , Jiri Sedlar , Masatoshi Okutomi , Josef Sivic , Tomas Pajdla , Torsten Sattler , Akihiko Torii

Location retrieval based on visual information is to retrieve the location of an agent (e.g. human, robot) or the area they see by comparing the observations with a certain form of representation of the environment. Existing methods…

计算机视觉与模式识别 · 计算机科学 2022-08-02 Lijun Wei , Valerie Gouet-Brunet , Anthony Cohn

In this work we present a novel approach to joint semantic localisation and scene understanding. Our work is motivated by the need for localisation algorithms which not only predict 6-DoF camera pose but also simultaneously recognise…

计算机视觉与模式识别 · 计算机科学 2019-09-24 Ignas Budvytis , Marvin Teichmann , Tomas Vojir , Roberto Cipolla

Visual place recognition methods struggle with occlusions and partial visual overlaps. We propose a novel visual place recognition approach based on overlap prediction, called VOP, shifting from traditional reliance on global image…

计算机视觉与模式识别 · 计算机科学 2024-12-05 Tong Wei , Philipp Lindenberger , Jiri Matas , Daniel Barath

Supervised keypoint localization methods rely on large manually labeled image datasets, where objects can deform, articulate, or occlude. However, creating such large keypoint labels is time-consuming and costly, and is often error-prone…

计算机视觉与模式识别 · 计算机科学 2023-03-31 Xingzhe He , Gaurav Bharaj , David Ferman , Helge Rhodin , Pablo Garrido

The success of deep face recognition (FR) systems has raised serious privacy concerns due to their ability to enable unauthorized tracking of users in the digital world. Previous studies proposed introducing imperceptible adversarial noises…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Minghui Li , Jiangxiong Wang , Hao Zhang , Ziqi Zhou , Shengshan Hu , Xiaobing Pei

Camera pose estimation in known scenes is a 3D geometry task recently tackled by multiple learning algorithms. Many regress precise geometric quantities, like poses or 3D points, from an input image. This either fails to generalize to new…

Visual localization is critical to many applications in computer vision and robotics. To address single-image RGB localization, state-of-the-art feature-based methods match local descriptors between a query image and a pre-built 3D model.…

计算机视觉与模式识别 · 计算机科学 2020-04-02 Xiaotian Li , Shuzhe Wang , Yi Zhao , Jakob Verbeek , Juho Kannala

Classical Visual Servoing (VS) rely on handcrafted visual features, which limit their generalizability. Recently, a number of approaches, some based on Deep Neural Networks, have been proposed to overcome this limitation by comparing…

机器人学 · 计算机科学 2022-01-21 Nicholas Adrian , Van-Thach Do , Quang-Cuong Pham

This paper presents an approach for creating a visual place recognition (VPR) database for localization in indoor environments from RGBD scanning sequences. The proposed approach is formulated as a minimization problem in terms of…

计算机视觉与模式识别 · 计算机科学 2024-11-01 Anastasiia Kornilova , Ivan Moskalenko , Timofei Pushkin , Fakhriddin Tojiboev , Rahim Tariverdizadeh , Gonzalo Ferrer

Visual localization plays a critical role in the functionality of low-cost autonomous mobile robots. Current state-of-the-art approaches for achieving accurate visual localization are 3D scene-specific, requiring additional computational…

机器人学 · 计算机科学 2023-09-06 Yanmei Jiao , Binxin Zhang , Peng Jiang , Chaoqun Wang , Rong Xiong , Yue Wang

Image AutoRegressive (IAR) models have achieved state-of-the-art performance in speed and quality of generated images. However, they also raise concerns about memorization of their training data and its implications for privacy. This work…

机器学习 · 计算机科学 2025-09-03 Aditya Kasliwal , Franziska Boenisch , Adam Dziedzic

Visual Simultaneous Localization and Mapping (SLAM) plays a vital role in real-time localization for autonomous systems. However, traditional SLAM methods, which assume a static environment, often suffer from significant localization drift…

机器人学 · 计算机科学 2025-07-30 Haolan Zhang , Thanh Nguyen Canh , Chenghao Li , Nak Young Chong

In recent years, camera-based 3D object detection has gained widespread attention for its ability to achieve high performance with low computational cost. However, the robustness of these methods to adversarial attacks has not been…

计算机视觉与模式识别 · 计算机科学 2024-01-22 Shaoyuan Xie , Zichao Li , Zeyu Wang , Cihang Xie

An appearance-based robot self-localization problem is considered in the machine learning framework. The appearance space is composed of all possible images, which can be captured by a robot's visual system under all robot localizations.…

计算机视觉与模式识别 · 计算机科学 2017-10-06 Alexander Kuleshov , Alexander Bernstein , Evgeny Burnaev , Yury Yanovich

Camera localization methods based on retrieval, local feature matching, and 3D structure-based pose estimation are accurate but require high storage, are slow, and are not privacy-preserving. A method based on scene landmark detection (SLD)…

计算机视觉与模式识别 · 计算机科学 2024-02-01 Tien Do , Sudipta N. Sinha

Camera pose estimation in large-scale environments is still an open question and, despite recent promising results, it may still fail in some situations. The research so far has focused on improving subcomponents of estimation pipelines, to…

计算机视觉与模式识别 · 计算机科学 2020-10-02 Luca Ferranti , Xiaotian Li , Jani Boutellier , Juho Kannala

Vision-Language Models (VLMs) have shown remarkable capabilities across diverse visual tasks, including image recognition, video understanding, and Visual Question Answering (VQA) when explicitly trained for these tasks. Despite these…

Object localization is an important task in computer vision but requires a large amount of computational power due mainly to an exhaustive multiscale search on the input image. In this paper, we describe a near real-time multiscale search…

计算机视觉与模式识别 · 计算机科学 2016-04-14 Hyungtae Lee , Heesung Kwon , Archith J. Bency , William D. Nothwang
‹ 上一页 1 8 9 10 下一页 ›