English
Related papers

Related papers: 3rd Place Solution to Google Landmark Recognition …

200 papers

In this paper, we introduce a data-efficient instance segmentation method we used in the 2021 VIPriors Instance Segmentation Challenge. Our solution is a modified version of Swin Transformer, based on the mmdetection which is a powerful…

Computer Vision and Pattern Recognition · Computer Science 2022-11-08 Pengyu Chen , Wanhua Li

We present an efficient end-to-end pipeline for largescale landmark recognition and retrieval. We show how to combine and enhance concepts from recent research in image retrieval and introduce two architectures especially suited for…

Computer Vision and Pattern Recognition · Computer Science 2021-10-28 Christof Henkel

This report details our solution to the Google AI Open Images Challenge 2019 Object Detection Track. Based on our detailed analysis on the Open Images dataset, it is found that there are four typical features: large-scale, hierarchical tag…

Computer Vision and Pattern Recognition · Computer Science 2019-10-29 Xingyuan Bu , Junran Peng , Changbao Wang , Cunjun Yu , Guoliang Cao

In this technical report, we present our 1st place solution for the ICDAR 2021 competition on mathematical formula detection (MFD). The MFD task has three key challenges including a large scale span, large variation of the ratio between…

Computer Vision and Pattern Recognition · Computer Science 2021-07-13 Yuxiang Zhong , Xianbiao Qi , Shanjun Li , Dengyi Gu , Yihao Chen , Peiyang Ning , Rong Xiao

In this paper, we present our champion solution to the Global Artificial Intelligence Technology Innovation Competition Track 1: Medical Imaging Diagnosis Report Generation. We select CPT-BASE as our base model for the text generation task.…

Computation and Language · Computer Science 2024-07-08 Xiangyu Wu , Hailiang Zhang , Yang Yang , Jianfeng Lu

We address the challenging problem of RGB image-based head pose estimation. We first reformulate head pose representation learning to constrain it to a bounded space. Head pose represented as vector projection or vector angles shows helpful…

Computer Vision and Pattern Recognition · Computer Science 2020-05-25 Donggen Dai , Wangkit Wong , Zhuojun Chen

We present our solution to Landmark Image Retrieval Challenge 2019. This challenge was based on the large Google Landmarks Dataset V2[9]. The goal was to retrieve all database images containing the same landmark for every provided query…

Computer Vision and Pattern Recognition · Computer Science 2019-06-13 Cheng Chang , Himanshu Rai , Satya Krishna Gorti , Junwei Ma , Chundi Liu , Guangwei Yu , Maksims Volkovs

Landmark localization in images and videos is a classic problem solved in various ways. Nowadays, with deep networks prevailing throughout machine learning, there are revamped interests in pushing facial landmark detection technologies to…

Computer Vision and Pattern Recognition · Computer Science 2019-08-16 Joseph P Robinson , Yuncheng Li , Ning Zhang , Yun Fu , and Sergey Tulyakov

This is a short technical report introducing the solution of Team Rat for Short-video Parsing Face Parsing Track of The 3rd Person in Context (PIC) Workshop and Challenge at CVPR 2021. In this report, we propose an Edge-Aware Network…

Computer Vision and Pattern Recognition · Computer Science 2021-07-15 Xiao Liu , Xiaofei Si , Jiangtao Xie

In this article, we introduce the solution we used in the VSPW 2021 Challenge. Our experiments are based on two baseline models, Swin Transformer and MaskFormer. To further boost performance, we adopt stochastic weight averaging technique…

Computer Vision and Pattern Recognition · Computer Science 2021-12-14 Jiafan Zhuang , Yixin Zhang , Xinyu Hu , Junjie Li , Zilei Wang

High-precision positioning is vital for cellular networks to support innovative applications such as extended reality, unmanned aerial vehicles (UAVs), and industrial Internet of Things (IoT) systems. Existing positioning algorithms using…

Signal Processing · Electrical Eng. & Systems 2025-09-03 Shugong Xu , Jun Jiang , Wenjun Yu , Yilin Gao , Guangjin Pan , Shiyi Mu , Zhiqi Ai , Yuan Gao , Peigang Jiang , Cheng-Xiang Wang

Camera localization methods based on retrieval, local feature matching, and 3D structure-based pose estimation are accurate but require high storage, are slow, and are not privacy-preserving. A method based on scene landmark detection (SLD)…

Computer Vision and Pattern Recognition · Computer Science 2024-02-01 Tien Do , Sudipta N. Sinha

Convolutional neural networks (CNNs) have achieved significant success in image classification by utilizing large-scale datasets. However, it is still of great challenge to learn from scratch on small-scale datasets efficiently and…

Computer Vision and Pattern Recognition · Computer Science 2022-06-14 Yilu Guo , Shicai Yang , Weijie Chen , Liang Ma , Di Xie , Shiliang Pu

The 2021 Image Similarity Challenge introduced a dataset to serve as a new benchmark to evaluate recent image copy detection methods. There were 200 participants to the competition. This paper presents a quantitative and qualitative…

Face alignment aims to estimate the locations of a set of landmarks for a given image. This problem has received much attention as evidenced by the recent advancement in both the methodology and performance. However, most of the existing…

Computer Vision and Pattern Recognition · Computer Science 2015-06-12 Amin Jourabloo , Xiaoming Liu

In this paper, we address the problem of global-scale image geolocation, proposing a mixed classification-retrieval scheme. Unlike other methods that strictly tackle the problem as a classification or retrieval task, we combine the two…

Computer Vision and Pattern Recognition · Computer Science 2021-05-18 Giorgos Kordopatis-Zilos , Panagiotis Galopoulos , Symeon Papadopoulos , Ioannis Kompatsiaris

Although 3D-aware GANs based on neural radiance fields have achieved competitive performance, their applicability is still limited to objects or scenes with the ground-truths or prediction models for clearly defined canonical camera poses.…

Computer Vision and Pattern Recognition · Computer Science 2023-07-04 Mijeong Kim , Hyunjoon Lee , Bohyung Han

3D semantic segmentation is one of the most crucial tasks in driving perception. The ability of a learning-based model to accurately perceive dense 3D surroundings often ensures the safe operation of autonomous vehicles. However, existing…

Computer Vision and Pattern Recognition · Computer Science 2025-01-13 Qing Wu

Recently, it was shown that excellent results can be achieved in both face landmark localization and pose-invariant face recognition. These breakthroughs are attributed to the efforts of the community to manually annotate facial images in…

Computer Vision and Pattern Recognition · Computer Science 2015-02-04 Christos Sagonas , Yannis Panagakis , Stefanos Zafeiriou , Maja Pantic

In this paper, we introduce our approach to the 5th CLVision Challenge, which presents distinctive challenges beyond traditional class incremental learning. Unlike standard settings, this competition features the recurrence of previously…

Computer Vision and Pattern Recognition · Computer Science 2024-06-25 Sishun Pan , Tingmin Li , Yang Yang