English
Related papers

Related papers: Simple, Effective and General: A New Backbone for …

200 papers

Vision based localization is the problem of inferring the pose of the camera given a single image. One solution to this problem is to learn a deep neural network to infer the pose of a query image after learning on a dataset of images with…

Machine Learning · Computer Science 2019-11-11 Carlos Lassance , Yasir Latif , Ravi Garg , Vincent Gripon , Ian Reid

Due to the rapid increase in the diversity of image data, the problem of domain generalization has received increased attention recently. While domain generalization is a challenging problem, it has achieved great development thanks to the…

Computer Vision and Pattern Recognition · Computer Science 2021-10-14 Cuicui Kang , Karthik Nandakumar

Planet-scale photo geolocalization involves the intricate task of estimating the geographic location depicted in an image purely based on its visual features. While deep learning models, particularly convolutional neural networks (CNNs),…

Computer Vision and Pattern Recognition · Computer Science 2026-03-26 David Faget , José Luis Lisani , Miguel Colom

Development of human machine interface has become a necessity for modern day machines to catalyze more autonomy and more efficiency. Gaze driven human intervention is an effective and convenient option for creating an interface to alleviate…

Computer Vision and Pattern Recognition · Computer Science 2023-03-23 Somsukla Maiti , Akshansh Gupta

The rapid development of urban low-altitude unmanned aerial vehicle (UAV) economy poses new challenges for dynamic site selection of UAV landing points and supply stations. Traditional deep reinforcement learning methods face computational…

Machine Learning · Computer Science 2025-07-16 Jianing Zhi , Xinghua Li , Zidong Chen

Most works on person re-identification (ReID) take advantage of large backbone networks such as ResNet, which are designed for image classification instead of ReID, for feature extraction. However, these backbones may not be computationally…

Computer Vision and Pattern Recognition · Computer Science 2021-04-12 Hanjun Li , Gaojie Wu , Wei-Shi Zheng

Network backbones provide useful sparse representations of weighted networks by keeping only their most important links, permitting a range of computational speedups and simplifying network visualizations. A key limitation of existing…

Social and Information Networks · Computer Science 2025-06-13 Alec Kirkley

Learning continuous image representations is recently gaining popularity for image super-resolution (SR) because of its ability to reconstruct high-resolution images with arbitrary scales from low-resolution inputs. Existing methods mostly…

Computer Vision and Pattern Recognition · Computer Science 2023-04-14 Jiezhang Cao , Qin Wang , Yongqin Xian , Yawei Li , Bingbing Ni , Zhiming Pi , Kai Zhang , Yulun Zhang , Radu Timofte , Luc Van Gool

The existing work in cross-view geo-localization is based on images where a ground panorama is matched to an aerial image. In this work, we focus on ground videos instead of images which provides additional contextual cues which are…

Computer Vision and Pattern Recognition · Computer Science 2022-07-07 Shruti Vyas , Chen Chen , Mubarak Shah

Visual localization plays an important role for intelligent robots and autonomous driving, especially when the accuracy of GNSS is unreliable. Recently, camera localization in LiDAR maps has attracted more and more attention for its low…

Computer Vision and Pattern Recognition · Computer Science 2024-10-28 Zhipeng Zhao , Huai Yu , Chenwei Lyv , Wen Yang , Sebastian Scherer

Existing spatial localization techniques for autonomous vehicles mostly use a pre-built 3D-HD map, often constructed using a survey-grade 3D mapping vehicle, which is not only expensive but also laborious. This paper shows that by using an…

Computer Vision and Pattern Recognition · Computer Science 2023-04-21 Shan Wang , Yanhao Zhang , Ankit Vora , Akhil Perincherry , Hongdong Li

The recent proliferation of photorealistic AI-generated images (AIGI) has raised urgent concerns about their potential misuse, particularly on social media platforms. Current state-of-the-art AIGI detection methods typically rely on large,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-08 Nicholas Chivaran , Jianbing Ni

Deep clustering has attracted increasing attention in recent years due to its capability of joint representation learning and clustering via deep neural networks. In its latest developments, the contrastive learning has emerged as an…

Machine Learning · Computer Science 2022-07-15 Xiaozhi Deng , Dong Huang , Ding-Hua Chen , Chang-Dong Wang , Jian-Huang Lai

We introduce a novel backbone architecture to improve target-perception ability of feature representation for tracking. Specifically, having observed that de facto frameworks perform feature matching simply using the outputs from backbone…

Computer Vision and Pattern Recognition · Computer Science 2022-01-10 Mingzhe Guo , Zhipeng Zhang , Heng Fan , Liping Jing , Yilin Lyu , Bing Li , Weiming Hu

We study the image-based geolocalization problem, aiming to localize ground-view query images on cartographic maps. Current methods often utilize cross-view localization techniques to match ground-view query images with 2D maps. However,…

Computer Vision and Pattern Recognition · Computer Science 2023-11-06 Mengjie Zhou , Liu Liu , Yiran Zhong , Andrew Calway

Cross-view geo-localization is a promising solution for large-scale localization problems, requiring the sequential execution of retrieval and metric localization tasks to achieve fine-grained predictions. However, existing methods…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Zhuo Song , Ye Zhang , Kunhong Li , Longguang Wang , Yulan Guo

Existing vehicle re-identification methods commonly use spatial pooling operations to aggregate feature maps extracted via off-the-shelf backbone networks. They ignore exploring the spatial significance of feature maps, eventually degrading…

Computer Vision and Pattern Recognition · Computer Science 2021-07-13 Fei Shen , Jianqing Zhu , Xiaobin Zhu , Yi Xie , Jingchang Huang

Convolutions (Convs) and multi-head self-attentions (MHSAs) are typically considered alternatives to each other for building vision backbones. Although some works try to integrate both, they apply the two operators simultaneously at the…

Computer Vision and Pattern Recognition · Computer Science 2024-11-22 Lei Zhu , Xinjiang Wang , Wayne Zhang , Rynson W. H. Lau

Matching local features across images is a fundamental problem in computer vision. Targeting towards high accuracy and efficiency, we propose Seeded Graph Matching Network, a graph neural network with sparse structure to reduce redundant…

Computer Vision and Pattern Recognition · Computer Science 2021-08-20 Hongkai Chen , Zixin Luo , Jiahui Zhang , Lei Zhou , Xuyang Bai , Zeyu Hu , Chiew-Lan Tai , Long Quan

Transformers have recently gained significant attention in the computer vision community. However, the lack of scalability of self-attention mechanisms with respect to image size has limited their wide adoption in state-of-the-art vision…

Computer Vision and Pattern Recognition · Computer Science 2022-09-12 Zhengzhong Tu , Hossein Talebi , Han Zhang , Feng Yang , Peyman Milanfar , Alan Bovik , Yinxiao Li
‹ Prev 1 4 5 6 7 8 10 Next ›