中文
相关论文

相关论文: Global-Supervised Contrastive Loss and View-Aware-…

200 篇论文

Visual Place Recognition (VPR) is a scene-oriented image retrieval problem in computer vision in which re-ranking based on local features is commonly employed to improve performance. In robotics, VPR is also referred to as Loop Closure…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Bingxi Liu , Hao Chen , Shiyi Guo , Yihong Wu , Jinqiang Cui , Hong Zhang

Vehicle Re-identification (re-id) over surveillance camera network with non-overlapping field of view is an exciting and challenging task in intelligent transportation systems (ITS). Due to its versatile applicability in metropolitan…

计算机视觉与模式识别 · 计算机科学 2021-02-22 Zakria , Jianhua Deng , Muhammad Saddam Khokhar , Muhammad Umar Aftab , Jingye Cai , Rajesh Kumar , Jay Kumar

Language-supervised vision models have recently attracted great attention in computer vision. A common approach to build such models is to use contrastive learning on paired data across the two modalities, as exemplified by Contrastive…

机器学习 · 计算机科学 2023-03-16 Ryumei Nakada , Halil Ibrahim Gulluk , Zhun Deng , Wenlong Ji , James Zou , Linjun Zhang

Multi-UAV collaborative 3D detection enables accurate and robust perception by fusing multi-view observations from aerial platforms, offering significant advantages in coverage and occlusion handling, while posing new challenges for…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Zhongyao Li , Peirui Cheng , Liangjin Zhao , Chen Chen , Yundu Li , Zhechao Wang , Xue Yang , Xian Sun , Zhirui Wang

Contrastive pre-training on distant supervision has shown remarkable effectiveness in improving supervised relation extraction tasks. However, the existing methods ignore the intrinsic noise of distant supervision during the pre-training…

计算与语言 · 计算机科学 2023-02-13 Zhen Wan , Fei Cheng , Qianying Liu , Zhuoyuan Mao , Haiyue Song , Sadao Kurohashi

Visual-only self-supervised learning has achieved significant improvement in video representation learning. Existing related methods encourage models to learn video representations by utilizing contrastive learning or designing specific…

计算机视觉与模式识别 · 计算机科学 2022-04-12 Jinyu Liu , Ying Cheng , Yuejie Zhang , Rui-Wei Zhao , Rui Feng

We tackle the cross-modal retrieval problem, where learning is only supervised by relevant multi-modal pairs in the data. Although the contrastive learning is the most popular approach for this task, it makes potentially wrong assumption…

机器学习 · 计算机科学 2022-10-13 Minyoung Kim

Multimodal image-text contrastive learning has shown that joint representations can be learned across modalities. Here, we show how leveraging multiple views of image data with contrastive learning can improve downstream fine-grained…

计算机视觉与模式识别 · 计算机科学 2024-10-01 Andy V. Huynh , Lauren E. Gillespie , Jael Lopez-Saucedo , Claire Tang , Rohan Sikand , Moisés Expósito-Alonso

In autonomous driving, using a variety of sensors to recognize preceding vehicles in middle and long distance is helpful for improving driving performance and developing various functions. However, if only LiDAR or camera is used in the…

机器人学 · 计算机科学 2021-03-26 Hyunjin Bae , Gu Lee , Jaeseung Yang , Gwanjun Shin , Yongseob Lim , Gyeungho Choi

Foundation models have recently gained attention within the field of machine learning thanks to its efficiency in broad data processing. While researchers had attempted to extend this success to time series models, the main challenge is…

机器学习 · 计算机科学 2023-11-22 Trang H. Tran , Lam M. Nguyen , Kyongmin Yeo , Nam Nguyen , Roman Vaculin

With the development of smart cities, urban surveillance video analysis will play a further significant role in intelligent transportation systems. Identifying the same target vehicle in large datasets from non-overlapping cameras should be…

计算机视觉与模式识别 · 计算机科学 2020-03-17 Huibing Wang , Jinjia Peng , Guangqi Jiang , Fengqiang Xu , Xianping Fu

The advent of graph convolutional network (GCN)-based multi-view learning provides a powerful framework for integrating structural information from heterogeneous views, enabling effective modeling of complex multi-view data. However,…

机器学习 · 计算机科学 2025-12-17 Huaiyuan Xiao , Fadi Dornaika , Jingjun Bi

Visual localization is a crucial component in the application of mobile robot and autonomous driving. Image retrieval is an efficient and effective technique in image-based localization methods. Due to the drastic variability of…

计算机视觉与模式识别 · 计算机科学 2021-10-15 Hanjiang Hu , Hesheng Wang , Zhe Liu , Weidong Chen

In the past few years, contrastive learning has played a central role for the success of visual unsupervised representation learning. Around the same time, high-performance non-contrastive learning methods have been developed as well. While…

计算机视觉与模式识别 · 计算机科学 2024-01-12 Jaeill Kim , Duhun Hwang , Eunjung Lee , Jangwon Suh , Jimyeong Kim , Wonjong Rhee

Cross-Domain Few-Shot Learning (CD-FSL) aims to transfer knowledge from seen source domains to unseen target domains, which is crucial for evaluating the generalization and robustness of models. Recent studies focus on utilizing visual…

计算机视觉与模式识别 · 计算机科学 2024-12-13 Wenqian Li , Pengfei Fang , Hui Xue

Understanding visual inputs for a given task amidst varied changes is a key challenge posed by visual reinforcement learning agents. We propose \textit{Value Explicit Pretraining} (VEP), a method that learns generalizable representations…

机器学习 · 计算机科学 2026-05-04 Kiran Lekkala , Henghui Bao , Sumedh A. Sontakke , Erdem Biyik , Laurent Itti

Vehicle re-identification (V-reID) has become significantly popular in the community due to its applications and research significance. In particular, the V-reID is an important problem that still faces numerous open challenges. This paper…

计算机视觉与模式识别 · 计算机科学 2019-06-03 Sultan Daud Khan , Habib Ullah

In visual place recognition (VPR), filtering and sequence-based matching approaches can improve performance by integrating temporal information across image sequences, especially in challenging conditions. While these methods are commonly…

计算机视觉与模式识别 · 计算机科学 2025-07-30 Somayeh Hussaini , Tobias Fischer , Michael Milford

Vehicle re-identification (re-ID) focuses on matching images of the same vehicle across different cameras. It is fundamentally challenging because differences between vehicles are sometimes subtle. While several studies incorporate…

计算机视觉与模式识别 · 计算机科学 2020-10-13 Tsai-Shien Chen , Chih-Ting Liu , Chih-Wei Wu , Shao-Yi Chien

Vehicle Re-ID has recently attracted enthusiastic attention due to its potential applications in smart city and urban surveillance. However, it suffers from large intra-class variation caused by view variations and illumination changes, and…

计算机视觉与模式识别 · 计算机科学 2021-11-11 Hongchao Li , Xianmin Lin , Aihua Zheng , Chenglong Li , Bin Luo , Ran He , Amir Hussain