中文
相关论文

相关论文: XoFTR: Cross-modal Feature Matching Transformer

200 篇论文

Existing deep Thermal InfraRed (TIR) trackers only use semantic features to describe the TIR object, which lack the sufficient discriminative capacity for handling distractors. This becomes worse when the feature extraction network is only…

计算机视觉与模式识别 · 计算机科学 2019-06-11 Qiao Liu , Xin Li , Zhenyu He , Nana Fan , Di Yuan , Hongpeng Wang

Graph-based models have achieved great success in person re-identification tasks recently, which compute the graph topology structure (affinities) among different people first and then pass the information across them to achieve stronger…

计算机视觉与模式识别 · 计算机科学 2022-11-15 Xulin Li , Yan Lu , Bin Liu , Yating Liu , Guojun Yin , Qi Chu , Jinyang Huang , Feng Zhu , Rui Zhao , Nenghai Yu

Cross-modal place recognition methods are flexible GPS-alternatives under varying environment conditions and sensor setups. However, this task is non-trivial since extracting consistent and robust global descriptors from different…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Yun-Jin Li , Mariia Gladkova , Yan Xia , Rui Wang , Daniel Cremers

Current multispectral object detection methods often retain extraneous background or noise during feature fusion, limiting perceptual performance. To address this, we propose an innovative feature fusion framework based on cross-modal…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Jifeng Shen , Haibo Zhan , Xin Zuo , Heng Fan , Xiaohui Yuan , Jun Li , Wankou Yang

Thermal Infrared (TIR) technology involves the use of sensors to detect and measure infrared radiation emitted by objects, and it is widely utilized across a broad spectrum of applications. The advancements in object detection methods…

计算机视觉与模式识别 · 计算机科学 2025-05-19 Jinke Li , Yue Wu , Xiaoyan Yang

Visible-infrared person re-identification (VI-ReID) technique could associate the pedestrian images across visible and infrared modalities in the practical scenarios of background illumination changes. However, a substantial gap inherently…

计算机视觉与模式识别 · 计算机科学 2025-11-05 Chao Yuan , Zanwu Liu , Guiwei Zhang , Haoxuan Xu , Yujian Zhao , Guanglin Niu , Bo Li

This paper presents a novel approach for visible-thermal infrared stereoscopy, focusing on the estimation of disparities of human silhouettes. Visible-thermal infrared stereo poses several challenges, including occlusions and differently…

计算机视觉与模式识别 · 计算机科学 2023-04-25 Noreen Anwar , Philippe Duplessis-Guindon , Guillaume-Alexandre Bilodeau , Wassim Bouachir

The goal of multimodal image fusion is to integrate complementary information from infrared and visible images, generating multimodal fused images for downstream tasks. Existing downstream pre-training models are typically trained on…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Yushen Xu , Xiaosong Li , Zhenyu Kuang , Xiaoqi Cheng , Haishu Tan , Huafeng Li

Thermal infrared (IR) images represent the heat patterns emitted from hot object and they do not consider the energies reflected from an object. Objects living or non-living emit different amounts of IR energy according to their body…

计算机视觉与模式识别 · 计算机科学 2013-09-06 Ayan Seal , Suranjan Ganguly , Debotosh Bhattacharjee , Mita Nasipuri , Dipak kr. Basu

The resolving ability of widefield fluorescence microscopy is fundamentally limited by out-of-focus background owing to its low axial resolution, particularly for densely labeled biological samples. Although total internal reflection…

The ability to form images through hair-thin optical fibres promises to open up new applications from biomedical imaging to industrial inspection. Unfortunately, deployment has been limited because small changes in mechanical deformation…

Interest in thermal to visible face recognition has grown significantly over the last decade due to advancements in thermal infrared cameras and analytics beyond the visible spectrum. Despite large discrepancies between thermal and visible…

计算机视觉与模式识别 · 计算机科学 2022-11-18 Cedric Nimpa Fondje , Shuowen Hu , Benjamin S. Riggan

We present an effective method for the matching of multimodal images. Accurate image matching is the basis of various applications, such as image registration and structure from motion. Conventional matching methods fail when handling noisy…

计算机视觉与模式识别 · 计算机科学 2023-03-01 Zhongli Fan , Li Zhang , Yuxuan Liu

Integrating multispectral data in object detection, especially visible and infrared images, has received great attention in recent years. Since visible (RGB) and infrared (IR) images can provide complementary information to handle light…

计算机视觉与模式识别 · 计算机科学 2022-09-29 Maoxun Yuan , Yinyan Wang , Xingxing Wei

Visible-infrared person re-identification (VI-ReID) aims to match individuals across different camera modalities, a critical task in modern surveillance systems. While current VI-ReID methods focus on cross-modality matching, real-world…

计算机视觉与模式识别 · 计算机科学 2025-01-24 Mahdi Alehdaghi , Rajarshi Bhattacharya , Pourya Shamsolmoali , Rafael M. O. Cruz , Eric Granger

Accurate multispectral image matching presents significant challenges due to non-linear intensity variations across spectral modalities, extreme viewpoint changes, and the scarcity of labeled datasets. Current state-of-the-art methods are…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Ismail Can Yagmur , Hasan F. Ates , Bahadir K. Gunturk

We propose X-NeRF, a novel method to learn a Cross-Spectral scene representation given images captured from cameras with different light spectrum sensitivity, based on the Neural Radiance Fields formulation. X-NeRF optimizes camera poses…

计算机视觉与模式识别 · 计算机科学 2022-09-02 Matteo Poggi , Pierluigi Zama Ramirez , Fabio Tosi , Samuele Salti , Stefano Mattoccia , Luigi Di Stefano

Image fusion aims to synthesize a single high-quality image from a pair of inputs captured under challenging conditions, such as differing exposure levels or focal depths. A core challenge lies in effectively handling disparities in dynamic…

计算机视觉与模式识别 · 计算机科学 2025-12-24 Mingwei Tang , Jiahao Nie , Guang Yang , Ziqing Cui , Jie Li

In this work, we present a unified framework for multi-modality 3D object detection, named UVTR. The proposed method aims to unify multi-modality representations in the voxel space for accurate and robust single- or cross-modality 3D…

计算机视觉与模式识别 · 计算机科学 2022-10-14 Yanwei Li , Yilun Chen , Xiaojuan Qi , Zeming Li , Jian Sun , Jiaya Jia

Visual tracking often faces challenges such as invalid targets and decreased performance in low-light conditions when relying solely on RGB image sequences. While incorporating additional modalities like depth and infrared data has proven…

计算机视觉与模式识别 · 计算机科学 2023-12-25 Lei Liu , Mengya Zhang , Cheng Li , Chenglong Li , Jin Tang