中文
相关论文

相关论文: Dynamic Patch-aware Enrichment Transformer for Occ…

200 篇论文

Person re-identification is a crucial task of identifying pedestrians of interest across multiple surveillance camera views. In person re-identification, a pedestrian is usually represented with features extracted from a rectangular image…

计算机视觉与模式识别 · 计算机科学 2021-03-09 Yiheng Liu , Wengang Zhou , Jianzhuang Liu , Guojun Qi , Qi Tian , Houqiang Li

With growing concerns over data privacy, researchers have started using virtual data as an alternative to sensitive real-world images for training person re-identification (Re-ID) models. However, existing virtual datasets produced by game…

计算机视觉与模式识别 · 计算机科学 2025-11-10 Ruolin Li , Min Liu , Yuan Bian , Zhaoyang Li , Yuzhen Li , Xueping Wang , Yaonan Wang

Single-object tracking (SOT) on edge devices is a critical computer vision task, requiring accurate and continuous target localization across video frames under occlusion, distractor interference, and fast motion. However, recent…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Syed Muhammad Raza , Syed Murtaza Hussain Abidi , Khawar Islam , Muhammad Ibrahim , Ajmal Saeed Mian

Reconstructing 3D face models from a single image is an inherently ill-posed problem, which becomes even more challenging in the presence of occlusions. In addition to fewer available observations, occlusions introduce an extra source of…

计算机视觉与模式识别 · 计算机科学 2025-03-04 Pratheba Selvaraju , Victoria Fernandez Abrevaya , Timo Bolkart , Rick Akkerman , Tianyu Ding , Faezeh Amjadi , Ilya Zharkov

Person re-identification (re-id) has made great progress in recent years, but occlusion is still a challenging problem which significantly degenerates the identification performance. In this paper, we design a teacher-student learning…

计算机视觉与模式识别 · 计算机科学 2019-07-09 Jiaxuan Zhuo , Jianhuang Lai , Peijia Chen

State-of-the-art pedestrian detectors have achieved significant progress on non-occluded pedestrians, yet they are still struggling under heavy occlusions. The recent occlusion handling strategy of popular two-stage approaches is to build a…

计算机视觉与模式识别 · 计算机科学 2020-10-22 Ye He , Chao Zhu , Xu-Cheng Yin

In this work, we present Eformer - Edge enhancement based transformer, a novel architecture that builds an encoder-decoder network using transformer blocks for medical image denoising. Non-overlapping window-based self-attention is used in…

图像与视频处理 · 电气工程与系统科学 2021-11-10 Achleshwar Luthra , Harsh Sulakhe , Tanish Mittal , Abhishek Iyer , Santosh Yadav

Sequential DeepFake detection is an emerging task that predicts the manipulation sequence in order. Existing methods typically formulate it as an image-to-sequence problem, employing conventional Transformer architectures. However, these…

计算机视觉与模式识别 · 计算机科学 2025-07-30 Yunfei Li , Yuezun Li , Baoyuan Wu , Junyu Dong , Guopu Zhu , Siwei Lyu

Federated Domain Generalization for Person Re-Identification (FedDG-ReID) learns domain-invariant representations from decentralized data. While Vision Transformer (ViT) is widely adopted, its global attention often fails to distinguish…

计算机视觉与模式识别 · 计算机科学 2026-03-16 Xin Xu , Weilong Li , Wei Liu , Wenke Huang , Zhixi Yu , Bin Yang , Xiaoying Liao , Kui Jiang

Visible-infrared person re-identification (VI-ReID) aims to retrieve images of the same pedestrian from different modalities, where the challenges lie in the significant modality discrepancy. To alleviate the modality gap, recent methods…

计算机视觉与模式识别 · 计算机科学 2024-05-01 Zhihao Qian , Yutian Lin , Bo Du

While deep learning-based models like transformers, have revolutionized time-series and vision tasks, they remain highly susceptible to noise and often overfit on noisy patterns rather than robust features. This issue is exacerbated in…

计算机视觉与模式识别 · 计算机科学 2026-01-12 Ashish Bastola , Nishant Luitel , Hao Wang , Danda Pani Paudel , Roshani Poudel , Abolfazl Razi

Monocular 3D object detection is an important yet challenging task in autonomous driving. Some existing methods leverage depth information from an off-the-shelf depth estimator to assist 3D detection, but suffer from the additional…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Kuan-Chih Huang , Tsung-Han Wu , Hung-Ting Su , Winston H. Hsu

Unsupervised Camoflaged Object Detection (UCOD) has gained attention since it doesn't need to rely on extensive pixel-level labels. Existing UCOD methods typically generate pseudo-labels using fixed strategies and train 1 x1 convolutional…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Weiqi Yan , Lvhai Chen , Huaijia Kou , Shengchuan Zhang , Yan Zhang , Liujuan Cao

Occluded person re-identification is a challenging task as the appearance varies substantially with various obstacles, especially in the crowd scenario. To address this issue, we propose a Pose-guided Visible Part Matching (PVPM) method…

计算机视觉与模式识别 · 计算机科学 2020-04-02 Shang Gao , Jingya Wang , Huchuan Lu , Zimo Liu

Since acquiring large amounts of realistic blurry-sharp image pairs is difficult and expensive, learning blind image deblurring from unpaired data is a more practical and promising solution. Unfortunately, dominant approaches rely heavily…

计算机视觉与模式识别 · 计算机科学 2025-07-21 Chengxu Liu , Lu Qi , Jinshan Pan , Xueming Qian , Ming-Hsuan Yang

Person re-identification (ReID) has made great strides thanks to the data-driven deep learning techniques. However, the existing benchmark datasets lack diversity, and models trained on these data cannot generalize well to dynamic wild…

计算机视觉与模式识别 · 计算机科学 2024-03-25 Lei Zhang , Xiaowei Fu , Fuxiang Huang , Yi Yang , Xinbo Gao

Recent open-world representation learning approaches have leveraged CLIP to enable zero-shot 3D object recognition. However, performance on real point clouds with occlusions still falls short due to unrealistic pretraining settings.…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Khanh Nguyen , Ghulam Mubashar Hassan , Ajmal Mian

Accurate surgical instrument segmentation in endoscopy is crucial for computer-assisted interventions, yet remains challenging due to frequent occlusions, rapid motion, and long-term instrument re-entry. While SAM3 provides a powerful…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Valay Bundele , Mehran Hosseinzadeh , Hendrik P. A. Lensch

Facial expression analysis is central to understanding human behavior, yet existing coding systems such as the Facial Action Coding System (FACS) are constrained by limited coverage and costly manual annotation. In this work, we introduce…

计算机视觉与模式识别 · 计算机科学 2025-10-03 Minh Tran , Maksim Siniukov , Zhangyu Jin , Mohammad Soleymani

Driver distraction remains a leading cause of traffic accidents, posing a critical threat to road safety globally. As intelligent transportation systems evolve, accurate and real-time identification of driver distraction has become…

计算机视觉与模式识别 · 计算机科学 2024-09-13 Junzhou Chen , Zirui Zhang , Jing Yu , Heqiang Huang , Ronghui Zhang , Xuemiao Xu , Bin Sheng , Hong Yan