中文
相关论文

相关论文: Bridging Domain Gaps for Fine-Grained Moth Classif…

200 篇论文

The primary requirement for cross-modal data fusion is the precise alignment of data from different sensors. However, the calibration between LiDAR point clouds and camera images is typically time-consuming and needs external calibration…

计算机视觉与模式识别 · 计算机科学 2025-07-11 Yuanchao Yue , Hui Yuan , Zhengxin Li , Shuai Li , Wei Zhang

Recent advances in large-scale visual representation learning have significantly improved performance in plant species and plant disease recognition tasks. However, state-of-the-art models, often based on high-capacity vision transformers…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Ilyass Moummad , Reda Bensaid , Kawtar Zaher , Hervé Goëau , Jean-Christophe Lombardo , Joseph Salmon , Pierre Bonnet , Alexis Joly

Few-shot learning for fine-grained image classification has gained recent attention in computer vision. Among the approaches for few-shot learning, due to the simplicity and effectiveness, metric-based methods are favorably state-of-the-art…

计算机视觉与模式识别 · 计算机科学 2021-02-03 Xiaoxu Li , Jijie Wu , Zhuo Sun , Zhanyu Ma , Jie Cao , Jing-Hao Xue

Multi-modal (vision-language) models, such as CLIP, are replacing traditional supervised pre-training models (e.g., ImageNet-based pre-training) as the new generation of visual foundation models. These models with robust and aligned…

计算机视觉与模式识别 · 计算机科学 2024-01-05 Fan Liu , Tianshu Zhang , Wenwen Dai , Wenwen Cai , Xiaocong Zhou , Delong Chen

Foundation models have achieved remarkable results in 2D and language tasks like image segmentation, object detection, and visual-language understanding. However, their potential to enrich 3D scene representation learning is largely…

计算机视觉与模式识别 · 计算机科学 2023-11-03 Zhimin Chen , Longlong Jing , Yingwei Li , Bing Li

CLIP models pretrained on natural images with billion-scale image-text pairs have demonstrated impressive capabilities in zero-shot classification, cross-modal retrieval, and open-ended visual answering. However, transferring this success…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Shansong Wang , Zhecheng Jin , Mingzhe Hu , Mojtaba Safari , Feng Zhao , Chih-Wei Chang , Richard LJ Qiu , Justin Roper , David S. Yu , Xiaofeng Yang

Effective pest management is crucial for enhancing agricultural productivity, especially for crops such as sugarcane and wheat that are highly vulnerable to pest infestations. Traditional pest management methods depend heavily on manual…

计算机视觉与模式识别 · 计算机科学 2026-01-05 Anirudha Ghosh , Ritam Sarkar , Debaditya Barman

Adapting vision-language models to remote sensing imagery presents a fundamental challenge: both the visual and linguistic distributions of satellite data lie far outside natural image pretraining corpora. Despite this, prompting remains…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Harshith Kethavath , Weiming Hu

The detection and classification of bacterial colonies in images of agar-plates is important in microbiology, but is hindered by the lack of labeled datasets. Therefore, we propose Colony Grounded SAM2, a zero-shot inference pipeline to…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Daan Korporaal , Patrick de Kruijf , Ralph H. G. M. Litjens , Bas H. M. van der Velden

With the popularity of foundational models, parameter efficient fine tuning has become the defacto approach to leverage pretrained models to perform downstream tasks. Taking inspiration from recent advances in large language models, Visual…

图像与视频处理 · 电气工程与系统科学 2025-01-08 Aadya Arora , Vinay Namboodiri

Automatic plant classification is a challenging problem due to the wide biodiversity of the existing plant species in a fine-grained scenario. Powerful deep learning architectures have been used to improve the classification performance in…

计算机视觉与模式识别 · 计算机科学 2021-10-05 Voncarlos M. Araujo , Alceu S. Britto , Luiz E. S. Oliveira , Alessandro L. Koerich

Transformer-based models have achieved strong performance in remote sensing image captioning by capturing long-range dependencies and contextual information. However, their practical deployment is hindered by high computational costs,…

计算机视觉与模式识别 · 计算机科学 2025-06-12 Swadhin Das , Divyansh Mundra , Priyanshu Dayal , Raksha Sharma

Skin cancer is a major concern to public health, accounting for one-third of the reported cancers. If not detected early, the cancer has the potential for severe consequences. Recognizing the critical need for effective skin cancer…

图像与视频处理 · 电气工程与系统科学 2024-07-01 Niful Islam , Khan Md Hasib , Fahmida Akter Joti , Asif Karim , Sami Azam

Wildlife monitoring is crucial for studying biodiversity loss and climate change. Camera trap images provide a non-intrusive method for analyzing animal populations and identifying ecological patterns over time. However, manual analysis is…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Julian D. Santamaria , Claudia Isaza , Jhony H. Giraldo

Camera traps are vital for large-scale biodiversity monitoring, yet accurate automated analysis remains challenging due to diverse deployment environments. While the computer vision community has mostly framed this challenge as cross-domain…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Sooyoung Jeon , Hongjie Tian , Lemeng Wang , Zheda Mai , Vidhi Bakshi , Jiacheng Hou , Ping Zhang , Arpita Chowdhury , Jianyang Gu , Wei-Lun Chao

In this paper, we propose the LiDAR Distillation to bridge the domain gap induced by different LiDAR beams for 3D object detection. In many real-world applications, the LiDAR points used by mass-produced robots and vehicles usually have…

计算机视觉与模式识别 · 计算机科学 2022-08-16 Yi Wei , Zibu Wei , Yongming Rao , Jiaxin Li , Jie Zhou , Jiwen Lu

In colonoscopy, 80% of the missed polyps could be detected with the help of Deep Learning models. In the search for algorithms capable of addressing this challenge, foundation models emerge as promising candidates. Their zero-shot or…

Model compression and knowledge distillation have been successfully applied for cross-architecture and cross-domain transfer learning. However, a key requirement is that training examples are in correspondence across the domains. We show…

计算机视觉与模式识别 · 计算机科学 2017-08-30 Jong-Chyi Su , Subhransu Maji

Camera-based 3D object detection and tracking are essential for perception in autonomous driving. Current state-of-the-art approaches often rely exclusively on either perspective-view (PV) or bird's-eye-view (BEV) features, limiting their…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Markus Käppeler , Özgün Çiçek , Daniele Cattaneo , Claudius Gläser , Yakov Miron , Abhinav Valada

Accurate and timely identification of plant leaf diseases is essential for resilient and sustainable agriculture, yet most deep learning approaches rely on large annotated datasets and computationally intensive models that are unsuitable…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Anika Islam , Tasfia Tahsin , Zaarin Anjum , Md. Bakhtiar Hasan , Md. Hasanul Kabir