English
Related papers

Related papers: DINOv3 Beats Specialized Detectors: A Simple Found…

200 papers

Extracting a Bird's Eye View (BEV) representation from multiple camera images offers a cost-effective, scalable alternative to LIDAR-based solutions in autonomous driving. However, the performance of the existing BEV methods drops…

Computer Vision and Pattern Recognition · Computer Science 2024-09-17 Merve Rabia Barın , Görkay Aydemir , Fatma Güney

Generative models have shown a giant leap in synthesizing photo-realistic images with minimal expertise, sparking concerns about the authenticity of online information. This study aims to develop a universal AI-generated image detector…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Zihan Liu , Hanyi Wang , Yaoyu Kang , Shilin Wang

In this work, we present a panoramic metric depth foundation model that generalizes across diverse scene distances. We explore a data-in-the-loop paradigm from the view of both data construction and framework design. We collect a…

Computer Vision and Pattern Recognition · Computer Science 2025-12-19 Xin Lin , Meixi Song , Dizhe Zhang , Wenxuan Lu , Haodong Li , Bo Du , Ming-Hsuan Yang , Truong Nguyen , Lu Qi

The demand for high-resolution subsurface imaging and continuous Earth monitoring has driven rapid growth in active and passive seismic data from dense geophone deployments, distributed acoustic sensing (DAS) arrays, and large-scale 2D and…

Geophysics · Physics 2026-05-13 Jiahua Zhao , Umair bin Waheed , Jing Sun , Yang Cui , Nikos Savva , Eric Verschuur

Lacking rich and realistic data, learned single image denoising algorithms generalize poorly to real raw images that do not resemble the data used for training. Although the problem can be alleviated by the heteroscedastic Gaussian model…

Image and Video Processing · Electrical Eng. & Systems 2020-04-10 Kaixuan Wei , Ying Fu , Jiaolong Yang , Hua Huang

Existing state-of-the-art AI-Generated image detection methods mostly consider extracting low-level information from RGB images to help improve the generalization of AI-Generated image detection, such as noise patterns. However, these…

Computer Vision and Pattern Recognition · Computer Science 2025-07-18 Ziyin Zhou , Ke Sun , Zhongxi Chen , Xianming Lin , Yunpeng Luo , Ke Yan , Shouhong Ding , Xiaoshuai Sun

In this paper, we propose to reformulate the blind image deblurring task to directly learn an inverse of the degradation model represented by a deep linear network. We introduce Deep Identity Learning (DIL), a novel learning strategy that…

Computer Vision and Pattern Recognition · Computer Science 2024-11-06 Vamsidhar Saraswathula , Rama Krishna Gorthi

The lack of large-scale noisy-clean image pairs restricts supervised denoising methods' deployment in actual applications. While existing unsupervised methods are able to learn image denoising without ground-truth clean images, they either…

Computer Vision and Pattern Recognition · Computer Science 2022-03-23 Yi Zhang , Dasong Li , Ka Lung Law , Xiaogang Wang , Hongwei Qin , Hongsheng Li

Fine-tuning a deep network trained with the standard cross-entropy loss is a strong baseline for few-shot learning. When fine-tuned transductively, this outperforms the current state-of-the-art on standard datasets such as Mini-ImageNet,…

Machine Learning · Computer Science 2020-10-23 Guneet S. Dhillon , Pratik Chaudhari , Avinash Ravichandran , Stefano Soatto

LiDAR registration is a fundamental task in robotic mapping and localization. A critical component of aligning two point clouds is identifying robust point correspondences using point descriptors. This step becomes particularly challenging…

Robotics · Computer Science 2025-02-27 Niclas Vödisch , Giovanni Cioffi , Marco Cannici , Wolfram Burgard , Davide Scaramuzza

The deep learning field is converging towards the use of general foundation models that can be easily adapted for diverse tasks. While this paradigm shift has become common practice within the field of natural language processing, progress…

Computer Vision and Pattern Recognition · Computer Science 2023-11-15 Joana Palés Huix , Adithya Raju Ganeshan , Johan Fredin Haslum , Magnus Söderberg , Christos Matsoukas , Kevin Smith

Recently, there have been significant advancements in Image Restoration based on CNN and transformer. However, the inherent characteristics of the Image Restoration task are often overlooked in many works. They, instead, tend to focus on…

Computer Vision and Pattern Recognition · Computer Science 2024-06-25 Dongqi Fan , Ting Yue , Xin Zhao , Renjing Xu , Liang Chang

Local feature matching has long been a fundamental component of 3D vision systems such as Structure-from-Motion (SfM), yet progress has lagged behind the rapid advances of modern data-driven approaches. The newer approaches, such as…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 David Nordström , Johan Edstedt , Georg Bökman , Jonathan Astermark , Anders Heyden , Viktor Larsson , Mårten Wadenbäck , Michael Felsberg , Fredrik Kahl

Foundation models such as DINOv2 have shown strong performance in few-shot anomaly detection, yet two key questions remain unexamined: (i) how susceptible are these detectors to adversarial perturbations; and (ii) how well do their anomaly…

Computer Vision and Pattern Recognition · Computer Science 2025-10-16 Akib Mohammed Khan , Bartosz Krawczyk

Remote tiny face detection applied in unmanned system is a challeng-ing work. The detector cannot obtain sufficient context semantic information due to the relatively long distance. The received poor fine-grained features make the face…

Computer Vision and Pattern Recognition · Computer Science 2020-10-12 Jia-Yi Chang , Yan-Feng Lu , Ya-Jun Liu , Bo Zhou , Hong Qiao

Parameter-efficient fine-tuning (PEFT) has emerged as a popular strategy for adapting large vision foundation models, such as the Segment Anything Model (SAM) and LLaVA, to downstream tasks like image forgery detection and localization…

Computer Vision and Pattern Recognition · Computer Science 2025-08-19 Rongxuan Peng , Shunquan Tan , Chenqi Kong , Anwei Luo , Alex C. Kot , Jiwu Huang

Anomaly detection methods typically require extensive normal samples from the target class for training, limiting their applicability in scenarios that require rapid adaptation, such as cold start. Zero-shot and few-shot anomaly detection…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Zhaopeng Gu , Bingke Zhu , Guibo Zhu , Yingying Chen , Ming Tang , Jinqiao Wang

Foundation models pre-trained on large-scale datasets demonstrate strong transfer learning capabilities; however, their adaptation to complex multi-label diagnostic tasks-such as comprehensive head CT finding detection-remains understudied.…

Foundation models (e.g., CLIP or DINOv2) have shown their impressive learning and transfer capabilities in a wide range of visual tasks, by training on a large corpus of data and adapting to specific downstream tasks. It is, however,…

Machine Learning · Computer Science 2023-11-06 Bin Deng , Kui Jia

Despite the significant advancements in general image segmentation achieved by large-scale pre-trained foundation models (such as Meta's Segment Any-thing Model (SAM) series and DINOv2), their performance in specialized fields remains…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Yimin Xu , Fan Yang , Bin Xu