English
Related papers

Related papers: Parameter-Efficient Adaptation of Pre-Trained Visi…

200 papers

Foundational models are trained on extensive datasets to capture the general trends of a domain. However, in medical imaging, the scarcity of data makes pre-training for every domain, modality, or task challenging. Instead of building…

Computer Vision and Pattern Recognition · Computer Science 2025-11-17 Mohammad Areeb Qazi , Munachiso S Nwadike , Ibrahim Almakky , Mohammad Yaqub , Numan Saeed

Existing vehicle detectors are usually obtained by training a typical detector (e.g., YOLO, RCNN, DETR series) on vehicle images based on a pre-trained backbone (e.g., ResNet, ViT). Some researchers also exploit and enhance the detection…

Computer Vision and Pattern Recognition · Computer Science 2024-08-26 Wentao Wu , Fanghua Hong , Xiao Wang , Chenglong Li , Jin Tang

Seismic full waveform inversion (FWI) has seen promising advancements through deep learning. Existing approaches typically focus on task-specific models trained and evaluated in isolation that lead to limited generalization across different…

Computational Engineering, Finance, and Science · Computer Science 2024-12-30 Koustav Ghosal , Abhranta Panigrahi , Arnav Chavan , ArunSingh , Deepak Gupta

Deep functional map frameworks are widely employed for 3D shape matching. However, most existing deep functional map methods cannot adaptively capture important frequency information for functional map estimation in specific matching…

Computer Vision and Pattern Recognition · Computer Science 2024-06-26 Feifan Luo , Qinsong Li , Ling Hu , Haibo Wang , Xinru Liu , Shengjun Liu , Hongyang Chen

Fluoroscopy is critical for real-time X-ray visualization in medical imaging. However, low-dose images are compromised by noise, potentially affecting diagnostic accuracy. Noise reduction is crucial for maintaining image quality, especially…

Image and Video Processing · Electrical Eng. & Systems 2024-11-05 Sun-Young Jeon , Sen Wang , Adam S. Wang , Garry E. Gold , Jang-Hwan Choi

Contemporary deep learning models have demonstrated promising results across various applications within seismology and earthquake engineering. These models rely primarily on utilizing ground motion records for tasks such as earthquake…

Signal Processing · Electrical Eng. & Systems 2025-05-06 Ümit Mert Çağlar , Baris Yilmaz , Melek Türkmen , Erdem Akagündüz , Salih Tileylioglu

The extraction of weak signals plays a crucial role in quantum precision measurement, where the estimation results are often limited by low signal-to-noise ratios. Here, we demonstrate a parameter-estimation framework based on the adaptive…

Quantum Physics · Physics 2026-05-19 Yihan Wang , Xiaofeng Jin , Yuchuan Ming , Jianxiang Miao , Xiao-Ming Lu , M. W. Mitchell , Jia Kong

Remote Sensing (RS) data encapsulates rich multi-dimensional information essential for Earth observation. Its vast volume, diverse sources, and temporal continuity make it particularly well-suited for developing large Visual Foundation…

Computer Vision and Pattern Recognition · Computer Science 2025-04-22 Xuyang Li , Chenyu Li , Gemine Vivone , Danfeng Hong

Large pre-trained models, also known as foundation models (FMs), are trained in a task-agnostic manner on large-scale data and can be adapted to a wide range of downstream tasks by fine-tuning, few-shot, or even zero-shot learning. Despite…

Artificial Intelligence · Computer Science 2023-04-17 Gengchen Mai , Weiming Huang , Jin Sun , Suhang Song , Deepak Mishra , Ninghao Liu , Song Gao , Tianming Liu , Gao Cong , Yingjie Hu , Chris Cundy , Ziyuan Li , Rui Zhu , Ni Lao

In recent years, deep learning (DL) has emerged as a promising alternative approach for various seismic processing tasks, including primary estimation (or multiple elimination), a crucial step for accurate subsurface imaging. In geophysics,…

Geophysics · Physics 2025-02-11 Jing Sun , Tiexing Wang , Eric Verschuur , Ivan Vasconcelos

Vision Foundation Models (VFMs) are large-scale, pre-trained models that serve as general-purpose backbones for various computer vision tasks. As VFMs' popularity grows, there is an increasing interest in understanding their effectiveness…

Computer Vision and Pattern Recognition · Computer Science 2025-05-06 Volodymyr Havrylov , Haiwen Huang , Dan Zhang , Andreas Geiger

Seismic exploration is currently the most mature approach for studying subsurface structures, yet the presence of noise greatly restricts its imaging accuracy. Previous methods still face significant challenges: traditional computational…

Geophysics · Physics 2025-11-25 Junheng Peng , Yong Li , Yingtian LIu , Mingwei Wang

Dense ground displacement measurements are crucial for geological studies but are impractical to collect directly. Traditionally, displacement fields are estimated using patch matching on optical satellite images from different acquisition…

Computer Vision and Pattern Recognition · Computer Science 2025-04-21 Juliette Bertrand , Sophie Giffard-Roisin , James Hollingsworth , Julien Mairal

Face Anti-Spoofing (FAS) remains challenging due to the requirement for robust domain generalization across unseen environments. While recent trends leverage Vision-Language Models (VLMs) for semantic supervision, these multimodal…

Computer Vision and Pattern Recognition · Computer Science 2026-04-22 Mika Feng , Pierre Gallin-Martel , Koichi Ito , Takafumi Aoki

Reliable displacement measurement is fundamental for structural health monitoring and digital engineering workflows, as it provides direct structural response information. Vision-based measurement has emerged as a promising approach for…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Qingyu Xian , Hao Cheng , Berend Jan van der Zwaag , Rolands Kromanis , Ozlem Durmaz Incel

With the expanding application scope of unmanned aerial vehicles (UAVs), the demand for stable UAV control has significantly increased. However, in complex environments, GPS signals are prone to interference, resulting in ineffective UAV…

Computer Vision and Pattern Recognition · Computer Science 2025-02-25 Mingkun Li , Ziming Wang , Guang Huo , Wei Chen , Xiaoning Zhao

Recent research on fine-tuning vision-language models has demonstrated impressive performance in various downstream tasks. However, the challenge of obtaining accurately labeled data in real-world applications poses a significant obstacle…

Machine Learning · Computer Science 2024-10-01 Tong Wei , Hao-Tian Li , Chun-Shu Li , Jiang-Xin Shi , Yu-Feng Li , Min-Ling Zhang

Advances in Earth observation (EO) foundation models have unlocked the potential of big satellite data to learn generic representations from space, benefiting a wide range of downstream applications crucial to our planet. However, most…

Pre-trained Foundation Models (PFMs) have ushered in a paradigm-shift in Artificial Intelligence, due to their ability to learn general-purpose representations that can be readily employed in a wide range of downstream tasks. While PFMs…

Databases · Computer Science 2024-11-13 Pasquale Balsebre , Weiming Huang , Gao Cong , Yi Li

Zero-shot anomaly detection aims to detect and localise abnormal regions in the image without access to any in-domain training images. While recent approaches leverage vision-language models (VLMs), such as CLIP, to transfer high-level…

Computer Vision and Pattern Recognition · Computer Science 2026-04-10 Matic Fučka , Vitjan Zavrtanik , Danijel Skočaj