English
Related papers

Related papers: Multi-Scale Target-Aware Representation Learning f…

200 papers

Multi-task feature learning aims to identity the shared features among tasks to improve generalization. It has been shown that by minimizing non-convex learning models, a better solution than the convex alternatives can be obtained.…

Machine Learning · Computer Science 2015-06-03 Yaru Fan , Yilun Wang

Pixel-aligned implicit models, such as PIFu, PIFuHD, and ICON, are used for single-view clothed human reconstruction. These models need to be trained using a sampling training scheme. Existing sampling training schemes either fail to…

Computer Vision and Pattern Recognition · Computer Science 2024-11-12 Kennard Yanting Chan , Fayao Liu , Guosheng Lin , Chuan Sheng Foo , Weisi Lin

Direct RAW-based object detection offers great promise by utilizing RAW data (unprocessed sensor data), but faces inherent challenges due to its wide dynamic range and linear response, which tends to suppress crucial object details. In…

Computer Vision and Pattern Recognition · Computer Science 2025-08-07 Zhuohua Ye , Liming Zhang , Hongru Han

The development of federated learning (FL) methods, which aim to learn from distributed databases (i.e., clients) without accessing data on clients, has recently attracted great attention. Most of these methods assume that the clients are…

Computer Vision and Pattern Recognition · Computer Science 2023-06-02 Barış Büyüktaş , Gencer Sumbul , Begüm Demir

Masked Image Modeling (MIM) has garnered significant attention in self-supervised learning, thanks to its impressive capacity to learn scalable visual representations tailored for downstream tasks. However, images inherently contain…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Wenzhao Xiang , Chang Liu , Hongyang Yu , Xilin Chen

Image restoration aims to recover high-quality images from their corrupted counterparts. Many existing methods primarily focus on the spatial domain, neglecting the understanding of frequency variations and ignoring the impact of implicit…

Computer Vision and Pattern Recognition · Computer Science 2024-07-15 Hu Gao , Depeng Dang

This paper introduces Multi-Level feature learning alongside the Embedding layer of Convolutional Autoencoder (CAE-MLE) as a novel approach in deep clustering. We use agglomerative clustering as the multi-level feature learning that…

Computer Vision and Pattern Recognition · Computer Science 2020-10-07 Behzad Ghazanfari , Fatemeh Afghah

In the realm of medical image fusion, integrating information from various modalities is crucial for improving diagnostics and treatment planning, especially in retinal health, where the important features exhibit differently in different…

Image and Video Processing · Electrical Eng. & Systems 2024-07-22 Xin Tian , Nantheera Anantrasirichai , Lindsay Nicholson , Alin Achim

Existing deep learning methods in multimode fiber (MMF) imaging often focus on simpler datasets, limiting their applicability to complex, real-world imaging tasks. These models are typically data-intensive, a challenge that becomes more…

Computer Vision and Pattern Recognition · Computer Science 2025-11-26 Jawaria Maqbool , M. Imran Cheema

Multiple Sclerosis (MS) is a chronic progressive neurological disease characterized by the development of lesions in the white matter of the brain. T2-fluid-attenuated inversion recovery (FLAIR) brain magnetic resonance imaging (MRI)…

Image and Video Processing · Electrical Eng. & Systems 2022-09-12 Jueqi Wang , Derek Berger , Erin Mazerolle , Othman Soufan , Jacob Levman

Deformable medical image registration is a crucial aspect of medical image analysis. In recent years, researchers have begun leveraging auxiliary tasks (such as supervised segmentation) to provide anatomical structure information for the…

Computer Vision and Pattern Recognition · Computer Science 2024-10-01 Hongchao Zhou , Shunbo Hu

Automatic diagnosis techniques have evolved to identify age-related macular degeneration (AMD) by employing single modality Fundus images or optical coherence tomography (OCT). To classify ocular diseases, fundus and OCT images are the most…

Image and Video Processing · Electrical Eng. & Systems 2024-09-04 Pragya Gupta , Subhamoy Mandal , Debashree Guha , Debjani Chakraborty

Automated radiology report generation is essential for improving diagnostic efficiency and reducing the workload of medical professionals. However, existing methods face significant challenges, such as disease class imbalance and…

Methodology · Statistics 2025-07-11 Qin Zhou , Guoyan Liang , Xindi Li , Jingyuan Chen , Wang Zhe , Chang Yao , Sai Wu

Federated learning (FL) can be used to improve data privacy and efficiency in magnetic resonance (MR) image reconstruction by enabling multiple institutions to collaborate without needing to aggregate local data. However, the domain shift…

Image and Video Processing · Electrical Eng. & Systems 2022-08-24 Chun-Mei Feng , Yunlu Yan , Shanshan Wang , Yong Xu , Ling Shao , Huazhu Fu

As FMs drive progress toward Artificial General Intelligence (AGI), fine-tuning them under privacy and resource constraints has become increasingly critical particularly when highquality training data resides on distributed edge devices.…

Machine Learning · Computer Science 2025-08-27 Gang Hu , Yinglei Teng , Pengfei Wu , Nan Wang

Medical ultrasound provides images which are the spatial map of the tissue echogenicity. Unfortunately, an ultrasound image is a low-quality version of the expected Tissue Reflectivity Function (TRF) mainly due to the non-ideal Point Spread…

Image and Video Processing · Electrical Eng. & Systems 2021-09-28 Sobhan Goudarzi , Hassan Rivaz

The growing burden of myopia and retinal diseases necessitates more accessible and efficient eye screening solutions. This study presents a compact, dual-function optical device that integrates fundus photography and refractive error…

Image and Video Processing · Electrical Eng. & Systems 2025-04-29 Boyuan Peng , Jiaju Chen , Yiwei Zhang , Cuiyi Peng , Junyang Li , Jiaming Deng , Peiwu Qin

Deep learning-based methods have achieved encouraging performances in the field of magnetic resonance (MR) image reconstruction. Nevertheless, to properly learn a powerful and robust model, these methods generally require large quantities…

Image and Video Processing · Electrical Eng. & Systems 2023-04-18 Ruoyou Wu , Cheng Li , Juan Zou , Qiegen Liu , Hairong Zheng , Shanshan Wang

Cross-modality magnetic resonance (MR) image synthesis can be used to generate missing modalities from given ones. Existing (supervised learning) methods often require a large number of paired multi-modal data to train an effective…

Image and Video Processing · Electrical Eng. & Systems 2023-06-21 Yonghao Li , Tao Zhou , Kelei He , Yi Zhou , Dinggang Shen

Magnetic Resonance Imaging (MRI) field-strength enhancement holds immense value for both clinical diagnostics and advanced research. However, existing methods typically focus on isolated enhancement tasks, such as specific 64mT-to-3T or…

Computer Vision and Pattern Recognition · Computer Science 2026-03-11 Yiyang Lin , Chenhui Wang , Zhihao Peng , Yixuan Yuan