English
Related papers

Related papers: Investigating the Impact of Rational Dilated Wavel…

200 papers

MRI super-resolution (SR) and denoising tasks are fundamental challenges in the field of deep learning, which have traditionally been treated as distinct tasks with separate paired training data. In this paper, we propose an innovative…

Image and Video Processing · Electrical Eng. & Systems 2023-08-24 Qi Wang , Lucas Mahler , Julius Steiglechner , Florian Birk , Klaus Scheffler , Gabriele Lohmann

While dust significantly affects the environmental perception of automated agricultural machines, the existing deep learning-based methods for dust removal require further research and improvement in this area to improve the performance and…

Computer Vision and Pattern Recognition · Computer Science 2024-01-11 Shengli Zhang , Zhiyong Tao , Sen Lin

Diffeomorphic deformable multi-modal image registration is a challenging task which aims to bring images acquired by different modalities to the same coordinate space and at the same time to preserve the topology and the invertibility of…

Image and Video Processing · Electrical Eng. & Systems 2022-03-16 Vasiliki Sideri-Lampretsa , Georgios Kaissis , Daniel Rueckert

Surface roughness and texture are critical to the functional performance of engineering components. The ability to analyze roughness and texture effectively and efficiently is much needed to ensure surface quality in many surface generation…

Signal Processing · Electrical Eng. & Systems 2023-03-15 Melih C. Yesilli , Jisheng Chen , Firas A. Khasawneh , Yang Guo

Deep learning models extract, before a final classification layer, features or patterns which are key for their unprecedented advantageous performance. However, the process of complex nonlinear feature extraction is not well understood, a…

Computer Vision and Pattern Recognition · Computer Science 2020-06-18 Roozbeh Yousefzadeh , Furong Huang

In recent years, many research achievements are made in the medical image fusion field. Medical Image fusion means that several of various modality image information is comprehended together to form one image to express its information. The…

Image and Video Processing · Electrical Eng. & Systems 2020-07-23 T Deepika

The Transformer structures have been widely used in computer vision and have recently made an impact in the area of medical image registration. However, the use of Transformer in most registration networks is straightforward. These networks…

Computer Vision and Pattern Recognition · Computer Science 2024-03-26 Haiqiao Wang , Dong Ni , Yi Wang

This paper presents a lightweight framework for classifying brain stroke types from Diffusion-Weighted Imaging (DWI) MRI scans, employing a Multi-Layer Perceptron (MLP) neural network with Wavelet Transform for feature extraction. Accurate…

Image and Video Processing · Electrical Eng. & Systems 2025-06-19 Mana Mohammadi , Amirhesam Jafari Rad , Ashkan Behrouzi

Deep Learning (DL) methods have been used for electrocardiogram (ECG) processing in a wide variety of tasks, demonstrating good performance compared with traditional signal processing algorithms. These methods offer an efficient framework…

Signal Processing · Electrical Eng. & Systems 2024-07-31 Adrian Atienza , Jakob Bardram , Sadasivan Puthusserypady

Multimodal (e.g., RGB-Depth/RGB-Thermal) fusion has shown great potential for improving semantic segmentation in complex scenes (e.g., indoor/low-light conditions). Existing approaches often fully fine-tune a dual-branch encoder-decoder…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Shaohua Dong , Yunhe Feng , Qing Yang , Yan Huang , Dongfang Liu , Heng Fan

Large Language Models (LLMs) have demonstrated exceptional capabilities across diverse natural language processing benchmarks. However, the escalating scale of model parameters imposes prohibitive memory overheads during training,…

Machine Learning · Computer Science 2026-04-28 Ziqing Wen , Ping Luo , Jiahuan Wang , Kun Yuan , Dongsheng Li , Tao Sun

Objective: Currently, only behavioral speech understanding tests are available, which require active participation of the person being tested. As this is infeasible for certain populations, an objective measure of speech intelligibility is…

Audio and Speech Processing · Electrical Eng. & Systems 2021-11-30 Bernd Accou , Mohammad Jalilpour Monesi , Hugo Van hamme , Tom Francart

The analysis of gravitational-wave (GW) signals is one of the most challenging application areas of signal processing. Wavelet transforms are specially helpful in detecting and analyzing GW transients and several analysis pipelines are…

General Relativity and Quantum Cosmology · Physics 2024-05-27 Andrea Virtuoso , Edoardo Milotti

Image deblurring plays a crucial role in enhancing visual clarity across various applications. Although most deep learning approaches primarily focus on sRGB images, which inherently lose critical information during the image signal…

Image and Video Processing · Electrical Eng. & Systems 2025-09-22 Wenlong Jiao , Binglong Li , Wei Shang , Ping Wang , Dongwei Ren

Recently, learning equivariant representations has attracted considerable research attention. Dieleman et al. introduce four operations which can be inserted into convolutional neural network to learn deep representations equivariant to…

Computer Vision and Pattern Recognition · Computer Science 2018-03-01 Junying Li , Zichen Yang , Haifeng Liu , Deng Cai

Electroencephalogram (EEG) signals are often corrupted with unintended artifacts which need to be removed for extracting meaningful clinical information from them. Typically a priori knowledge of the nature of the artifacts is needed for…

Medical Physics · Physics 2018-03-02 Valentina Bono , Saptarshi Das , Wasifa Jamal , Koushik Maharatna

In this study, we demonstrate the application of a hybrid Vision Transformer (ViT) model, pretrained on ImageNet, on an electroencephalogram (EEG) regression task. Despite being originally trained for image classification tasks, when…

Computer Vision and Pattern Recognition · Computer Science 2023-08-02 Ruiqi Yang , Eric Modesitt

Existing RGB-based imitation learning approaches typically employ traditional vision encoders such as ResNet or ViT, which lack explicit 3D reasoning capabilities. Recent geometry-grounded vision models, such as VGGT~\cite{wang2025vggt},…

Robotics · Computer Science 2025-09-22 An Dinh Vuong , Minh Nhat Vu , Ian Reid

In this study, we focus on the training process and inference improvements of deep neural networks (DNNs), specifically Autoencoders (AEs) and Variational Autoencoders (VAEs), using Random Fourier Transformation (RFT). We further explore…

Machine Learning · Computer Science 2026-02-26 Ata Akbari Asanjan , Milad Memarzadeh , Bryan Matthews , Nikunj Oza

The task of recalibrating the illumination settings in an image to a target configuration is known as relighting. Relighting techniques have potential applications in digital photography, gaming industry and in augmented reality. In this…

Computer Vision and Pattern Recognition · Computer Science 2020-09-16 Densen Puthussery , Hrishikesh P. S. , Melvin Kuriakose , Jiji C.