English
Related papers

Related papers: LDSF: Lightweight Dual-Stream Framework for SAR Ta…

200 papers

This study introduces LRDif, a novel diffusion-based framework designed specifically for facial expression recognition (FER) within the context of under-display cameras (UDC). To address the inherent challenges posed by UDC's image…

Computer Vision and Pattern Recognition · Computer Science 2024-02-02 Zhifeng Wang , Kaihao Zhang , Ramesh Sankaranarayana

Nonlinear electromagnetic (EM) inverse scattering is a quantitative and super-resolution imaging technique, in which more realistic interactions between the internal structure of scene and EM wavefield are taken into account in the imaging…

Information Retrieval · Computer Science 2019-05-01 Lianlin Li , Long Gang Wang , Fernando L. Teixeira , Che Liu , Arye Nehora , Tie Jun Cui

Ultrasound imaging is a prevalent diagnostic tool known for its simplicity and non-invasiveness. However, its inherent characteristics often introduce substantial noise, posing considerable challenges for automated lesion or organ…

Computer Vision and Pattern Recognition · Computer Science 2025-07-11 Ling Zhou , Runtian Yuan , Yi Liu , Yuejie Zhang , Rui Feng , Shang Gao

Laryngo-pharyngeal cancer (LPC) is a highly fatal malignant disease affecting the head and neck region. Previous studies on endoscopic tumor detection, particularly those leveraging dual-branch network architectures, have shown significant…

Computer Vision and Pattern Recognition · Computer Science 2024-08-16 Jia Wei , Yun Li , Meiyu Qiu , Hongyu Chen , Xiaomao Fan , Wenbin Lei

Lensless imaging stands out as a promising alternative to conventional lens-based systems, particularly in scenarios demanding ultracompact form factors and cost-effective architectures. However, such systems are fundamentally governed by…

Image and Video Processing · Electrical Eng. & Systems 2025-05-06 Jiesong Bai , Yuhao Yin , Yihang Dong , Xiaofeng Zhang , Chi-Man Pun , Xuhang Chen

We focus on the word-level visual lipreading, which requires recognizing the word being spoken, given only the video but not the audio. State-of-the-art methods explore the use of end-to-end neural networks, including a shallow (up to three…

Computer Vision and Pattern Recognition · Computer Science 2019-07-22 Xinshuo Weng , Kris Kitani

Microphone array techniques are widely used in sound source localization and smart city acoustic-based traffic monitoring, but these applications face significant challenges due to the scarcity of labeled real-world traffic audio data and…

Audio and Speech Processing · Electrical Eng. & Systems 2024-12-30 Shitong Fan , Feiyang Xiao , Wenbo Wang , Shuhan Qi , Qiaoxi Zhu , Wenwu Wang , Jian Guan

Recent advances of semantic image segmentation greatly benefit from deeper and larger Convolutional Neural Network (CNN) models. Compared to image segmentation in the wild, properties of both medical images themselves and of existing…

Computer Vision and Pattern Recognition · Computer Science 2022-07-28 Xin Chen , Ke Ding

LiDAR and camera fusion techniques are promising for achieving 3D object detection in autonomous driving. Most multi-modal 3D object detection frameworks integrate semantic knowledge from 2D images into 3D LiDAR point clouds to enhance…

Computer Vision and Pattern Recognition · Computer Science 2023-06-21 Shaoqing Xu , Fang Li , Ziying Song , Jin Fang , Sifen Wang , Zhi-Xin Yang

Local feature extraction is a standard approach in computer vision for tackling important tasks such as image matching and retrieval. The core assumption of most methods is that images undergo affine transformations, disregarding more…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Guilherme Potje , Felipe Cadar , Andre Araujo , Renato Martins , Erickson R. Nascimento

Action segmentation is a challenging yet active research area that involves identifying when and where specific actions occur in continuous video streams. Most existing work has focused on single-stream approaches that model the…

Computer Vision and Pattern Recognition · Computer Science 2025-10-10 Harshala Gammulle , Clinton Fookes , Sridha Sridharan , Simon Denman

Deep learning has driven significant progress in object detection using Synthetic Aperture Radar (SAR) imagery. Existing methods, while achieving promising results, often struggle to effectively integrate local and global information,…

Computer Vision and Pattern Recognition · Computer Science 2024-10-15 Mingxiang Cao , Weiying Xie , Jie Lei , Jiaqing Zhang , Daixun Li , Yunsong Li

Deep learning and Convolutional Neural Networks (CNNs) have driven major transformations in diverse research areas. However, their limitations in handling low-frequency information present obstacles in certain tasks like interpreting global…

Computer Vision and Pattern Recognition · Computer Science 2024-03-14 Fuzhi Wu , Jiasong Wu , Youyong Kong , Chunfeng Yang , Guanyu Yang , Huazhong Shu , Guy Carrault , Lotfi Senhadji

We propose a fast and generalizable solution to Multi-view Photometric Stereo (MVPS), called MVPSNet. The key to our approach is a feature extraction network that effectively combines images from the same view captured under multiple…

Computer Vision and Pattern Recognition · Computer Science 2023-05-19 Dongxu Zhao , Daniel Lichy , Pierre-Nicolas Perrin , Jan-Michael Frahm , Soumyadip Sengupta

The multichannel electrode array used for electromyogram (EMG) pattern recognition provides good performance, but it has a high cost, is computationally expensive, and is inconvenient to wear. Therefore, researchers try to use as few…

Signal Processing · Electrical Eng. & Systems 2022-05-24 Md. Johirul Islam , Shamim Ahmad , Fahmida Haque , Mamun Bin Ibne Reaz , Mohammad A. S. Bhuiyan , Md. Rezaul Islam

Face Anti-Spoofing (FAS) is essential to secure face recognition systems and has been extensively studied in recent years. Although deep neural networks (DNNs) for the FAS task have achieved promising results in intra-dataset experiments…

Computer Vision and Pattern Recognition · Computer Science 2022-05-18 Rizhao Cai , Zhi Li , Renjie Wan , Haoliang Li , Yongjian Hu , Alex Chichung Kot

Automatically detecting violence from surveillance footage is a subset of activity recognition that deserves special attention because of its wide applicability in unmanned security monitoring systems, internet video filtration, etc. In…

Computer Vision and Pattern Recognition · Computer Science 2021-10-08 Zahidul Islam , Mohammad Rukonuzzaman , Raiyan Ahmed , Md. Hasanul Kabir , Moshiur Farazi

Stereo image super-resolution (SR) refers to the reconstruction of a high-resolution (HR) image from a pair of low-resolution (LR) images as typically captured by a dual-camera device. To enhance the quality of SR images, most previous…

Image and Video Processing · Electrical Eng. & Systems 2024-05-15 Yihong Chen , Zhen Fan , Shuai Dong , Zhiwei Chen , Wenjie Li , Minghui Qin , Min Zeng , Xubing Lu , Guofu Zhou , Xingsen Gao , Jun-Ming Liu

We study multi-sensor fusion for 3D semantic segmentation that is important to scene understanding for many applications, such as autonomous driving and robotics. Existing fusion-based methods, however, may not achieve promising performance…

Computer Vision and Pattern Recognition · Computer Science 2024-09-10 Mingkui Tan , Zhuangwei Zhuang , Sitao Chen , Rong Li , Kui Jia , Qicheng Wang , Yuanqing Li

Existing deep learning-based methods can capture shared features from optical and synthetic aperture radar (SAR) images for spatial alignment. However, optical-SAR registration remains challenging under large geometric deformations, because…

Computer Vision and Pattern Recognition · Computer Science 2026-05-13 Zhuoyu Cai , Dou Quan , Ning Huyan , Pei He , Shuang Wang , Licheng Jiao
‹ Prev 1 3 4 5 6 7 10 Next ›