中文
相关论文

相关论文: Action Unit Detection with Region Adaptation, Mult…

200 篇论文

High-resolution remote sensing imagery increasingly contains dense clusters of tiny objects, the detection of which is extremely challenging due to severe mutual occlusion and limited pixel footprints. Existing detection methods typically…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Zhicheng Zhao , Xuanang Fan , Lingma Sun , Chenglong Li , Jin Tang

Thoracic aortic dissection and aneurysms are the most lethal diseases of the aorta. The major hindrance to treatment lies in the accurate analysis of the medical images. More particularly, aortic segmentation of the 3D image is often…

图像与视频处理 · 电气工程与系统科学 2026-01-14 Loris Giordano , Ine Dirks , Tom Lenaerts , Jef Vandemeulebroucke

Breast cancer screening with mammography remains central to early detection and mortality reduction. Deep learning has shown strong potential for automating mammogram interpretation, yet limited-resolution datasets and small sample sizes…

计算机视觉与模式识别 · 计算机科学 2025-10-07 Farbod Bigdeli , Mohsen Mohammadagha , Ali Bigdeli

Multi-label multi-view action recognition aims to recognize multiple concurrent or sequential actions from untrimmed videos captured by multiple cameras. Existing work has focused on multi-view action recognition in a narrow area with…

计算机视觉与模式识别 · 计算机科学 2024-10-21 Trung Thanh Nguyen , Yasutomo Kawanishi , Takahiro Komamizu , Ichiro Ide

Accurate motion tracking of snow particles in avalanche events requires robust localization in global navigation satellite system (GNSS)-denied outdoor environments. This paper introduces AoI-FusionNet, a tightly coupled deep learning-based…

信号处理 · 电气工程与系统科学 2026-03-16 Tehmina Bibi , Anselm Köhler , Jan-Thomas Fischer , Falko Dressler

Extensive efforts have been devoted to recognizing facial action units (AUs). However, it is still challenging to recognize AUs from spontaneous facial displays especially when they are accompanied with speech. Different from all prior work…

计算机视觉与模式识别 · 计算机科学 2017-09-20 Zibo Meng , Shizhong Han , Yan Tong

Localizing anatomical landmarks are important tasks in medical image analysis. However, the landmarks to be localized often lack prominent visual features. Their locations are elusive and easily confused with the background, and thus…

图像与视频处理 · 电气工程与系统科学 2022-12-23 Xiaofeng Lei , Shaohua Li , Xinxing Xu , Huazhu Fu , Yong Liu , Yih-Chung Tham , Yangqin Feng , Mingrui Tan , Yanyu Xu , Jocelyn Hui Lin Goh , Rick Siow Mong Goh , Ching-Yu Cheng

Human behavior expression and experience are inherently multi-modal, and characterized by vast individual and contextual heterogeneity. To achieve meaningful human-computer and human-robot interactions, multi-modal models of the users…

机器学习 · 计算机科学 2019-06-10 Ognjen Rudovic , Meiru Zhang , Bjorn Schuller , Rosalind W. Picard

The classification of airborne laser scanning (ALS) point clouds is a critical task of remote sensing and photogrammetry fields. Although recent deep learning-based methods have achieved satisfactory performance, they have ignored the…

计算机视觉与模式识别 · 计算机科学 2022-07-22 Yongqiang Mao , Kaiqiang Chen , Wenhui Diao , Xian Sun , Xiaonan Lu , Kun Fu , Martin Weinmann

Facial action unit recognition is an important task for facial analysis. Owing to the complex collection environment, facial action unit recognition in the wild is still challenging. The 3rd competition on affective behavior analysis…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Shangfei Wang , Yanan Chang , Jiahe Wang

Expressions and facial action units (AUs) are two levels of facial behavior descriptors. Expression auxiliary information has been widely used to improve the AU detection performance. However, most existing expression representations can…

计算机视觉与模式识别 · 计算机科学 2022-10-31 Rudong An , Wei Zhang , Hao Zeng , Wei Chen , Zhigang Deng , Yu Ding

Face parsing computes pixel-wise label maps for different semantic components (e.g., hair, mouth, eyes) from face images. Existing face parsing literature have illustrated significant advantages by focusing on individual regions of interest…

计算机视觉与模式识别 · 计算机科学 2019-06-05 Jinpeng Lin , Hao Yang , Dong Chen , Ming Zeng , Fang Wen , Lu Yuan

In this paper, we aim to tackle the task of semi-supervised video object segmentation across a sequence of frames where only the ground-truth segmentation of the first frame is provided. The challenges lie in how to online update the…

计算机视觉与模式识别 · 计算机科学 2019-09-30 Mingjie Sun , Jimin Xiao , Eng Gee Lim , Yanchu Xie , Jiashi Feng

Handling varying computational resources is a critical issue in modern AI applications. Adaptive deep networks, featuring the dynamic employment of multiple classifier heads among different layers, have been proposed to address…

计算机视觉与模式识别 · 计算机科学 2024-08-30 Xu Zhang , Zhipeng Xie , Haiyang Yu , Qitong Wang , Peng Wang , Wei Wang

Remote sensing change detection between bi-temporal images receives growing concentration from researchers. However, comparing two bi-temporal images for detecting changes is challenging, as they demonstrate different appearances. In this…

计算机视觉与模式识别 · 计算机科学 2023-10-04 Luyi Qiu , Xiaofeng Zhang , ChaoChen Gu , and ShanYing Zhu

Heterogeneous Face Recognition (HFR) aims to expand the applicability of Face Recognition (FR) systems to challenging scenarios, enabling the matching of face images across different domains, such as matching thermal images to visible…

计算机视觉与模式识别 · 计算机科学 2024-04-23 Anjith George , Sebastien Marcel

The most widely used activation functions in current deep feed-forward neural networks are rectified linear units (ReLU), and many alternatives have been successfully applied, as well. However, none of the alternatives have managed to…

机器学习 · 计算机科学 2018-06-27 Leon René Sütfeld , Flemming Brieger , Holger Finger , Sonja Füllhase , Gordon Pipa

Recently, the joint learning framework (JOINT) integrates matching based transductive reasoning and online inductive learning to achieve accurate and robust semi-supervised video object segmentation (SVOS). However, using the mask embedding…

计算机视觉与模式识别 · 计算机科学 2022-12-06 Meng Lan , Jing Zhang , Lefei Zhang , Dacheng Tao

Magnetic Resonance Imaging (MRI) is an essential diagnostic tool for assessing knee injuries. However, manual interpretation of MRI slices remains time-consuming and prone to inter-observer variability. This study presents a systematic…

图像与视频处理 · 电气工程与系统科学 2025-08-22 Justin Yiu , Kushank Arora , Daniel Steinberg , Rohit Ghiya

Typical detection-free methods for image-to-point cloud registration leverage transformer-based architectures to aggregate cross-modal features and establish correspondences. However, they often struggle under challenging conditions, where…

计算机视觉与模式识别 · 计算机科学 2025-11-11 Zhixin Cheng , Xiaotian Yin , Jiacheng Deng , Bohao Liao , Yujia Chen , Xu Zhou , Baoqun Yin , Tianzhu Zhang