English
Related papers

Related papers: Semantic-Aware Transformation-Invariant RoI Align

200 papers

Zero-Shot Learning (ZSL) is achieved via aligning the semantic relationships between the global image feature vector and the corresponding class semantic descriptions. However, using the global features to represent fine-grained images may…

Computer Vision and Pattern Recognition · Computer Science 2018-05-22 Yunlong Yu , Zhong Ji , Yanwei Fu , Jichang Guo , Yanwei Pang , Zhongfei Zhang

Up to the present, an enormous number of advanced techniques have been developed to enhance and extract the spatially semantic information in hyperspectral image processing and analysis. However, locally semantic change, such as scene…

Computer Vision and Pattern Recognition · Computer Science 2019-12-20 Danfeng Hong , Xin Wu , Pedram Ghamisi , Jocelyn Chanussot , Naoto Yokoya , Xiao Xiang Zhu

Enabling object detectors to recognize out-of-distribution (OOD) objects is vital for building reliable systems. A primary obstacle stems from the fact that models frequently do not receive supervisory signals from unfamiliar data, leading…

Computer Vision and Pattern Recognition · Computer Science 2025-03-31 Bin Zhang , Jinggang Chen , Xiaoyang Qu , Guokuan Li , Kai Lu , Jiguang Wan , Jing Xiao , Jianzong Wang

Few-shot semantic segmentation (FSS) aims to segment objects of unseen classes in query images with only a few annotated support images. Existing FSS algorithms typically focus on mining category representations from the single-view support…

Computer Vision and Pattern Recognition · Computer Science 2023-09-29 Qinglong Cao , Yuntian Chen , Chao Ma , Xiaokang Yang

Sign language recognition (SLR) refers to interpreting sign language glosses from given videos automatically. This research area presents a complex challenge in computer vision because of the rapid and intricate movements inherent in sign…

Computer Vision and Pattern Recognition · Computer Science 2025-03-27 Muxin Pu , Mei Kuan Lim , Chun Yong Chong

Semantic segmentation is a fundamental task in multimedia processing, which can be used for analyzing, understanding, editing contents of images and videos, among others. To accelerate the analysis of multimedia data, existing segmentation…

Computer Vision and Pattern Recognition · Computer Science 2024-12-13 Zhiyan Wang , Deyin Liu , Lin Yuanbo Wu , Song Wang , Xin Guo , Lin Qi

Skeleton-aware sign language recognition (SLR) has gained popularity due to its ability to remain unaffected by background information and its lower computational requirements. Current methods utilize spatial graph modules and temporal…

Computer Vision and Pattern Recognition · Computer Science 2024-03-20 Lianyu Hu , Liqing Gao , Zekang Liu , Wei Feng

Radiance Fields (RF) are popular to represent casually-captured scenes for new view synthesis and several applications beyond it. Mixed reality on personal spaces needs understanding and manipulating scenes represented as RFs, with semantic…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Rahul Goel , Dhawal Sirikonda , Saurabh Saini , PJ Narayanan

Automatic radiology report generation has attracted enormous research interest due to its practical value in reducing the workload of radiologists. However, simultaneously establishing global correspondences between the image (e.g., Chest…

Computer Vision and Pattern Recognition · Computer Science 2023-07-19 Yaowei Li , Bang Yang , Xuxin Cheng , Zhihong Zhu , Hongxiang Li , Yuexian Zou

Iris segmentation is a deterministic part of the iris recognition system. Unreliable segmentation of iris regions especially the limbic area is still the bottleneck problem, which impedes more accurate recognition. To make further efforts…

Computer Vision and Pattern Recognition · Computer Science 2021-11-19 Jianze Wei , Huaibo Huang , Muyi Sun , Yunlong Wang , Min Ren , Ran He , Zhenan Sun

High-resolution remote sensing images (HRRSIs) contain substantial ground object information, such as texture, shape, and spatial location. Semantic segmentation, which is an important task for element extraction, has been widely used in…

Computer Vision and Pattern Recognition · Computer Science 2020-05-08 Haifeng Li , Kaijian Qiu , Li Chen , Xiaoming Mei , Liang Hong , Chao Tao

Object extraction and segmentation from remote sensing (RS) images is a critical yet challenging task in urban environment monitoring. Urban morphology is inherently complex, with irregular objects of diverse shapes and varying scales.…

Computer Vision and Pattern Recognition · Computer Science 2025-02-24 Chenyu Li , Danfeng Hong , Bing Zhang , Yuxuan Li , Gustau Camps-Valls , Xiao Xiang Zhu , Jocelyn Chanussot

Existing real-world super-resolution (RSR) methods based on generative priors have achieved remarkable progress in producing high-quality and globally consistent reconstructions. However, they often struggle to recover fine-grained details…

Computer Vision and Pattern Recognition · Computer Science 2026-03-26 Zixin Guo , Kai Zhao , Luyan Zhang

Accurate localization and mapping in outdoor environments remains challenging when using consumer-grade hardware, particularly with rolling-shutter cameras and low-precision inertial navigation systems (INS). We present a novel semantic…

Robotics · Computer Science 2025-04-04 Yuchen Zhang , Miao Fan , Shengtong Xu , Xiangzeng Liu , Haoyi Xiong

Recently salient object detection has witnessed remarkable improvement owing to the deep convolutional neural networks which can harvest powerful features for images. In particular, state-of-the-art salient object detection methods enjoy…

Computer Vision and Pattern Recognition · Computer Science 2019-05-10 Haofeng Li , Guanbin Li , Yizhou Yu

The past decade has witnessed significant progress on detecting objects in aerial images that are often distributed with large scale variations and arbitrary orientations. However most of existing methods rely on heuristically defined…

Computer Vision and Pattern Recognition · Computer Science 2021-07-13 Jiaming Han , Jian Ding , Jie Li , Gui-Song Xia

Researchers in functional neuroimaging mostly use activation coordinates to formulate their hypotheses. Instead, we propose to use the full statistical images to define regions of interest (ROIs). This paper presents two machine learning…

Machine Learning · Statistics 2012-09-10 Yannick Schwartz , Gaël Varoquaux , Bertrand Thirion

Recent advances in rotation-invariant (RI) learning for 3D point clouds typically replace raw coordinates with handcrafted RI features to ensure robustness under arbitrary rotations. However, these approaches often suffer from the loss of…

Computer Vision and Pattern Recognition · Computer Science 2025-11-13 Jiaxun Guo , Manar Amayri , Nizar Bouguila , Xin Liu , Wentao Fan

To better detect pedestrians of various scales, deep multi-scale methods usually detect pedestrians of different scales by different in-network layers. However, the semantic levels of features from different layers are usually inconsistent.…

Computer Vision and Pattern Recognition · Computer Science 2018-04-04 Jiale Cao , Yanwei Pang , Xuelong Li

ROI extraction is an active but challenging task in remote sensing because of the complicated landform, the complex boundaries and the requirement of annotations. Weakly supervised learning (WSL) aims at learning a mapping from input image…

Computer Vision and Pattern Recognition · Computer Science 2023-05-11 Lingfeng He , Mengze Xu , Jie Ma