English
Related papers

Related papers: GMFL-Net: A Global Multi-geometric Feature Learnin…

200 papers

As our population ages, neurological impairments and degeneration of the musculoskeletal system yield gait abnormalities, which can significantly reduce quality of life. Gait rehabilitative therapy has been widely adopted to help patients…

Computer Vision and Pattern Recognition · Computer Science 2017-02-02 Ioannis Papavasileiou , Wenlong Zhang , Xin Wang , Jinbo Bi , Li Zhang , Song Han

Multishot Magnetic Resonance Imaging (MRI) is a promising imaging modality that can produce a high-resolution image with relatively less data acquisition time. The downside of multishot MRI is that it is very sensitive to subject motion and…

Computer Vision and Pattern Recognition · Computer Science 2020-03-13 Muhammad Usman , Muhammad Umar Farooq , Siddique Latif , Muhammad Asim , Junaid Qadir

Extracting and binding salient information from different sensory modalities to determine common features in the environment is a significant challenge in robotics. Here we present MuPNet (Multi-modal Predictive Coding Network), a…

Both accuracy and efficiency are significant for pose estimation and tracking in videos. State-of-the-art performance is dominated by two-stages top-down methods. Despite the leading results, these methods are impractical for real-world…

Computer Vision and Pattern Recognition · Computer Science 2019-08-16 Jiabin Zhang , Zheng Zhu , Wei Zou , Peng Li , Yanwei Li , Hu Su , Guan Huang

We introduce a new generative approach for synthesizing 3D geometry and images from single-view collections. Most existing approaches predict volumetric density to render multi-view consistent images. By employing volumetric rendering using…

Computer Vision and Pattern Recognition · Computer Science 2024-06-17 Salvatore Esposito , Qingshan Xu , Kacper Kania , Charlie Hewitt , Octave Mariotti , Lohit Petikam , Julien Valentin , Arno Onken , Oisin Mac Aodha

Safe manipulation-oriented navigation for humanoid robots requires scene memory that remains reliable under locomotion-induced perceptual distortion, environmental changes, and interaction-level geometric safety constraints. Existing…

Robotics · Computer Science 2026-05-22 Peifeng Jiang , Hong Liu , Jin Jin , Wenshuai Wang , Xia Li

Contemporary grasp detection approaches employ deep learning to achieve robustness to sensor and object model uncertainty. The two dominant approaches design either grasp-quality scoring or anchor-based grasp recognition networks. This…

Robotics · Computer Science 2021-12-16 Ruinian Xu , Fu-Jen Chu , Patricio A. Vela

Category-agnostic pose estimation aims to locate keypoints on query images according to a few annotated support images for arbitrary novel classes. Existing methods generally extract support features via heatmap pooling, and obtain…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Junjie Chen , Weilong Chen , Yifan Zuo , Yuming Fang

We present KDFNet, a novel method for 6D object pose estimation from RGB images. To handle occlusion, many recent works have proposed to localize 2D keypoints through pixel-wise voting and solve a Perspective-n-Point (PnP) problem for pose…

Computer Vision and Pattern Recognition · Computer Science 2021-09-22 Xingyu Liu , Shun Iwase , Kris M. Kitani

Description-based person re-identification (Re-id) is an important task in video surveillance that requires discriminative cross-modal representations to distinguish different people. It is difficult to directly measure the similarity…

Computer Vision and Pattern Recognition · Computer Science 2019-06-25 Kai Niu , Yan Huang , Wanli Ouyang , Liang Wang

Learning representations of multimodal data that are both informative and robust to missing modalities at test time remains a challenging problem due to the inherent heterogeneity of data obtained from different channels. To address it, we…

Machine Learning · Computer Science 2022-11-21 Petra Poklukar , Miguel Vasco , Hang Yin , Francisco S. Melo , Ana Paiva , Danica Kragic

Detecting multiple change points in functional data sequences has been increasingly popular and critical in various scientific fields. In this article, we propose a novel two-stage framework for detecting multiple change points in…

Methodology · Statistics 2025-05-27 Zhiqing Fang , Xin Liu

Despite substantial progress in 3D human pose estimation from a single-view image, prior works rarely explore global and local correlations, leading to insufficient learning of human skeleton representations. To address this issue, we…

Computer Vision and Pattern Recognition · Computer Science 2023-04-28 Ti Wang , Hong Liu , Runwei Ding , Wenhao Li , Yingxuan You , Xia Li

Graph convolutional networks (GCNs), which can model the human body skeletons as spatial and temporal graphs, have shown remarkable potential in skeleton-based action recognition. However, in the existing GCN-based methods, graph-structured…

Computer Vision and Pattern Recognition · Computer Science 2022-10-13 Han Chen , Yifan Jiang , Hanseok Ko

Neural surface reconstruction is sensitive to the camera pose noise, even if state-of-the-art pose estimators like COLMAP or ARKit are used. More importantly, existing Pose-NeRF joint optimisation methods have struggled to improve pose…

Computer Vision and Pattern Recognition · Computer Science 2024-03-14 Jia-Wang Bian , Wenjing Bian , Victor Adrian Prisacariu , Philip Torr

We propose a novel transformer-style architecture called Global-Local Filter Network (GLFNet) for medical image segmentation and demonstrate its state-of-the-art performance. We replace the self-attention mechanism with a combination of…

Computer Vision and Pattern Recognition · Computer Science 2024-09-02 Athanasios Tragakis , Qianying Liu , Chaitanya Kaul , Swalpa Kumar Roy , Hang Dai , Fani Deligianni , Roderick Murray-Smith , Daniele Faccio

We present the first single-network approach for 2D~whole-body pose estimation, which entails simultaneous localization of body, face, hands, and feet keypoints. Due to the bottom-up formulation, our method maintains constant real-time…

Computer Vision and Pattern Recognition · Computer Science 2019-10-01 Gines Hidalgo , Yaadhav Raaj , Haroon Idrees , Donglai Xiang , Hanbyul Joo , Tomas Simon , Yaser Sheikh

An open problem in mobile manipulation is how to represent objects and scenes in a unified manner so that robots can use both for navigation and manipulation. The latter requires capturing intricate geometry while understanding fine-grained…

We describe a novel metric-based learning approach that introduces a multimodal framework and uses deep audio and geophone encoders in siamese configuration to design an adaptable and lightweight supervised model. This framework eliminates…

Sound · Computer Science 2021-11-16 Muhammad Shakeel , Katsutoshi Itoyama , Kenji Nishida , Kazuhiro Nakadai

Marker-free human pose estimation (HPE) has found increasing applications in various fields. Current HPE suffers from occasional errors in keypoint recognition and random fluctuation in keypoint trajectories when analyzing kinematic human…

Computer Vision and Pattern Recognition · Computer Science 2025-07-16 Chang Peng , Yifei Zhou , Huifeng Xi , Shiqing Huang , Chuangye Chen , Jianming Yang , Bao Yang , Zhenyu Jiang