English
Related papers

Related papers: HSTFormer: Hierarchical Spatial-Temporal Transform…

200 papers

Movement synchrony reflects the coordination of body movements between interacting dyads. The estimation of movement synchrony has been automated by powerful deep learning models such as transformer networks. However, instead of designing a…

Computer Vision and Pattern Recognition · Computer Science 2022-08-03 Jicheng Li , Anjana Bhat , Roghayeh Barmaki

In this paper we present a novel method to estimate 3D human pose and shape from monocular videos. This task requires directly recovering pixel-alignment 3D human pose and body shape from monocular images or videos, which is challenging due…

Computer Vision and Pattern Recognition · Computer Science 2023-03-02 Sen Yang , Wen Heng , Gang Liu , Guozhong Luo , Wankou Yang , Gang Yu

We propose a novel Transformer-based architecture for the task of generative modelling of 3D human motion. Previous work commonly relies on RNN-based models considering shorter forecast horizons reaching a stationary and often implausible…

Computer Vision and Pattern Recognition · Computer Science 2021-11-30 Emre Aksan , Manuel Kaufmann , Peng Cao , Otmar Hilliges

3D Human Pose Estimation (HPE) is the task of locating keypoints of the human body in 3D space from 2D or 3D representations such as RGB images, depth maps or point clouds. Current HPE methods from depth and point clouds predominantly rely…

Computer Vision and Pattern Recognition · Computer Science 2024-09-04 Irene Ballester , Ondřej Peterka , Martin Kampel

Multi-person pose forecasting remains a challenging problem, especially in modeling fine-grained human body interaction in complex crowd scenarios. Existing methods typically represent the whole pose sequence as a temporal series, yet…

Computer Vision and Pattern Recognition · Computer Science 2023-03-14 Xiaogang Peng , Siyuan Mao , Zizhao Wu

Multivariate time series analysis has long been one of the key research topics in the field of artificial intelligence. However, analyzing complex time series data remains a challenging and unresolved problem due to its high dimensionality,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Hao Si , Xiao Wang , Fan Zhang , Xiaoya Zhou , Dengdi Sun , Wanli Lyu , Qingquan Yang , Jin Tang

Registering point clouds of dressed humans to parametric human models is a challenging task in computer vision. Traditional approaches often rely on heavily engineered pipelines that require accurate manual initialization of human poses and…

Computer Vision and Pattern Recognition · Computer Science 2021-04-19 Shaofei Wang , Andreas Geiger , Siyu Tang

The state-of-the-art for monocular 3D human pose estimation in videos is dominated by the paradigm of 2D-to-3D pose uplifting. While the uplifting methods themselves are rather efficient, the true computational complexity depends on the…

Computer Vision and Pattern Recognition · Computer Science 2022-10-24 Moritz Einfalt , Katja Ludwig , Rainer Lienhart

Neural preconditioners for real-time physics simulation offer promising data-driven priors, but they often fail to capture long-range couplings efficiently because they inherit local message passing or sparse-operator access patterns. We…

Graphics · Computer Science 2026-05-15 Carl Osborne , Minghao Guo , Crystal Owens , Wojciech Matusik

Crucial to the success of training a depth-based 3D hand pose estimator (HPE) is the availability of comprehensive datasets covering diverse camera perspectives, shapes, and pose variations. However, collecting such annotated datasets is…

Computer Vision and Pattern Recognition · Computer Science 2018-05-14 Seungryul Baek , Kwang In Kim , Tae-Kyun Kim

Multi-person pose understanding from RGB videos involves three complex tasks: pose estimation, tracking and motion forecasting. Intuitively, accurate multi-person pose estimation facilitates robust tracking, and robust tracking builds…

Computer Vision and Pattern Recognition · Computer Science 2023-09-14 Shihao Zou , Yuanlu Xu , Chao Li , Lingni Ma , Li Cheng , Minh Vo

Recent approaches for monocular 3D human pose estimation (3D HPE) have achieved leading performance by directly regressing 3D poses from 2D keypoint sequences. Despite the rapid progress in 3D HPE, existing methods are typically trained and…

Computer Vision and Pattern Recognition · Computer Science 2025-12-17 Qingyuan Cai , Linxin Zhang , Xuecai Hu , Saihui Hou , Yongzhen Huang

Recent years have witnessed a trend of applying context frames to boost the performance of object detection as video object detection. Existing methods usually aggregate features at one stroke to enhance the feature. These methods, however,…

Computer Vision and Pattern Recognition · Computer Science 2022-09-07 Han Wang , Jun Tang , Xiaodong Liu , Shanyan Guan , Rong Xie , Li Song

Significant advancements have been achieved in the realm of understanding poses and interactions of two hands manipulating an object. The emergence of augmented reality (AR) and virtual reality (VR) technologies has heightened the demand…

Computer Vision and Pattern Recognition · Computer Science 2025-02-28 Elkhan Ismayilzada , MD Khalequzzaman Chowdhury Sayem , Yihalem Yimolal Tiruneh , Mubarrat Tajoar Chowdhury , Muhammadjon Boboev , Seungryul Baek

Convolutional neural networks (CNNs) have been the consensus for medical image segmentation tasks. However, they suffer from the limitation in modeling long-range dependencies and spatial correlations due to the nature of convolution…

Computer Vision and Pattern Recognition · Computer Science 2023-01-10 Moein Heidari , Amirhossein Kazerouni , Milad Soltany , Reza Azad , Ehsan Khodapanah Aghdam , Julien Cohen-Adad , Dorit Merhof

Transformers have significantly advanced the field of 3D human pose estimation (HPE). However, existing transformer-based methods primarily use self-attention mechanisms for spatio-temporal modeling, leading to a quadratic complexity,…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Yunlong Huang , Junshuo Liu , Ke Xian , Robert Caiming Qiu

Recently, deep learning-based tooth segmentation methods have been limited by the expensive and time-consuming processes of data collection and labeling. Achieving high-precision segmentation with limited datasets is critical. A viable…

Computer Vision and Pattern Recognition · Computer Science 2023-10-24 Yuan Li , Huan Liu , Yubo Tao , Xiangyang He , Haifeng Li , Xiaohu Guo , Hai Lin

Vision Transformers (ViTs) have recently achieved state-of-the-art performance in 2D human pose estimation due to their strong global modeling capability. However, existing ViT-based pose estimators are designed for static images and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-09 Hongwei Fang , Jiahang Cai , Xun Wang , Wenwu Yang

Dynamic representation learning plays a pivotal role in understanding the evolution of linguistic content over time. On this front both context and time dynamics as well as their interplay are of prime importance. Current approaches model…

Computation and Language · Computer Science 2024-10-23 Talia Tseriotou , Adam Tsakalidis , Maria Liakata

Effective learning of spatial-temporal information within a point cloud sequence is highly important for many down-stream tasks such as 4D semantic segmentation and 3D action recognition. In this paper, we propose a novel framework named…

Computer Vision and Pattern Recognition · Computer Science 2021-10-20 Yimin Wei , Hao Liu , Tingting Xie , Qiuhong Ke , Yulan Guo
‹ Prev 1 4 5 6 7 8 10 Next ›