English
Related papers

Related papers: Re-evaluating Position and Velocity Decoding for H…

200 papers

We propose Parabolic Position Encoding (PaPE), a parabola-based position encoding for vision modalities in attention-based architectures. Given a set of vision tokens-such as from videos, event camera streams, images, or point clouds-our…

Visual re-localization means using a single image as input to estimate the camera's location and orientation relative to a pre-recorded environment. The highest-scoring methods are "structure based," and need the query camera's intrinsics…

Computer Vision and Pattern Recognition · Computer Science 2021-04-13 Mehmet Ozgur Turkoglu , Eric Brachmann , Konrad Schindler , Gabriel Brostow , Aron Monszpart

Recent advances in camera-controlled video diffusion models have significantly improved video-camera alignment. However, the camera controllability still remains limited. In this work, we build upon Reward Feedback Learning and aim to…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Wenhang Ge , Guibao Shen , Jiawei Feng , Luozhou Wang , Hao Lu , Xingye Tian , Xin Tao , Ying-Cong Chen

The proliferation of XR devices has made egocentric hand pose estimation a vital task, yet this perspective is inherently challenged by frequent finger occlusions. To address this, we propose a novel approach that leverages the rich…

Computer Vision and Pattern Recognition · Computer Science 2026-01-27 William Huang , Siyou Pei , Leyi Zou , Eric J. Gonzalez , Ishan Chatterjee , Yang Zhang

Surface electromyography (sEMG) recordings can be contaminated by electrocardiogram (ECG) signals when the monitored muscle is closed to the heart. Traditional signal processing-based approaches, such as high-pass filtering and template…

Signal Processing · Electrical Eng. & Systems 2025-02-20 Yu-Tung Liu , Kuan-Chen Wang , Rong Chao , Sabato Marco Siniscalchi , Ping-Cheng Yeh , Yu Tsao

The development of EEG decoding algorithms confronts challenges such as data sparsity, subject variability, and the need for precise annotations, all of which are vital for advancing brain-computer interfaces and enhancing the diagnosis of…

Signal Processing · Electrical Eng. & Systems 2025-01-15 Zirui Wang , Zhenxi Song , Yi Guo , Yuxin Liu , Guoyang Xu , Min Zhang , Zhiguo Zhang

Positional encodings (PEs) are essential for building powerful and expressive graph neural networks and graph transformers, as they effectively capture the relative spatial relationships between nodes. Although extensive research has been…

Machine Learning · Computer Science 2026-03-16 Yinan Huang , Haoyu Wang , Pan Li

Establishing correspondences from image to 3D has been a key task of 6DoF object pose estimation for a long time. To predict pose more accurately, deeply learned dense maps replaced sparse templates. Dense methods also improved pose…

Computer Vision and Pattern Recognition · Computer Science 2022-03-31 Yongzhi Su , Mahdi Saleh , Torben Fetzer , Jason Rambach , Nassir Navab , Benjamin Busam , Didier Stricker , Federico Tombari

We propose the Encoder-Recurrent-Decoder (ERD) model for recognition and prediction of human body pose in videos and motion capture. The ERD model is a recurrent neural network that incorporates nonlinear encoder and decoder networks before…

Computer Vision and Pattern Recognition · Computer Science 2015-09-30 Katerina Fragkiadaki , Sergey Levine , Panna Felsen , Jitendra Malik

Positional Encodings (PEs) are essential for injecting structural information into Graph Neural Networks (GNNs), particularly Graph Transformers, yet their empirical impact remains insufficiently understood. We introduce a unified…

Machine Learning · Computer Science 2026-01-15 Florian Grötschla , Jiaqing Xie , Roger Wattenhofer

The aim of this work was to identify six basic movements of the hand using two systems. Being an interdisciplinary topic, there has been conducted studying in the anatomy of forearm muscles, biosignals, the method of electromyography (EMG)…

Signal Processing · Electrical Eng. & Systems 2019-06-20 Christos Sapsanis

Due to the large memory footprint of untrimmed videos, current state-of-the-art video localization methods operate atop precomputed video clip features. These features are extracted from video encoders typically trained for trimmed action…

Computer Vision and Pattern Recognition · Computer Science 2021-08-18 Humam Alwassel , Silvio Giancola , Bernard Ghanem

Surface electromyogram (sEMG) is arguably the most sought-after physiological signal with a broad spectrum of biomedical applications, especially in miniaturized rehabilitation robots such as multifunctional prostheses. The widespread use…

Category-level pose estimation is a challenging problem due to intra-class shape variations. Recent methods deform pre-computed shape priors to map the observed point cloud into the normalized object coordinate space and then retrieve the…

Computer Vision and Pattern Recognition · Computer Science 2022-08-16 Ruida Zhang , Yan Di , Fabian Manhardt , Federico Tombari , Xiangyang Ji

The rapid growth of streaming media and e-commerce has driven advancements in recommendation systems, particularly Sequential Recommendation Systems (SRS). These systems employ users' interaction histories to predict future preferences.…

Information Retrieval · Computer Science 2025-01-22 Alejo Lopez-Avila , Jinhua Du , Abbas Shimary , Ze Li

Micro-gestures are unconsciously performed body gestures that can convey the emotion states of humans and start to attract more research attention in the fields of human behavior understanding and affective computing as an emerging topic.…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Zhaoqiang Xia , Hexiang Huang , Haoyu Chen , Xiaoyi Feng , Guoying Zhao

Surface electromyography (s-EMG) sensors are a promising way to control upper-limb prostheses. However a training session is necessary in order to set up the controller that will make s-EMG based movement possible. All data recorded during…

Other Quantitative Biology · Quantitative Biology 2015-11-25 Marco Lampacrescia

3D reconstruction serves as the foundational layer for numerous robotic perception tasks, including 6D object pose estimation and grasp pose generation. Modern 3D reconstruction methods for objects can produce visually and geometrically…

Robotics · Computer Science 2026-02-20 Varun Burde , Pavel Burget , Torsten Sattler

Recent state-of-the-art Learned Image Compression methods feature spatial context models, achieving great rate-distortion improvements over hyperprior methods. However, the autoregressive context model requires serial decoding, limiting…

Computer Vision and Pattern Recognition · Computer Science 2023-02-21 Fangzheng Lin , Heming Sun , Jinming Liu , Jiro Katto

Electroencephalogram (EEG) signals have become a popular medium for decoding visual information due to their cost-effectiveness and high temporal resolution. However, current approaches face significant challenges in bridging the modality…

Machine Learning · Computer Science 2026-03-10 Sicheng Dai , Hongwang Xiao , Shan Yu , Qiwei Ye