中文
相关论文

相关论文: Optimizing Hand Region Detection in MediaPipe Holi…

200 篇论文

Inter-subject registration of cortical areas is necessary in functional imaging (fMRI) studies for making inferences about equivalent brain function across a population. However, many high-level visual brain areas are defined as peaks of…

神经元与认知 · 定量生物学 2016-06-09 Marius Cătălin Iordan , Armand Joulin , Diane M. Beck , Li Fei-Fei

Markerless tracking of hands and fingers is a promising enabler for human-computer interaction. However, adoption has been limited because of tracking inaccuracies, incomplete coverage of motions, low framerate, complex camera setups, and…

计算机视觉与模式识别 · 计算机科学 2016-02-15 Srinath Sridhar , Franziska Mueller , Antti Oulasvirta , Christian Theobalt

State-of-the-art object pose estimation methods are prone to generating geometrically infeasible pose hypotheses. This problem is prevalent in dexterous manipulation, where estimated poses often intersect with the robotic hand or are not…

机器人学 · 计算机科学 2026-03-24 Anil Zeybek , Rhys Newbury , Snehal Dikhale , Nawid Jamali , Soshi Iba , Akansel Cosgun

3D hand pose estimation from single depth image is an important and challenging problem for human-computer interaction. Recently deep convolutional networks (ConvNet) with sophisticated design have been employed to address it, but the…

计算机视觉与模式识别 · 计算机科学 2017-07-25 Hengkai Guo , Guijin Wang , Xinghao Chen , Cairong Zhang

Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. We present an offline hand-shadowing inverse-kinematics (IK) retargeting pipeline driven by a…

机器人学 · 计算机科学 2026-05-13 Hendrik Chiche , Antoine Jamme , Trevor Rigoberto Martinez , Gabriel Gomes

Thai Finger Spelling (TFS) sign recognition could benefit a community of hearing-difficulty people in bridging to a major hearing population. With a relatively large number of alphabets, TFS employs multiple signing schemes. Two schemes of…

计算机视觉与模式识别 · 计算机科学 2022-01-11 Jinnavat Sanalohit , Tatpong Katanyukul

Region anchors are the cornerstone of modern object detection techniques. State-of-the-art detectors mostly rely on a dense anchoring scheme, where anchors are sampled uniformly over the spatial domain with a predefined set of scales and…

计算机视觉与模式识别 · 计算机科学 2019-04-15 Jiaqi Wang , Kai Chen , Shuo Yang , Chen Change Loy , Dahua Lin

Hand pose estimation from a single depth image is an essential topic in computer vision and human computer interaction. Despite recent advancements in this area promoted by convolutional neural network, accurate hand pose estimation is…

计算机视觉与模式识别 · 计算机科学 2019-07-16 Xinghao Chen , Guijin Wang , Hengkai Guo , Cairong Zhang

Unlike traditional robotic hands, underactuated compliant hands are challenging to model due to inherent uncertainties. Consequently, pose estimation of a grasped object is usually performed based on visual perception. However, visual…

机器人学 · 计算机科学 2024-01-18 Osher Azulay , Inbar Ben-David , Avishai Sintov

Hand-specific localization has garnered significant interest within the computer vision community. Although there are numerous datasets with hand annotations from various angles and settings, domain transfer techniques frequently struggle…

计算机视觉与模式识别 · 计算机科学 2025-01-16 Roi Papo , Sapir Gershov , Tom Friedman , Itay Or , Gil Bolotin , Shlomi Laufer

Handwritten text recognition has been developed rapidly in the recent years, following the rise of deep learning and its applications. Though deep learning methods provide notable boost in performance concerning text recognition,…

计算机视觉与模式识别 · 计算机科学 2024-04-18 George Retsinas , Giorgos Sfikas , Basilis Gatos , Christophoros Nikou

Segmentation-based scene text detection algorithms can handle arbitrary shape scene texts and have strong robustness and adaptability, so it has attracted wide attention. Existing segmentation-based scene text detection algorithms usually…

计算机视觉与模式识别 · 计算机科学 2024-01-19 Jinzhi Zheng , Libo Zhang , Yanjun Wu , Chen Zhao

Prompt learning has gained significant attention as a parameter-efficient approach for adapting large pre-trained vision-language models to downstream tasks. However, when only partial labels are available, its performance is often limited…

计算机视觉与模式识别 · 计算机科学 2026-04-09 Yaqi Zhao , Haoliang Sun , Yating Wang , Yongshun Gong , Yilong Yin

We propose EgoGrasp, the first method to reconstruct world-space hand-object interactions (W-HOI) from dynamic egoview videos, supporting open-vocabulary objects. Accurate W-HOI reconstruction is critical for embodied intelligence yet…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Hongming Fu , Wenjia Wang , Xiaozhen Qiao , Rolandos Alexandros Potamias , Taku Komura , Shuo Yang , Zheng Liu , Bo Zhao

The dynamic hand gesture recognition task has seen studies on various unimodal and multimodal methods. Previously, researchers have explored depth and 2D-skeleton-based multimodal fusion CRNNs (Convolutional Recurrent Neural Networks) but…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Hasan Mahmud , Mashrur M. Morshed , Md. Kamrul Hasan

Accurate human motion prediction (HMP) is critical for seamless human-robot collaboration, particularly in handover tasks that require real-time adaptability. Despite the high accuracy of state-of-the-art models, their computational…

机器人学 · 计算机科学 2025-03-04 Gerard Gómez-Izquierdo , Javier Laplaza , Alberto Sanfeliu , Anaís Garrell

In this paper, we present a toolbox for a specific optimization problem that frequently arises in bioinformatics or genomics. In this specific optimisation problem, the state space is a set of words of specified length over a finite…

人工智能 · 计算机科学 2017-06-27 Régis Garnier , Christophe Guyeux , Stéphane Chrétien

Objects for detection usually have distinct characteristics in different sub-regions and different aspect ratios. However, in prevalent two-stage object detection methods, Region-of-Interest (RoI) features are extracted by RoI pooling with…

计算机视觉与模式识别 · 计算机科学 2017-11-27 Yao Zhai , Jingjing Fu , Yan Lu , Houqiang Li

Reliably planning fingertip grasps for multi-fingered hands lies as a key challenge for many tasks including tool use, insertion, and dexterous in-hand manipulation. This task becomes even more difficult when the robot lacks an accurate…

机器人学 · 计算机科学 2022-12-19 Martin Matak , Tucker Hermans

We propose to use a model-based generative loss for training hand pose estimators on depth images based on a volumetric hand model. This additional loss allows training of a hand pose estimator that accurately infers the entire set of 21…

计算机视觉与模式识别 · 计算机科学 2021-06-01 Jiayi Wang , Franziska Mueller , Florian Bernard , Christian Theobalt