中文
相关论文

相关论文: Two-stream Fusion Model for Dynamic Hand Gesture R…

200 篇论文

Perceiving and understanding 3D motion is a core technology in fields such as autonomous driving, robots, and motion prediction. This paper proposes a 3D motion perception method called ScaleFlow++ that is easy to generalize. With just a…

计算机视觉与模式识别 · 计算机科学 2024-10-15 Han Ling , Yinghui Sun , Quansen Sun , Yuhui Zheng

Perceiving and understanding 3D motion is a core technology in fields such as autonomous driving, robots, and motion prediction. This paper proposes a 3D motion perception method called ScaleFlow++ that is easy to generalize. With just a…

计算机视觉与模式识别 · 计算机科学 2024-10-17 Han Ling , Quansen Sun

We are concerned with a novel sensor-based gesture input/instruction technology which enables human beings to interact with computers conveniently. The human being wears an emitter on the finger or holds a digital pen that generates a time…

经典物理 · 物理学 2017-05-23 Yukun Guo , Jingzhi Li , Hongyu Liu , Xianchao Wang

In this paper, we propose a deep learning approach for smartphone user identification based on analyzing motion signals recorded by the accelerometer and the gyroscope, during a single tap gesture performed by the user on the screen. We…

机器学习 · 计算机科学 2020-03-24 Cezara Benegui , Radu Tudor Ionescu

This paper proposes the second version of the widespread Hand Gesture Recognition dataset HaGRID -- HaGRIDv2. We cover 15 new gestures with conversation and control functions, including two-handed ones. Building on the foundational concepts…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Anton Nuzhdin , Alexander Nagaev , Alexander Sautin , Alexander Kapitanov , Karina Kvanchiani

Scene flow represents the 3D motion of every point in the dynamic environments. Like the optical flow that represents the motion of pixels in 2D images, 3D motion representation of scene flow benefits many applications, such as autonomous…

计算机视觉与模式识别 · 计算机科学 2021-06-09 Guangming Wang , Xinrui Wu , Zhe Liu , Hesheng Wang

We present a 3D Convolutional Neural Networks (CNNs) based single shot detector for spatial-temporal action detection tasks. Our model includes: (1) two short-term appearance and motion streams, with single RGB and optical flow image input…

计算机视觉与模式识别 · 计算机科学 2019-08-23 Pengfei Zhang , Yu Cao , Benyuan Liu

Hand detection is essential for many hand related tasks, e.g. parsing hand pose, understanding gesture, which are extremely useful for robotics and human-computer interaction. However, hand detection in uncontrolled environments is…

计算机视觉与模式识别 · 计算机科学 2016-12-09 Xiaoming Deng , Ye Yuan , Yinda Zhang , Ping Tan , Liang Chang , Shuo Yang , Hongan Wang

Occlusions between consecutive frames have long posed a significant challenge in optical flow estimation. The inherent ambiguity introduced by occlusions directly violates the brightness constancy constraint and considerably hinders…

计算机视觉与模式识别 · 计算机科学 2023-11-30 Shangkun Sun , Jiaming Liu , Thomas H. Li , Huaxia Li , Guoqing Liu , Wei Gao

This paper presents an evaluation of deep neural networks for recognition of digits entered by users on a smartphone touchscreen. A new large dataset of Arabic numerals was collected for training and evaluation of the network. The dataset…

计算机视觉与模式识别 · 计算机科学 2017-09-21 Philip J. Corr , Guenole C. Silvestre , Chris J. Bleakley

Recently, medical image synthesis gains more and more popularity, along with the rapid development of generative models. Medical image synthesis aims to generate an unacquired image modality, often from other observed data modalities.…

图像与视频处理 · 电气工程与系统科学 2025-07-04 Zhe Xiong , Qiaoqiao Ding , Xiaoqun Zhang

We present a dual-pathway approach for recognizing fine-grained interactions from videos. We build on the success of prior dual-stream approaches, but make a distinction between the static and dynamic representations of objects and their…

计算机视觉与模式识别 · 计算机科学 2021-04-02 Tae Soo Kim , Jonathan Jones , Gregory D. Hager

In this paper, we revive the use of old-fashioned handcrafted video representations for action recognition and put new life into these techniques via a CNN-based hallucination step. Despite of the use of RGB and optical flow frames, the I3D…

计算机视觉与模式识别 · 计算机科学 2019-08-20 Lei Wang , Piotr Koniusz , Du Q. Huynh

With the rapid advancements in deep learning, computer vision tasks have seen significant improvements, making two-stream neural networks a popular focus for video based action recognition. Traditional models using RGB and optical flow…

计算机视觉与模式识别 · 计算机科学 2024-11-28 Song-Jiang Lai , Tsun-Hin Cheung , Ka-Chun Fung , Tian-Shan Liu , Kin-Man Lam

Human action recognition is one of the challenging tasks in computer vision. The current action recognition methods use computationally expensive models for learning spatio-temporal dependencies of the action. Models utilizing RGB channels…

计算机视觉与模式识别 · 计算机科学 2022-06-07 Labina Shrestha , Shikha Dubey , Farrukh Olimov , Muhammad Aasim Rafique , Moongu Jeon

3D hand pose estimation from a single depth image plays an important role in computer vision and human-computer interaction. Although recent hand pose estimation methods using convolution neural network (CNN) have shown notable improvements…

计算机视觉与模式识别 · 计算机科学 2020-08-28 Cheol-hwan Yoo , Seo-won Ji , Yong-goo Shin , Seung-wook Kim , Sung-jea Ko

Improved dense trajectories (iDT) have shown great performance in action recognition, and their combination with the two-stream approach has achieved state-of-the-art performance. It is, however, difficult for iDT to completely remove…

计算机视觉与模式识别 · 计算机科学 2016-05-02 Katsunori Ohnishi , Masatoshi Hidaka , Tatsuya Harada

Object pose tracking is one of the pivotal technologies in multimedia, attracting ever-growing attention in recent years. Existing methods employing traditional cameras encounter numerous challenges such as motion blur, sensor noise,…

计算机视觉与模式识别 · 计算机科学 2025-12-25 Zibin Liu , Banglei Guan , Yang Shang , Shunkun Liang , Zhenbao Yu , Qifeng Yu

Dual-arm cooperative manipulation holds great promise for tackling complex real-world tasks that demand seamless coordination and adaptive dynamics. Despite substantial progress in learning-based motion planning, most approaches struggle to…

机器人学 · 计算机科学 2025-11-24 Jiaming Chen , Yiyu Jiang , Aoshen Huang , Yang Li , Wei Pan

Existing event stream-based pattern recognition models usually represent the event stream as the point cloud, voxel, image, etc., and design various deep neural networks to learn their features. Although considerable results can be achieved…

计算机视觉与模式识别 · 计算机科学 2024-06-28 Lan Chen , Dong Li , Xiao Wang , Pengpeng Shao , Wei Zhang , Yaowei Wang , Yonghong Tian , Jin Tang