中文
相关论文

相关论文: Real-Time Glottis Detection Framework via Spatial-…

200 篇论文

We present a generalized velocity model to improve localization when using an Inertial Navigation System (INS). This algorithm was applied to correct the velocity of a smart phone based indoor INS system to increase the accuracy by…

其他计算机科学 · 计算机科学 2016-01-13 Rasika Lakmal Hettiarachchige Don , Jagath Samarabandu

The purpose of gesture recognition is to recognize meaningful movements of human bodies, and gesture recognition is an important issue in computer vision. In this paper, we present a multimodal gesture recognition method based on 3D densely…

计算机视觉与模式识别 · 计算机科学 2020-01-17 Yi Zhang , Chong Wang , Ye Zheng , Jieyu Zhao , Yuqi Li , Xijiong Xie

Humans can accurately determine whether the object in hand has slipped or not by visual and tactile perception. However, it is still a challenge for robots to detect in-hand object slip through visuo-tactile fusion. To address this issue, a…

机器人学 · 计算机科学 2023-02-28 Junli Gao , Zhaoji Huang , Zhaonian Tang , Haitao Song , Wenyu Liang

Deep neural networks are applied in more and more areas of everyday life. However, they still lack essential abilities, such as robustly dealing with spatially transformed input signals. Approaches to mitigate this severe robustness issue…

机器学习 · 计算机科学 2024-05-28 Johann Schmidt , Sebastian Stober

Image enhancement is a critical task in computer vision and photography that is often entangled with noise. This renders the traditional Image Signal Processing (ISP) ineffective compared to the advances in deep learning. However, the…

图像与视频处理 · 电气工程与系统科学 2026-01-21 Srinivas Miriyala , Sowmya Vajrala , Hitesh Kumar , Sravanth Kodavanti , Vikram Rajendiran

The word-level lipreading approach typically employs a two-stage framework with separate frontend and backend architectures to model dynamic lip movements. Each component has been extensively studied, and in the backend architecture,…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Byung Hoon Lee , Wooseok Shin , Sung Won Han

Pedestrian intention recognition is very important to develop robust and safe autonomous driving (AD) and advanced driver assistance systems (ADAS) functionalities for urban driving. In this work, we develop an end-to-end pedestrian…

Recently, data-driven inertial navigation approaches have demonstrated their capability of using well-trained neural networks to obtain accurate position estimates from inertial measurement units (IMU) measurements. In this paper, we…

机器人学 · 计算机科学 2021-12-22 Bingbing Rao , Ehsan Kazemi , Yifan Ding , Devu M Shila , Frank M. Tucker , Liqiang Wang

Reconstructing dynamic driving scenes from dashcam videos has attracted increasing attention due to its significance in autonomous driving and scene understanding. While recent advances have made impressive progress, most methods still…

计算机视觉与模式识别 · 计算机科学 2025-10-30 Hongyuan Liu , Haochen Yu , Bochao Zou , Jianfei Jiang , Qiankun Liu , Jiansheng Chen , Huimin Ma

In recent years, mobile Internet has accelerated the proliferation of smart mobile development. The mobile payment, mobile security and privacy protection have become the focus of widespread attention. Iris recognition becomes a…

计算机视觉与模式识别 · 计算机科学 2020-06-16 Siming Zheng , Rahmita Wirza O. K. Rahmat , Fatimah Khalid , Nurul Amelina Nasharuddin

Strapdown inertial navigation systems are sensitive to the quality of the data provided by the accelerometer and gyroscope. Low-grade IMUs in handheld smart-devices pose a problem for inertial odometry on these devices. We propose a scheme…

计算机视觉与模式识别 · 计算机科学 2018-08-13 Santiago Cortés , Arno Solin , Juho Kannala

Accurate real-time catheter segmentation is an important pre-requisite for robot-assisted endovascular intervention. Most of the existing learning-based methods for catheter segmentation and tracking are only trained on small-scale datasets…

图像与视频处理 · 电气工程与系统科学 2020-06-17 Anh Nguyen , Dennis Kundrat , Giulio Dagnino , Wenqiang Chi , Mohamed E. M. K. Abdelaziz , Yao Guo , YingLiang Ma , Trevor M. Y. Kwok , Celia Riga , Guang-Zhong Yang

Transferring existing image-based detectors to the video is non-trivial since the quality of frames is always deteriorated by part occlusion, rare pose, and motion blur. Previous approaches exploit to propagate and aggregate features across…

计算机视觉与模式识别 · 计算机科学 2020-07-17 Zhengkai Jiang , Yu Liu , Ceyuan Yang , Jihao Liu , Peng Gao , Qian Zhang , Shiming Xiang , Chunhong Pan

Spiking neural networks (SNNs), as one of the brain-inspired models, has spatio-temporal information processing capability, low power feature, and high biological plausibility. The effective spatio-temporal feature makes it suitable for…

神经与进化计算 · 计算机科学 2022-03-21 Changqing Xu , Yi Liu , Yintang Yang

As the volume of image data grows, data-oriented cloud computing in Internet of Video Things (IoVT) systems encounters latency issues. Task-oriented edge computing addresses this by shifting data analysis to the edge. However, limited…

计算机视觉与模式识别 · 计算机科学 2024-11-05 Jiaqi Wu , Simin Chen , Zehua Wang , Wei Chen , Zijian Tian , F. Richard Yu , Victor C. M. Leung

Surgical guide plate is an important tool for the dental implant surgery. However, the design process heavily relies on the dentist to manually simulate the implant angle and depth. When deep neural networks have been applied to assist the…

计算机视觉与模式识别 · 计算机科学 2024-06-19 Xinquan Yang , Xuguang Li , Xiaoling Luo , Leilei Zeng , Yudi Zhang , Linlin Shen , Yongqiang Deng

This paper addresses the task of segmenting class-agnostic objects in semi-supervised setting. Although previous detection based methods achieve relatively good performance, these approaches extract the best proposal by a greedy strategy,…

计算机视觉与模式识别 · 计算机科学 2020-12-11 Daizong Liu , Shuangjie Xu , Xiao-Yang Liu , Zichuan Xu , Wei Wei , Pan Zhou

Medical image segmentation plays a crucial role in clinical diagnosis and treatment planning. Although models based on convolutional neural networks (CNNs) and Transformers have achieved remarkable success in medical image segmentation…

图像与视频处理 · 电气工程与系统科学 2024-10-04 Jiashu Xu

This paper proposes a novel framework for lung sound event detection, segmenting continuous lung sound recordings into discrete events and performing recognition on each event. Exploiting the lightweight nature of Temporal Convolution…

Deep learning has emerged as a transformative tool in healthcare, offering significant advancements in dental diagnostics by analyzing complex imaging data. This paper presents an enhanced ResNet50 architecture, integrated with the SimAM…

计算机视觉与模式识别 · 计算机科学 2024-07-12 Shahriar Rezaie , Neda Saberitabar , Elnaz Salehi