中文
相关论文

相关论文: A Multi-Stage Temporal Convolutional Network for V…

200 篇论文

Ureteroscopy and cystoscopy are the gold standard methods to identify and treat tumors along the urinary tract. It has been reported that during a normal procedure a rate of 10-20 % of the lesions could be missed. In this work we study the…

图像与视频处理 · 电气工程与系统科学 2021-04-09 Jorge F. Lazo , Sara Moccia , Aldo Marzullo , Michele Catellani , Ottavio De Cobelli , Benoit Rosa , Michel de Mathelin , Elena De Momi

Together with the rapid development of the Internet of Things (IoT), human activity recognition (HAR) using wearable Inertial Measurement Units (IMUs) becomes a promising technology for many research areas. Recently, deep learning-based…

计算机视觉与模式识别 · 计算机科学 2020-10-12 Ling Pei , Songpengcheng Xia , Lei Chu , Fanyi Xiao , Qi Wu , Wenxian Yu , Robert Qiu

Human activity recognition (HAR) with wearables is promising research that can be widely adopted in many smart healthcare applications. In recent years, the deep learning-based HAR models have achieved impressive recognition performance.…

计算机视觉与模式识别 · 计算机科学 2022-08-17 Songpengcheng Xia , Lei Chu , Ling Pei , Wenxian Yu , Robert C. Qiu

Inertial Measurement Unit (IMU) sensors are being increasingly used to detect human gestures and movements. Using a single IMU sensor, whole body movement recognition remains a hard problem because movements may not be adequately captured…

机器学习 · 计算机科学 2020-09-07 Varun Badrinath Krishna

This research proposes and evaluates scoring and assessment methods for Virtual Reality (VR) training simulators. VR simulators capture detailed n-dimensional human motion data which is useful for performance analysis. Custom made medical…

信号处理 · 电气工程与系统科学 2020-06-23 Neil Vaughan , Bogdan Gabrys

Ultrasound Localization Microscopy (ULM) can map microvessels at a resolution of a few micrometers (\mu m). Transcranial ULM remains challenging in presence of aberrations caused by the skull, which lead to localization errors. Herein, we…

图像与视频处理 · 电气工程与系统科学 2023-09-20 Paul Xing , Jonathan Porée , Brice Rauby , Antoine Malescot , Éric Martineau , Vincent Perrot , Ravi L. Rungta , Jean Provost

Most multimodal large language models (MLLMs) treat visual tokens as "a sequence of text", integrating them with text tokens into a large language model (LLM). However, a great quantity of visual tokens significantly increases the demand…

计算机视觉与模式识别 · 计算机科学 2025-03-28 Dongchen Lu , Yuyao Sun , Zilu Zhang , Leping Huang , Jianliang Zeng , Mao Shu , Huo Cao

The dominant paradigm for video-based action segmentation is composed of two steps: first, for each frame, compute low-level features using Dense Trajectories or a Convolutional Neural Network that encode spatiotemporal information locally,…

计算机视觉与模式识别 · 计算机科学 2016-08-31 Colin Lea , Rene Vidal , Austin Reiter , Gregory D. Hager

Deep learning has been demonstrated to achieve excellent results for image classification and object detection. However, the impact of deep learning on video analysis (e.g. action detection and recognition) has been limited due to…

计算机视觉与模式识别 · 计算机科学 2017-08-03 Rui Hou , Chen Chen , Mubarak Shah

In many sports, it is useful to analyse video of an athlete in competition for training purposes. In swimming, stroke rate is a common metric used by coaches; requiring a laborious labelling of each individual stroke. We show that using a…

计算机视觉与模式识别 · 计算机科学 2017-05-30 Brandon Victor , Zhen He , Stuart Morgan , Dino Miniutti

Sports action classification representing complex body postures and player-object interactions is an emerging area in image-based sports analysis. Some works have contributed to automated sports action recognition using machine learning…

计算机视觉与模式识别 · 计算机科学 2025-07-16 Palash Ray , Mahuya Sasmal , Asish Bera

Instructed Visual Segmentation (IVS) tasks require segmenting objects in images or videos based on natural language instructions. While recent multimodal large language models (MLLMs) have achieved strong performance on IVS, their inference…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Wenhui Zhu , Xiwen Chen , Zhipeng Wang , Shao Tang , Sayan Ghosh , Xuanzhao Dong , Rajat Koner , Yalin Wang

Infant sleep is critical to brain and behavioral development. Prior studies on infant sleep/wake classification have been largely limited to reliance on expensive and burdensome polysomnography (PSG) tests in the laboratory or wearable…

多媒体 · 计算机科学 2023-06-29 Kai Chieh Chang , Mark Hasegawa-Johnson , Nancy L. McElwain , Bashima Islam

This paper illustrates how multilevel functional models can detect and characterize biomechanical changes along different sport training sessions. Our analysis focuses on the relevant cases to identify differences in knee biomechanics in…

应用统计 · 统计学 2021-04-07 Marcos Matabuena , Sherveen Riazati , Nick Caplan , Phil Hayes

In multi-person pose estimation, the left/right joint type discrimination is always a hard problem because of the similar appearance. Traditionally, we solve this problem by stacking multiple refinement modules to increase network's…

计算机视觉与模式识别 · 计算机科学 2019-11-27 Ying Huang , Jiankai Zhuang , Zengchang Qin

The issue of failed weaning is a critical concern in the intensive care unit (ICU) setting. This scenario occurs when a patient experiences difficulty maintaining spontaneous breathing and ensuring a patent airway within the first 48 hours…

机器学习 · 计算机科学 2025-03-05 Hernando Gonzalez , Carlos Julio Arizmendi , Beatriz F. Giraldo

Objective: A novel structure based on channel-wise attention mechanism is presented in this paper. Embedding with the proposed structure, an efficient classification model that accepts multi-lead electrocardiogram (ECG) as input is…

信号处理 · 电气工程与系统科学 2020-03-27 Hao Tung , Chao Zheng , Xinsheng Mao , Dahong Qian

Automatic detection of liver lesions in CT images poses a great challenge for researchers. In this work we present a deep learning approach that models explicitly the variability within the non-lesion class, based on prior knowledge of the…

计算机视觉与模式识别 · 计算机科学 2017-07-21 Maayan Frid-Adar , Idit Diamant , Eyal Klang , Michal Amitai , Jacob Goldberger , Hayit Greenspan

Human action recognition has become an important research focus in computer vision due to the wide range of applications where it is used. 3D Resnet-based CNN models, particularly MC3, R3D, and R(2+1)D, have different convolutional filters…

计算机视觉与模式识别 · 计算机科学 2026-01-19 Mohammad Rasras , Iuliana Marin , Serban Radu , Irina Mocanu

An abdominal ultrasound examination, which is the most common ultrasound examination, requires substantial manual efforts to acquire standard abdominal organ views, annotate the views in texts, and record clinically relevant organ…

计算机视觉与模式识别 · 计算机科学 2018-06-06 Zhoubing Xu , Yuankai Huo , JinHyeong Park , Bennett Landman , Andy Milkowski , Sasa Grbic , Shaohua Zhou