中文
相关论文

相关论文: YOLOv10-Based Multi-Task Framework for Hand Locali…

200 篇论文

This paper presents a generalized model for real-time detection of flying objects that can be used for transfer learning and further research, as well as a refined model that achieves state-of-the-art results for flying object detection. We…

计算机视觉与模式识别 · 计算机科学 2024-05-24 Dillon Reis , Jordan Kupec , Jacqueline Hong , Ahmad Daoudi

The reliable identification of mitotic figures in whole-slide histopathological images remains difficult, owing to their low prevalence, substantial morphological heterogeneity, and the inconsistencies introduced by tissue processing and…

图像与视频处理 · 电气工程与系统科学 2025-09-23 Navya Sri Kelam , Akash Parekh , Saikiran Bonthu , Nitin Singhal

The surgical usage of Mixed Reality (MR) has received growing attention in areas such as surgical navigation systems, skill assessment, and robot-assisted surgeries. For such applications, pose estimation for hand and surgical instruments…

计算机视觉与模式识别 · 计算机科学 2023-10-05 Rui Wang , Sophokles Ktistakis , Siwei Zhang , Mirko Meboldt , Quentin Lohmeyer

Camera traps offer enormous new opportunities in ecological studies, but current automated image analysis methods often lack the contextual richness needed to support impactful conservation outcomes. Here we present an integrated approach…

计算机视觉与模式识别 · 计算机科学 2024-11-22 Paul Fergus , Carl Chalmers , Naomi Matthews , Stuart Nixon , Andre Burger , Oliver Hartley , Chris Sutherland , Xavier Lambin , Steven Longmore , Serge Wich

This paper presents an architectural analysis of YOLOv12, a significant advancement in single-stage, real-time object detection building upon the strengths of its predecessors while introducing key improvements. The model incorporates an…

计算机视觉与模式识别 · 计算机科学 2025-02-21 Mujadded Al Rabbani Alif , Muhammad Hussain

Object detection and localization are crucial tasks for biomedical image analysis, particularly in the field of hematology where the detection and recognition of blood cells are essential for diagnosis and treatment decisions. While…

计算机视觉与模式识别 · 计算机科学 2023-12-19 Shun Liu , Jianan Zhang , Ruocheng Song , Teik Toe Teoh

Surgery is a high-stakes domain where surgeons must navigate critical anatomical structures and actively avoid potential complications while achieving the main task at hand. Such surgical activity has been shown to affect long-term patient…

We propose a method for hand pose estimation based on a deep regressor trained on two different kinds of input. Raw depth data is fused with an intermediate representation in the form of a segmentation of the hand into parts. This…

计算机视觉与模式识别 · 计算机科学 2017-09-18 Natalia Neverova , Christian Wolf , Florian Nebout , Graham Taylor

Image-based tracking of laparoscopic instruments plays a fundamental role in computer and robotic-assisted surgeries by aiding surgeons and increasing patient safety. Computer vision contests, such as the Robust Medical Instrument…

计算机视觉与模式识别 · 计算机科学 2022-01-21 Juan Carlos Angeles Ceron , Leonardo Chang , Gilberto Ochoa-Ruiz , Sharib Ali

In the construction sector, ensuring worker safety is of the utmost significance. In this study, a deep learning-based technique is presented for identifying safety gear worn by construction workers, such as helmets, goggles, jackets,…

计算机视觉与模式识别 · 计算机科学 2024-08-31 Md. Shariful Islam , SM Shaqib , Shahriar Sultan Ramit , Shahrun Akter Khushbu , Abdus Sattar , Sheak Rashed Haider Noori

Manual peripheral blood smear (PBS) analysis is labor intensive and subjective. While deep learning offers a promising alternative, a systematic evaluation of state of the art models such as YOLOv11 for fine grained PBS detection is still…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Mohamad Abou Ali , Mariam Abdulfattah , Baraah Al Hussein , Fadi Dornaika , Ali Cherry , Mohamad Hajj-Hassan , Lara Hamawy

In industrial scenarios, effective human-robot collaboration relies on multi-camera systems to robustly monitor human operators despite the occlusions that typically show up in a robotic workcell. In this scenario, precise localization of…

机器人学 · 计算机科学 2024-06-18 Davide Allegro , Matteo Terreran , Stefano Ghidoni

Satellite remote sensing images pose significant challenges for object detection due to their high resolution, complex scenes, and large variations in target scales. To address the insufficient detection accuracy of the YOLOv11n model in…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Shuaiyu Zhu , Sergey Ablameyko

Utilizing large language models (LLMs) to compose off-the-shelf visual tools represents a promising avenue of research for developing robust visual assistants capable of addressing diverse visual tasks. However, these methods often overlook…

计算机视觉与模式识别 · 计算机科学 2024-04-11 Zhi Gao , Yuntao Du , Xintong Zhang , Xiaojian Ma , Wenjuan Han , Song-Chun Zhu , Qing Li

We propose an approach to estimate arm and hand dynamics from monocular video by utilizing the relationship between arm and hand. Although monocular full human motion capture technologies have made great progress in recent years, recovering…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Shuying Liu , Wenbin Wu , Jiaxian Wu , Yue Lin

In surgical training for medical students, proficiency development relies on expert-led skill assessment, which is costly, time-limited, difficult to scale, and its expertise remains confined to institutions with available specialists.…

计算机视觉与模式识别 · 计算机科学 2026-03-30 Le Ma , Thiago Freitas dos Santos , Nadia Magnenat-Thalmann , Katarzyna Wac

Recovering world space 4D motion of two interacting hands from egocentric video is a fundamental capability for supervising robot policy learning, where wrist trajectories track the end-effector and finger articulations specify the grasp…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Huajian Zeng , Chaohua Yao , Yuantai Zhang , Jiaqi Yang , Rolandos Alexandros Potamias , Xingxing Zuo

This study examines the effectiveness of spatio-temporal modeling and the integration of spatial attention mechanisms in deep learning models for underwater object detection. Specifically, in the first phase, the performance of…

计算机视觉与模式识别 · 计算机科学 2025-10-31 Sai Likhith Karri , Ansh Saxena

While many recent hand pose estimation methods critically rely on a training set of labelled frames, the creation of such a dataset is a challenging task that has been overlooked so far. As a result, existing datasets are limited to a few…

计算机视觉与模式识别 · 计算机科学 2016-12-05 Markus Oberweger , Gernot Riegler , Paul Wohlhart , Vincent Lepetit

This paper presents a deep learning-based wound classification tool that can assist medical personnel in non-wound care specialization to classify five key wound conditions, namely deep wound, infected wound, arterial wound, venous wound,…

计算机视觉与模式识别 · 计算机科学 2023-03-30 Po-Hsuan Huang , Yi-Hsiang Pan , Ying-Sheng Luo , Yi-Fan Chen , Yu-Cheng Lo , Trista Pei-Chun Chen , Cherng-Kang Perng