中文
相关论文

相关论文: Open surgery tool classification and hand utilizat…

200 篇论文

Typical attempts to improve the capability of visual place recognition techniques include the use of multi-sensor fusion and integration of information over time from image sequences. These approaches can improve performance but have…

机器人学 · 计算机科学 2019-03-11 Stephen Hausler , Adam Jacobson , Michael Milford

Depth estimation is a core task in 3D computer vision. Recent methods investigate the task of monocular depth trained with various depth sensor modalities. Every sensor has its advantages and drawbacks caused by the nature of estimates. In…

Correct fusion of data from two sensors is not possible without an accurate estimate of their relative pose, which can be determined through the process of extrinsic calibration. When two or more sensors are capable of producing their own…

机器人学 · 计算机科学 2021-11-18 Emmett Wise , Matthew Giamou , Soroush Khoubyarian , Abhinav Grover , Jonathan Kelly

Plenoptic cameras enable the capturing of spatial as well as angular color information which can be used for various applications among which are image refocusing and depth calculations. However, these cameras are expensive and research in…

图像与视频处理 · 电气工程与系统科学 2022-04-12 Tim Michels , Arne Petersen , Luca Palmieri , Reinhard Koch

3D face reconstruction (3DFR) algorithms are based on specific assumptions tailored to distinct application scenarios. These assumptions limit their use when acquisition conditions, such as the subject's distance from the camera or the…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Simone Maurizio La Cava , Sara Concas , Ruben Tolosana , Roberto Casula , Giulia Orrù , Martin Drahansky , Julian Fierrez , Gian Luca Marcialis

Minimally invasive colorectal surgery is characterized by procedural variability, a difficult learning curve, and complications that impact quality and outcomes. Video-based assessment (VBA) offers an opportunity to generate data-driven…

A lensless camera is an imaging system that uses a mask in place of a lens, making it thinner, lighter, and less expensive than a lensed camera. However, additional complex computation and time are required for image reconstruction. This…

计算机视觉与模式识别 · 计算机科学 2022-10-26 Yinger Zhang , Zhouyi Wu , Peiying Lin , Yang Pan , Yuting Wu , Liufang Zhang , Jiangtao Huangfu

Occlusion poses a great threat to monocular multi-person 3D human pose estimation due to large variability in terms of the shape, appearance, and position of occluders. While existing methods try to handle occlusion with pose…

计算机视觉与模式识别 · 计算机科学 2022-08-02 Qihao Liu , Yi Zhang , Song Bai , Alan Yuille

Video and wearable sensor data provide complementary information about human movement. Video provides a holistic understanding of the entire body in the world while wearable sensors provide high-resolution measurements of specific body…

计算机视觉与模式识别 · 计算机科学 2024-11-18 J. D. Peiffer , Kunal Shah , Shawana Anarwala , Kayan Abdou , R. James Cotton

Hand pose estimation has matured rapidly in recent years. The introduction of commodity depth sensors and a multitude of practical applications have spurred new advances. We provide an extensive analysis of the state-of-the-art, focusing on…

计算机视觉与模式识别 · 计算机科学 2015-05-08 James Steven Supancic , Gregory Rogez , Yi Yang , Jamie Shotton , Deva Ramanan

Multichannel, infinite-conjugate optical systems easily allow implementation of multiple image paths and imaging modes into a single microscope. Traditional optical alignment methods which rely on additional hardware are not always simple…

光学 · 物理学 2024-07-16 Gemma S. Cairns , Brian R. Patton

We propose a multi-camera LiDAR-visual-inertial odometry framework, Multi-LVI-SAM, which fuses data from multiple fisheye cameras, LiDAR and inertial sensors for highly accurate and robust state estimation. To enable efficient and…

计算机视觉与模式识别 · 计算机科学 2025-09-09 Xinyu Zhang , Kai Huang , Junqiao Zhao , Zihan Yuan , Tiantian Feng

One of the challenges in evaluating multi-object video detection, tracking and classification systems is having publically available data sets with which to compare different systems. However, the measures of performance for tracking and…

计算机视觉与模式识别 · 计算机科学 2017-04-24 Avishek Chakraborty , Victor Stamatescu , Sebastien C. Wong , Grant Wigley , David Kearney

Current human pose estimation systems focus on retrieving an accurate 3D global estimate of a single person. Therefore, this paper presents one of the first 3D multi-person human pose estimation systems that is able to work in real-time and…

计算机视觉与模式识别 · 计算机科学 2024-03-15 Pawel Knap , Peter Hardy , Alberto Tamajo , Hwasup Lim , Hansung Kim

We revisit the study of a wrist-mounted camera system (referred to as HandCam) for recognizing activities of hands. HandCam has two unique properties as compared to egocentric systems (referred to as HeadCam): (1) it avoids the need to…

计算机视觉与模式识别 · 计算机科学 2016-03-23 Cheng-Sheng Chan , Shou-Zhong Chen , Pei-Xuan Xie , Chiung-Chih Chang , Min Sun

We introduce YOLO-pose, a novel heatmap-free approach for joint detection, and 2D multi-person pose estimation in an image based on the popular YOLO object detection framework. Existing heatmap based two-stage approaches are sub-optimal as…

计算机视觉与模式识别 · 计算机科学 2022-04-15 Debapriya Maji , Soyeb Nagori , Manu Mathew , Deepak Poddar

Articulated hand pose tracking is an under-explored problem that carries the potential for use in an extensive number of applications, especially in the medical domain. With a robust and accurate tracking system on surgical videos, the…

计算机视觉与模式识别 · 计算机科学 2025-02-10 Nathan Louis , Luowei Zhou , Steven J. Yule , Roger D. Dias , Milisa Manojlovich , Francis D. Pagani , Donald S. Likosky , Jason J. Corso

We address the challenge of accurate 3D human pose and shape estimation from monocular images. The key to accuracy and robustness lies in high-quality training data. Existing training datasets containing real images with pseudo ground truth…

计算机视觉与模式识别 · 计算机科学 2024-11-14 Priyanka Patel , Michael J. Black

Instructional cataract surgery videos are crucial for ophthalmologists and trainees to observe surgical details repeatedly. This paper presents a deep learning model for real-time identification of surgical instruments in these videos,…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Sanya Sinha , Michal Balazia , Francois Bremond

Self-captured full-body videos are popular, but most deployments require mounted cameras, carefully-framed shots, and repeated practice. We propose a more convenient solution that enables full-body video capture using handheld mobile…

计算机视觉与模式识别 · 计算机科学 2025-12-02 Bowei Chen , Brian Curless , Ira Kemelmacher-Shlizerman , Steven M. Seitz
‹ 上一页 1 8 9 10 下一页 ›