中文
相关论文

相关论文: A Deep Learning Based 6 Degree-of-Freedom Localiza…

200 篇论文

We present an approach for recognizing all objects in a scene and estimating their full pose from an accurate 3D instance-aware semantic reconstruction using an RGB-D camera. Our framework couples convolutional neural networks (CNNs) and a…

机器人学 · 计算机科学 2019-10-01 Dinh-Cuong Hoang , Todor Stoyanov , Achim J. Lilienthal

Robotic grasping, the ability of robots to reliably secure and manipulate objects of varying shapes, sizes and orientations, is a complex task that requires precise perception and control. Deep neural networks have shown remarkable success…

Deep learning has significantly advanced computer vision and natural language processing. While there have been some successes in robotics using deep learning, it has not been widely adopted. In this paper, we present a novel robotic grasp…

机器人学 · 计算机科学 2017-07-25 Sulabh Kumra , Christopher Kanan

Computed tomography (CT)-guided needle biopsies are critical for diagnosing a range of conditions, including lung cancer, but present challenges such as limited in-bore space, prolonged procedure times, and radiation exposure. Robotic…

机器人学 · 计算机科学 2025-06-02 Peihan Zhang , Florian Richter , Ishan Duriseti , Albert Hsiao , Sean Tutton , Alexander Norbash , Michael Yip

Automated patient positioning is a crucial step in streamlining MRI workflows and enhancing patient throughput. RGB-D camera-based systems offer a promising approach to automate this process by leveraging depth information to estimate…

计算机视觉与模式识别 · 计算机科学 2025-04-01 Eytan Kats , Kai Geißler , Jochen G. Hirsch , Stefan Heldman , Mattias P. Heinrich

We propose a novel deep convolutional neural network (CNN) based multi-task learning approach for open-set visual recognition. We combine a classifier network and a decoder network with a shared feature extractor network within a multi-task…

计算机视觉与模式识别 · 计算机科学 2019-03-11 Poojan Oza , Vishal M. Patel

Precise and real-time detection of gastrointestinal polyps during endoscopic procedures is crucial for early diagnosis and prevention of colorectal cancer. This work presents EndoSight AI, a deep learning architecture developed and…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Daniel Cavadia

Objective: Depth estimation is crucial for endoscopic navigation and manipulation, but obtaining ground-truth depth maps in real clinical scenarios, such as the colon, is challenging. This study aims to develop a robust framework that…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Sijia Du , Chengfeng Zhou , Suncheng Xiang , Jianwei Xu , Dahong Qian

Light-field microscopes are able to capture spatial and angular information of incident light rays. This allows reconstructing 3D locations of neurons from a single snap-shot.In this work, we propose a model-inspired deep learning approach…

图像与视频处理 · 电气工程与系统科学 2021-03-11 Pingfan Song , Herman Verinaz Jadan , Carmel L. Howe , Peter Quicke , Amanda J. Foust , Pier Luigi Dragotti

This paper proposes a deep learning (DL) model for automatic sleep stage classification based on single-channel EEG data. The DL model features a convolutional neural network (CNN) and transformers. The model was designed to run on energy…

信号处理 · 电气工程与系统科学 2022-11-24 Zongyan Yao , Xilin Liu

Deep convolutional neural networks (CNNs) are the backbone of state-of-art semantic image segmentation systems. Recent work has shown that complementing CNNs with fully-connected conditional random fields (CRFs) can significantly enhance…

计算机视觉与模式识别 · 计算机科学 2016-06-03 Liang-Chieh Chen , Jonathan T. Barron , George Papandreou , Kevin Murphy , Alan L. Yuille

Navigating surgical tools in the dynamic and tortuous anatomy of the lung's airways requires accurate, real-time localization of the tools with respect to the preoperative scan of the anatomy. Such localization can inform human operators or…

计算机视觉与模式识别 · 计算机科学 2018-09-18 Jake Sganga , David Eng , Chauncey Graetzel , David Camarillo

In this work we propose a novel approach to perform segmentation by leveraging the abstraction capabilities of convolutional neural networks (CNNs). Our method is based on Hough voting, a strategy that allows for fully automatic…

In contrast to fully connected networks, Convolutional Neural Networks (CNNs) achieve efficiency by learning weights associated with local filters with a finite spatial extent. An implication of this is that a filter may know what it is…

计算机视觉与模式识别 · 计算机科学 2020-01-24 Md Amirul Islam , Sen Jia , Neil D. B. Bruce

Video capsule endoscopy is a hot topic in computer vision and medicine. Deep learning can have a positive impact on the future of video capsule endoscopy technology. It can improve the anomaly detection rate, reduce physicians' time for…

图像与视频处理 · 电气工程与系统科学 2022-06-17 Abhishek Srivastava , Nikhil Kumar Tomar , Ulas Bagci , Debesh Jha

Endoscopic depth estimation is a critical technology for improving the safety and precision of minimally invasive surgery. It has attracted considerable attention from researchers in medical imaging, computer vision, and robotics. Over the…

计算机视觉与模式识别 · 计算机科学 2025-10-16 Ke Niu , Zeyun Liu , Xue Feng , Heng Li , Qika Lin , Kaize Shi

An accurate seizure prediction system enables early warnings before seizure onset of epileptic patients. It is extremely important for drug-refractory patients. Conventional seizure prediction works usually rely on features extracted from…

信号处理 · 电气工程与系统科学 2021-08-18 Yankun Xu , Jie Yang , Shiqi Zhao , Hemmings Wu , Mohamad Sawan

This paper provides an initial investigation on the application of convolutional neural networks (CNNs) for fingerprint-based positioning using measured massive MIMO channels. When represented in appropriate domains, massive MIMO channels…

机器学习 · 统计学 2017-08-22 Joao Vieira , Erik Leitinger , Muris Sarajlic , Xuhong Li , Fredrik Tufvesson

Augmented Reality has been subject to various integration efforts within industries due to its ability to enhance human machine interaction and understanding. Neural networks have achieved remarkable results in areas of computer vision,…

计算机视觉与模式识别 · 计算机科学 2020-01-17 Linh Kästner , Daniel Dimitrov , Jens Lambrecht

How can we effectively utilise the 2D monocular image information for recovering the 6D pose (6-DoF) of the visual objects? Deep learning has shown to be effective for robust and real-time monocular pose estimation. Oftentimes, the network…

计算机视觉与模式识别 · 计算机科学 2020-03-27 Di Wu , Yihao Chen , Xianbiao Qi , Yongjian Yu , Weixuan Chen , Rong Xiao