中文
相关论文

相关论文: Visual Gyroscope: Combination of Deep Learning Fea…

200 篇论文

Predictive world models that simulate future observations under explicit camera control are fundamental to interactive AI. Despite rapid advances, current systems lack spatial persistence: they fail to maintain stable scene structures over…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Chendong Xiang , Jiajun Liu , Jintao Zhang , Xiao Yang , Zhengwei Fang , Shizun Wang , Zijun Wang , Yingtian Zou , Hang Su , Jun Zhu

Trajectory prediction is, naturally, a key task for vehicle autonomy. While the number of traffic rules is limited, the combinations and uncertainties associated with each agent's behaviour in real-world scenarios are nearly impossible to…

计算机视觉与模式识别 · 计算机科学 2023-12-21 Sushil Sharma , Arindam Das , Ganesh Sistu , Mark Halton , Ciarán Eising

It's a practical approach using the ground-aerial collaborative system to enhance the localization robustness of flying robots in cluttered environments, especially when visual sensors degrade. Conventional approaches estimate the flying…

机器人学 · 计算机科学 2025-12-19 Sijia Chen , Wei Dong

Camera with a fisheye or ultra-wide lens covers a wide field of view that cannot be modeled by the perspective projection. Serious fisheye lens distortion in the peripheral region of the image leads to degraded performance of the existing…

计算机视觉与模式识别 · 计算机科学 2024-04-11 Bing Li , Dong Zhang , Cheng Huang , Yun Xian , Ming Li , Dah-Jye Lee

Creating mobile robots which are able to find and manipulate objects in large environments is an active topic of research. These robots not only need to be capable of searching for specific objects but also to estimate their poses often…

机器人学 · 计算机科学 2022-03-09 Jascha Hellwig , Mark Baierl , Joao Carvalho , Julen Urain , Jan Peters

We introduce a convolutional neural network model for unsupervised learning of depth and ego-motion from cylindrical panoramic video. Panoramic depth estimation is an important technology for applications such as virtual reality, 3D…

计算机视觉与模式识别 · 计算机科学 2020-08-20 Alisha Sharma , Jonathan Ventura

The increasing frequency of firearm-related incidents has necessitated advancements in security and surveillance systems, particularly in firearm detection within public spaces. Traditional gun detection methods rely on manual inspections…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Amulya Reddy Maligireddy , Manohar Reddy Uppula , Nidhi Rastogi , Yaswanth Reddy Parla

Wide-range and fine-grained vehicle detection plays a critical role in enabling active safety features in intelligent driving systems. However, existing vehicle detection methods based on rectangular bounding boxes (BBox) often struggle…

计算机视觉与模式识别 · 计算机科学 2023-09-18 Zhupeng Ye , Yinqi Li , Zejian Yuan

Estimating the 3D shape of an object from a single or multiple images has gained popularity thanks to the recent breakthroughs powered by deep learning. Most approaches regress the full object shape in a canonical pose, possibly…

计算机视觉与模式识别 · 计算机科学 2020-11-19 Riccardo Spezialetti , David Joseph Tan , Alessio Tonioni , Keisuke Tateno , Federico Tombari

Accurate and robust object pose estimation for robotics applications requires verification and refinement steps. In this work, we propose to integrate hypotheses verification with object pose refinement guided by physics simulation. This…

计算机视觉与模式识别 · 计算机科学 2020-05-19 Dominik Bauer , Timothy Patten , Markus Vincze

We present a novel two-view geometry estimation framework which is based on a differentiable robust loss function fitting. We propose to treat the robust fundamental matrix estimation as an implicit layer, which allows us to avoid…

计算机视觉与模式识别 · 计算机科学 2024-10-24 Vladislav Pyatov , Iaroslav Koshelev , Stamatis Lefkimmiatis

We present a novel Structure from Motion pipeline that is capable of reconstructing accurate camera poses for panorama-style video capture without prior camera intrinsic calibration. While panorama-style capture is common and convenient,…

计算机视觉与模式识别 · 计算机科学 2019-06-11 Chris Sweeney , Aleksander Holynski , Brian Curless , Steve M Seitz

In this paper we propose an end-to-end learnable approach that detects static urban objects from multiple views, re-identifies instances, and finally assigns a geographic position per object. Our method relies on a Graph Neural Network…

计算机视觉与模式识别 · 计算机科学 2020-03-25 Ahmed Samy Nassar , Stefano D'Aronco , Sébastien Lefèvre , Jan D. Wegner

We describe a learning-based approach to hand-eye coordination for robotic grasping from monocular images. To learn hand-eye coordination for grasping, we trained a large convolutional neural network to predict the probability that…

机器学习 · 计算机科学 2016-08-30 Sergey Levine , Peter Pastor , Alex Krizhevsky , Deirdre Quillen

We introduce ViewNeRF, a Neural Radiance Field-based viewpoint estimation method that learns to predict category-level viewpoints directly from images during training. While NeRF is usually trained with ground-truth camera poses, multiple…

计算机视觉与模式识别 · 计算机科学 2022-12-02 Octave Mariotti , Oisin Mac Aodha , Hakan Bilen

Vehicles of higher automation levels require the creation of situation awareness. One important aspect of this situation awareness is an understanding of the current risk of a driving situation. In this work, we present a novel approach for…

计算机视觉与模式识别 · 计算机科学 2018-06-22 Patrik Feth , Mohammed Naveed Akram , René Schuster , Oliver Wasenmüller

This work describes the automatic registration of a large network (approximately 40) of fixed, ceiling-mounted environment cameras spread over a large area (approximately 800 squared meters) using a mobile calibration robot equipped with a…

机器人学 · 计算机科学 2022-10-11 Subodh Mishra , Sushruth Nagesh , Sagar Manglani , Graham Mills , Punarjay Chakravarty , Gaurav Pandey

We propose a vision-based method that localizes a ground vehicle using publicly available satellite imagery as the only prior knowledge of the environment. Our approach takes as input a sequence of ground-level images acquired by the…

机器人学 · 计算机科学 2022-03-08 Dong-Ki Kim , Matthew R. Walter

Traditional approaches for Visual Simultaneous Localization and Mapping (VSLAM) rely on low-level vision information for state estimation, such as handcrafted local features or the image gradient. While significant progress has been made…

机器人学 · 计算机科学 2021-08-05 Huaiyang Huang , Haoyang Ye , Yuxiang Sun , Lujia Wang , Ming Liu

This paper aims to design a 3D object detection model from 2D images taken by monocular cameras by combining the estimated bird's-eye view elevation map and the deep representation of object features. The proposed model has a pre-trained…

计算机视觉与模式识别 · 计算机科学 2020-11-25 Ali Babolhavaeji , Mohammad Fanaei