中文
相关论文

相关论文: Virtual Worlds as Proxy for Multi-Object Tracking …

200 篇论文

We present a challenging dataset, the TartanAir, for robot navigation tasks and more. The data is collected in photo-realistic simulation environments with the presence of moving objects, changing light and various weather conditions. By…

机器人学 · 计算机科学 2020-08-11 Wenshan Wang , Delong Zhu , Xiangwei Wang , Yaoyu Hu , Yuheng Qiu , Chen Wang , Yafei Hu , Ashish Kapoor , Sebastian Scherer

We introduce and we analyze a new dataset which resembles the input to biological vision systems much more than most previously published ones. Our analysis leaded to several important conclusions. First, it is possible to disambiguate over…

计算机视觉与模式识别 · 计算机科学 2013-04-29 Alessandro Perina , Nebojsa Jojic

Accurate detection of 3D objects is a fundamental problem in computer vision and has an enormous impact on autonomous cars, augmented/virtual reality and many applications in robotics. In this work we present a novel fusion of neural…

计算机视觉与模式识别 · 计算机科学 2019-04-17 Martin Simon , Karl Amende , Andrea Kraus , Jens Honer , Timo Sämann , Hauke Kaulbersch , Stefan Milz , Horst Michael Gross

With the increasing global popularity of self-driving cars, there is an immediate need for challenging real-world datasets for benchmarking and training various computer vision tasks such as 3D object detection. Existing datasets either…

计算机视觉与模式识别 · 计算机科学 2019-09-18 Quang-Hieu Pham , Pierre Sevestre , Ramanpreet Singh Pahwa , Huijing Zhan , Chun Ho Pang , Yuda Chen , Armin Mustafa , Vijay Chandrasekhar , Jie Lin

The task of detecting 3D objects is important to various robotic applications. The existing deep learning-based detection techniques have achieved impressive performance. However, these techniques are limited to run with a graphics…

计算机视觉与模式识别 · 计算机科学 2020-08-14 Xuesong Li , Jose Guivant , Subhan Khan

Multi-People Tracking in an open-world setting requires a special effort in precise detection. Moreover, temporal continuity in the detection phase gains more importance when scene cluttering introduces the challenging problems of occluded…

计算机视觉与模式识别 · 计算机科学 2018-09-19 Matteo Fabbri , Fabio Lanzi , Simone Calderara , Andrea Palazzi , Roberto Vezzani , Rita Cucchiara

In the last decade many different algorithms have been proposed to track a generic object in videos. Their execution on recent large-scale video datasets can produce a great amount of various tracking behaviours. New trends in Reinforcement…

计算机视觉与模式识别 · 计算机科学 2020-03-10 Matteo Dunnhofer , Niki Martinel , Gian Luca Foresti , Christian Micheloni

Image captioning is a computer vision task that involves generating natural language descriptions for images. This method has numerous applications in various domains, including image retrieval systems, medicine, and various industries.…

计算机视觉与模式识别 · 计算机科学 2023-08-08 Sai Suprabhanu Nallapaneni , Subrahmanyam Konakanchi

The ability to recognize objects is an essential skill for a robotic system acting in human-populated environments. Despite decades of effort from the robotic and vision research communities, robots are still missing good visual perceptual…

机器人学 · 计算机科学 2018-05-23 Mohammad Reza Loghmani , Barbara Caputo , Markus Vincze

Recent progress in advanced driver assistance systems and the race towards autonomous vehicles is mainly driven by two factors: (1) increasingly sophisticated algorithms that interpret the environment around the vehicle and react…

计算机视觉与模式识别 · 计算机科学 2017-04-04 Marius Cordts , Timo Rehfeld , Lukas Schneider , David Pfeiffer , Markus Enzweiler , Stefan Roth , Marc Pollefeys , Uwe Franke

Interacting with real-world objects in Mixed Reality (MR) often proves difficult when they are crowded, distant, or partially occluded, hindering straightforward selection and manipulation. We observe that these difficulties stem from…

人机交互 · 计算机科学 2025-07-25 Xiaoan Liu , Difan Jia , Xianhao Carton Liu , Mar Gonzalez-Franco , Chen Zhu-Tian

Originally designed for applications in computer graphics, visual computing (VC) methods synthesize information about physical and virtual worlds, using prescribed algorithms optimized for spatial computing. VC is used to analyze geometry,…

In this dissertation, we investigated and enhanced Deep Learning (DL) techniques for counting objects, like pedestrians, cells or vehicles, in still images or video frames. In particular, we tackled the challenge related to the lack of data…

计算机视觉与模式识别 · 计算机科学 2022-06-09 Luca Ciampi

Dense prediction tasks hold significant importance of computer vision, aiming to learn pixel-wise annotated labels for input images. Despite advances in this field, existing methods primarily focus on idealized conditions, exhibiting…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Changliang Xia , Chengyou Jia , Zhuohang Dang , Minnan Luo , Zhihui Li , Xiaojun Chang

Data seems cheap to get, and in many ways it is, but the process of creating a high quality labeled dataset from a mass of data is time-consuming and expensive. With the advent of rich 3D repositories, photo-realistic rendering systems…

计算机视觉与模式识别 · 计算机科学 2016-09-09 Yair Movshovitz-Attias , Takeo Kanade , Yaser Sheikh

Visual tracking (VT) is the process of locating a moving object of interest in a video. It is a fundamental problem in computer vision, with various applications in human-computer interaction, security and surveillance, robot perception,…

量子物理 · 物理学 2019-02-06 Chao-Hua Yu , Fei Gao , Chenghuan Liu , Du Huynh , Mark Reynolds , Jingbo Wang

Classical visual simultaneous localization and mapping (SLAM) algorithms usually assume the environment to be rigid. This assumption limits the applicability of those algorithms as they are unable to accurately estimate the camera poses and…

机器人学 · 计算机科学 2022-09-28 Mathieu Gonzalez , Eric Marchand , Amine Kacete , Jérôme Royan

Mobile service robots are increasingly prevalent in human-centric, real-world domains, operating autonomously in unconstrained indoor environments. In such a context, robotic vision plays a central role in enabling service robots to…

机器人学 · 计算机科学 2025-10-20 Michele Antonazzi , Matteo Luperto , N. Alberto Borghese , Nicola Basilico

Vision Transformers (ViTs) have emerged as popular models in computer vision, demonstrating state-of-the-art performance across various tasks. This success typically follows a two-stage strategy involving pre-training on large-scale…

计算机视觉与模式识别 · 计算机科学 2024-02-07 Zijun Long , Zaiqiao Meng , Gerardo Aragon Camarasa , Richard McCreadie

When training object detection models on synthetic data, it is important to make the distribution of synthetic data as close as possible to the distribution of real data. We investigate specifically the impact of object placement…