中文
相关论文

相关论文: Implementing AI-powered semantic character recogni…

200 篇论文

Video Captioning is considered to be one of the most challenging problems in the field of computer vision. Video Captioning involves the combination of different deep learning models to perform object detection, action detection, and…

计算机视觉与模式识别 · 计算机科学 2021-04-08 Soheyla Amirian , Abolfazl Farahani , Hamid R. Arabnia , Khaled Rasheed , Thiab R. Taha

We propose a general way to integrate procedural knowledge of a domain into deep learning models. We apply it to the case of video prediction, building on top of object-centric deep models and show that this leads to a better performance…

计算机视觉与模式识别 · 计算机科学 2024-06-27 Patrick Takenaka , Johannes Maucher , Marco F. Huber

Automatically generating a summary of sports video poses the challenge of detecting interesting moments, or highlights, of a game. Traditional sports video summarization methods leverage editing conventions of broadcast sports video that…

计算机视觉与模式识别 · 计算机科学 2018-04-16 Antonio Tejero-de-Pablos , Yuta Nakashima , Tomokazu Sato , Naokazu Yokoya , Marko Linna , Esa Rahtu

Deep learning has been successfully applied to several problems related to autonomous driving, often relying on large databases of real target-domain images for proper training. The acquisition of such real-world data is not always possible…

计算机视觉与模式识别 · 计算机科学 2020-08-04 Lucas Tabelini , Rodrigo Berriel , Thiago M. Paixão , Alberto F. De Souza , Claudine Badue , Nicu Sebe , Thiago Oliveira-Santos

Driver inattention is a large problem on the roads around the world. The objective of this project was to develop an eye tracking algorithm with sufficient computational efficiency and accuracy, to successfully realize when the driver was…

计算机视觉与模式识别 · 计算机科学 2019-08-26 Matthew Kowal , Gillian Sandison , Len Yabuki-Soh , Raner la Bastide

This paper presents to the best of our knowledge the first end-to-end object tracking approach which directly maps from raw sensor input to object tracks in sensor space without requiring any feature engineering or system identification in…

机器学习 · 计算机科学 2016-03-10 Peter Ondruska , Ingmar Posner

The notion of learning underlies almost every evolution of Intelligent Agents. In this paper, we present an approach for searching and detecting a given entity in a video sequence. Specifically, we study how the deep learning technique by…

计算机视觉与模式识别 · 计算机科学 2025-12-11 Nzakiese Mbongo , Ngombo Armando

Scene understanding and object recognition is a difficult to achieve yet crucial skill for robots. Recently, Convolutional Neural Networks (CNN), have shown success in this task. However, there is still a gap between their performance on…

机器人学 · 计算机科学 2017-01-18 Sepehr Valipour , Camilo Perez , Martin Jagersand

Over the past few years, deep learning techniques have achieved tremendous success in many visual understanding tasks such as object detection, image segmentation, and caption generation. Despite this thriving in computer vision and natural…

计算机视觉与模式识别 · 计算机科学 2019-03-26 Anh Nguyen

We present a control approach for autonomous vehicles based on deep reinforcement learning. A neural network agent is trained to map its estimated state to acceleration and steering commands given the objective of reaching a specific target…

机器人学 · 计算机科学 2020-03-16 Andreas Folkers , Matthias Rick , Christof Büskens

With the advancement in computer vision deep learning, systems now are able to analyze an unprecedented amount of rich visual information from videos to enable applications such as autonomous driving, socially-aware robot assistant and…

计算机视觉与模式识别 · 计算机科学 2021-07-19 Junwei Liang

Images depicting complex, dynamic scenes are challenging to parse automatically, requiring both high-level comprehension of the overall situation and fine-grained identification of participating entities and their interactions. Current…

计算机视觉与模式识别 · 计算机科学 2025-05-06 Shahaf Pruss , Morris Alper , Hadar Averbuch-Elor

This paper addresses the problem of predicting hazards that drivers may encounter while driving a car. We formulate it as a task of anticipating impending accidents using a single input image captured by car dashcams. Unlike existing…

计算机视觉与模式识别 · 计算机科学 2024-07-02 Korawat Charoenpitaks , Van-Quang Nguyen , Masanori Suganuma , Masahiro Takahashi , Ryoma Niihara , Takayuki Okatani

Comma.ai's approach to Artificial Intelligence for self-driving cars is based on an agent that learns to clone driver behaviors and plans maneuvers by simulating future events in the road. This paper illustrates one of our research…

机器学习 · 计算机科学 2016-08-04 Eder Santana , George Hotz

The automated analysis of human behaviour provides many opportunities for the creation of interactive systems and the post-experiment investigations for user studies. Commodity depth cameras offer reasonable body tracking accuracy at a low…

人机交互 · 计算机科学 2025-06-25 Adrien Coppens , Valérie Maquil

This research focuses on real-time monitoring and analysis of track and field athletes, addressing the limitations of traditional monitoring systems in terms of real-time performance and accuracy. We propose an IoT-optimized system that…

机器学习 · 计算机科学 2024-11-12 Xiaowei Tang , Bin Long , Li Zhou

Fully autonomous driving has been widely studied and is becoming increasingly feasible. However, such autonomous driving has yet to be achieved on public roads, because of various uncertainties due to surrounding human drivers and…

机器人学 · 计算机科学 2023-05-19 Shunsuke Aoki , Issei Yamamoto , Daiki Shiotsuka , Yuichi Inoue , Kento Tokuhiro , Keita Miwa

Deep Learning (DL) has become a crucial technology for Artificial Intelligence (AI). It is a powerful technique to automatically extract high-level features from complex data which can be exploited for applications such as computer vision,…

计算机视觉与模式识别 · 计算机科学 2019-06-10 Gael Kamdem De Teyou

Current deep learning based autonomous driving approaches yield impressive results also leading to in-production deployment in certain controlled scenarios. One of the most popular and fascinating approaches relies on learning vehicle…

计算机视觉与模式识别 · 计算机科学 2020-06-08 Luca Cultrera , Lorenzo Seidenari , Federico Becattini , Pietro Pala , Alberto Del Bimbo

With the rising popularity of autonomous navigation research, Formula Student (FS) events are introducing a Driverless Vehicle (DV) category to their event list. This paper presents the initial investigation into utilising Deep…

机器人学 · 计算机科学 2023-08-28 Aakaash Salvaji , Harry Taylor , David Valencia , Trevor Gee , Henry Williams