English
Related papers

Related papers: Long-term Multi-granularity Deep Framework for Dri…

200 papers

Dynamic facial expression recognition (DFER) in the wild is an extremely challenging task, due to a large number of noisy frames in the video sequences. Previous works focus on extracting more discriminative features, but ignore…

Computer Vision and Pattern Recognition · Computer Science 2022-06-13 Hanting Li , Mingzhe Sui , Zhaoqing Zhu , Feng zhao

Drones are increasingly used in fields like industry, medicine, research, disaster relief, defense, and security. Technical challenges, such as navigation in GPS-denied environments, hinder further adoption. Research in visual odometry is…

Robotics · Computer Science 2024-04-30 Olivier Brochu Dufour , Abolfazl Mohebbi , Sofiane Achiche

A driver face monitoring system can detect driver fatigue, which is a significant factor in many accidents, using computer vision techniques. In this paper, we present a real-time technique for driver eye state detection. First, the face is…

Computer Vision and Pattern Recognition · Computer Science 2025-04-22 Deepak Ghimire , Sunghwan Jeong , Sunhong Yoon , Sanghyun Park , Juhwan Choi

Inspired by the recent development of deep network-based methods in semantic image segmentation, we introduce an end-to-end trainable model for face mask extraction in video sequence. Comparing to landmark-based sparse face shape…

Computer Vision and Pattern Recognition · Computer Science 2021-03-02 Yujiang Wang , Bingnan Luo , Jie Shen , Maja Pantic

Avoiding bottleneck situations in crowds is critical for the safety and comfort of people at large events or in public transportation. Based on the work of Lagrangian motion analysis we propose a novel video-based bottleneckdetector by…

Computer Vision and Pattern Recognition · Computer Science 2019-08-22 Maik Simon , Markus Küchhold , Tobias Senst , Erik Bochinski , Thomas Sikora

Accurate classification of autonomous vehicle (AV) driving behaviors is critical for safety validation, performance diagnosis, and traffic integration analysis. However, existing approaches primarily rely on numerical time-series modeling…

Artificial Intelligence · Computer Science 2026-03-04 Xiangyu Li , Tianyi Wang , Xi Cheng , Rakesh Chowdary Machineni , Zhaomiao Guo , Sikai Chen , Junfeng Jiao , Christian Claudel

Vehicle color recognition plays an important role in intelligent traffic management and criminal investigation assistance. However, the current vehicle color recognition research involves at most 13 types of colors and the recognition…

Image and Video Processing · Electrical Eng. & Systems 2021-07-22 Hu Ming-Di , Bai Long , Li Ying , Zhao Si-Rui , Chen En-Hong

In video compression, most of the existing deep learning approaches concentrate on the visual quality of a single frame, while ignoring the useful priors as well as the temporal information of adjacent frames. In this paper, we propose a…

Computer Vision and Pattern Recognition · Computer Science 2019-01-16 Xiandong Meng , Xuan Deng , Shuyuan Zhu , Shuaicheng Liu , Chuan Wang , Chen Chen , Bing Zeng

Since the number of cars has grown rapidly in recent years, driving safety draws more and more public attention. Drowsy driving is one of the biggest threatens to driving safety. Therefore, a simple but robust system that can detect drowsy…

Sound · Computer Science 2025-04-01 Yadong Xie , Fan Li , Yue Wu , Song Yang , Yu Wang

Moment retrieval in videos is a challenging task that aims to retrieve the most relevant video moment in an untrimmed video given a sentence description. Previous methods tend to perform self-modal learning and cross-modal interaction in a…

Computer Vision and Pattern Recognition · Computer Science 2023-02-21 Xin Sun , Xuan Wang , Jialin Gao , Qiong Liu , Xi Zhou

The rise of deepfake technology brings forth new questions about the authenticity of various forms of media found online today. Videos and images generated by artificial intelligence (AI) have become increasingly more difficult to…

Computer Vision and Pattern Recognition · Computer Science 2025-07-28 Benjamin Carter , Nathan Dilla , Micheal Callahan , Atuhaire Ambala

Deepfakes is a branch of malicious techniques that transplant a target face to the original one in videos, resulting in serious problems such as infringement of copyright, confusion of information, or even public panic. Previous efforts for…

Computer Vision and Pattern Recognition · Computer Science 2021-04-12 Zekun Sun , Yujie Han , Zeyu Hua , Na Ruan , Weijia Jia

Video temporal grounding aims to pinpoint a video segment that matches the query description. Despite the recent advance in short-form videos (\textit{e.g.}, in minutes), temporal grounding in long videos (\textit{e.g.}, in hours) is still…

Computer Vision and Pattern Recognition · Computer Science 2024-02-20 Yulin Pan , Xiangteng He , Biao Gong , Yiliang Lv , Yujun Shen , Yuxin Peng , Deli Zhao

Mobile vision systems such as smartphones, drones, and augmented-reality headsets are revolutionizing our lives. These systems usually run multiple applications concurrently and their available resources at runtime are dynamic due to events…

Computer Vision and Pattern Recognition · Computer Science 2018-10-25 Biyi Fang , Xiao Zeng , Mi Zhang

Detecting driver fatigue is critical for road safety, as drowsy driving remains a leading cause of traffic accidents. Many existing solutions rely on computationally demanding deep learning models, which result in high latency and are…

Computer Vision and Pattern Recognition · Computer Science 2025-08-14 Jing Ren , Suyu Ma , Hong Jia , Xiwei Xu , Ivan Lee , Haytham Fayek , Xiaodong Li , Feng Xia

In this paper, we present a new dataset for "distracted driver" posture estimation. In addition, we propose a novel system that achieves 95.98% driving posture estimation classification accuracy. The system consists of a…

Computer Vision and Pattern Recognition · Computer Science 2018-12-03 Yehya Abouelnaga , Hesham M. Eraqi , Mohamed N. Moustafa

This paper proposes a DNN-based system that detects multiple people from a single depth image. Our neural network processes a depth image and outputs a likelihood map in image coordinates, where each detection corresponds to a…

Computer Vision and Pattern Recognition · Computer Science 2020-07-15 David Fuentes-Jimenez , Cristina Losada-Gutierrez , David Casillas-Perez , Javier Macias-Guarasa , Roberto Martin-Lopez , Daniel Pizarro , Carlos A. Luna

The efficacy of video generation models heavily depends on the quality of their training datasets. Most previous video generation models are trained on short video clips, while recently there has been increasing interest in training long…

Computer Vision and Pattern Recognition · Computer Science 2024-10-15 Tianwei Xiong , Yuqing Wang , Daquan Zhou , Zhijie Lin , Jiashi Feng , Xihui Liu

Talking head video generation aims to animate a human face in a still image with dynamic poses and expressions using motion information derived from a target-driving video, while maintaining the person's identity in the source image.…

Computer Vision and Pattern Recognition · Computer Science 2023-08-21 Fa-Ting Hong , Dan Xu

Buses and heavy vehicles have more blind spots compared to cars and other road vehicles due to their large sizes. Therefore, accidents caused by these heavy vehicles are more fatal and result in severe injuries to other road users. These…

Computer Vision and Pattern Recognition · Computer Science 2022-08-22 Muhammad Muzammel , Mohd Zuki Yusoff , Mohamad Naufal Mohamad Saad , Faryal Sheikh , Muhammad Ahsan Awais