English
Related papers

Related papers: End-to-end Evaluation of Practical Video Analytics…

200 papers

This paper presents a pioneering exploration into the integration of fine-grained human supervision within the autonomous driving domain to enhance system performance. The current advances in End-to-End autonomous driving normally are…

Robotics · Computer Science 2024-08-21 Yiqun Duan , Zhuoli Zhuang , Jinzhao Zhou , Yu-Cheng Chang , Yu-Kai Wang , Chin-Teng Lin

Lensless cameras, innovatively replacing traditional lenses for ultra-thin, flat optics, encode light directly onto sensors, producing images that are not immediately recognizable. This compact, lightweight, and cost-effective imaging…

Computer Vision and Pattern Recognition · Computer Science 2024-06-07 Xin Cai , Hailong Zhang , Chenchen Wang , Wentao Liu , Jinwei Gu , Tianfan Xue

Despite the increasing research interest in end-to-end learning systems for speech emotion recognition, conventional systems either suffer from the overfitting due in part to the limited training data, or do not explicitly consider the…

Computation and Language · Computer Science 2019-04-01 Zixing Zhang , Bingwen Wu , Bjoern Schuller

Although deep learning approaches have achieved performance surpassing humans for still image-based face recognition, unconstrained video-based face recognition is still a challenging task due to large volume of data to be processed and…

Computer Vision and Pattern Recognition · Computer Science 2019-08-13 Jingxiao Zheng , Rajeev Ranjan , Ching-Hui Chen , Jun-Cheng Chen , Carlos D. Castillo , Rama Chellappa

Video holds significance in computer graphics applications. Because of the heterogeneous of digital devices, retargeting videos becomes an essential function to enhance user viewing experience in such applications. In the research of video…

Computer Vision and Pattern Recognition · Computer Science 2023-11-10 Thi-Ngoc-Hanh Le , HuiGuang Huang , Yi-Ru Chen , Tong-Yee Lee

Most existing autonomous-driving datasets (e.g., KITTI, nuScenes, and the Waymo Perception Dataset), collected by human-driving mode or unidentified driving mode, can only serve as early training for the perception and prediction of…

Computer Vision and Pattern Recognition · Computer Science 2025-11-20 Xiangyu Li , Chen Wang , Yumao Liu , Dengbo He , Jiahao Zhang , Ke Ma

Dialog systems need to understand dynamic visual scenes in order to have conversations with users about the objects and events around them. Scene-aware dialog systems for real-world applications could be developed by integrating…

Future mobile networks supporting Internet of Things are expected to provide both high throughput and low latency to user-specific services. One way to overcome this challenge is to adopt Network Function Virtualization (NFV) and…

Networking and Internet Architecture · Computer Science 2019-06-26 Emmanouil Fountoulakis , Qi Liao , Nikolaos Pappas

Focusing on the task of point-to-point navigation for an autonomous driving vehicle, we propose a novel deep learning model trained with end-to-end and multi-task learning manners to perform both perception and control tasks simultaneously.…

Robotics · Computer Science 2022-06-23 Oskar Natan , Jun Miura

In the pathway toward Artificial General Intelligence (AGI), understanding human's affection is essential to enhance machine's cognition abilities. For achieving more sensual human-AI interaction, Multimodal Affective Computing (MAC) in…

Computer Vision and Pattern Recognition · Computer Science 2024-08-15 Ronghao Lin , Ying Zeng , Sijie Mai , Haifeng Hu

Spatial analysis can generate both exogenous and endogenous biases, which will lead to ethics issues. Exogenous biases arise from external factors or environments and are unrelated to internal operating mechanisms, while endogenous biases…

Human-Computer Interaction · Computer Science 2024-12-20 Chuan Chen , Peng Luo , Bo Zhao , Yu Feng , Liqiu Meng

We propose an end-to-end driving model that integrates a multi-task UNet (MTUNet) architecture and control algorithms in a pipeline of data flow from a front camera through this model to driving decisions. It provides quantitative measures…

Machine Learning · Computer Science 2023-09-11 Der-Hau Lee , Jinn-Liang Liu

Modern vehicles equip dashcams that primarily collect visual evidence for traffic accidents. However, most of the video data collected by dashcams that is not related to traffic accidents is discarded without any use. In this paper, we…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-12-02 Seyul Lee , Jayden King , Young Choon Lee , Hyuck Han , Sooyong Kang

Traditionally, audio-visual automatic speech recognition has been studied under the assumption that the speaking face on the visual signal is the face matching the audio. However, in a more realistic setting, when multiple faces are…

Audio and Speech Processing · Electrical Eng. & Systems 2022-05-12 Otavio Braga , Takaki Makino , Olivier Siohan , Hank Liao

We propose an experimental method for measuring bias in face recognition systems. Existing methods to measure bias depend on benchmark datasets that are collected in the wild and annotated for protected (e.g., race, gender) and…

Computer Vision and Pattern Recognition · Computer Science 2023-08-11 Hao Liang , Pietro Perona , Guha Balakrishnan

Over many decades, researchers working in object recognition have longed for an end-to-end automated system that will simply accept 2D or 3D image or videos as inputs and output the labels of objects in the input data. Computer vision…

Computer Vision and Pattern Recognition · Computer Science 2016-01-29 Rama Chellappa , Jun-Cheng Chen , Rajeev Ranjan , Swami Sankaranarayanan , Amit Kumar , Vishal M. Patel , Carlos D. Castillo

Detecting digital face manipulation in images and video has attracted extensive attention due to the potential risk to public trust. To counteract the malicious usage of such techniques, deep learning-based deepfake detection methods have…

Computer Vision and Pattern Recognition · Computer Science 2023-04-14 Yuhang Lu , Touradj Ebrahimi

Facial analysis systems have been deployed by large companies and critiqued by scholars and activists for the past decade. Many existing algorithmic audits examine the performance of these systems on later stage elements of facial analysis…

Computers and Society · Computer Science 2022-11-30 Samuel Dooley , George Z. Wei , Tom Goldstein , John P. Dickerson

Visual recognition systems mounted on autonomous moving agents face the challenge of unconstrained data, but simultaneously have the opportunity to improve their performance by moving to acquire new views of test data. In this work, we…

Computer Vision and Pattern Recognition · Computer Science 2016-08-09 Dinesh Jayaraman , Kristen Grauman

End-to-end autonomous driving provides a feasible way to automatically maximize overall driving system performance by directly mapping the raw pixels from a front-facing camera to control signals. Recent advanced methods construct a latent…

Machine Learning · Computer Science 2024-05-21 Zeyu Gao , Yao Mu , Chen Chen , Jingliang Duan , Shengbo Eben Li , Ping Luo , Yanfeng Lu