中文
相关论文

相关论文: A speech-based driver assisting module for Intelli…

200 篇论文

In this project, we aim to build a Text-to-Speech system able to produce speech with a controllable emotional expressiveness. We propose a methodology for solving this problem in three main steps. The first is the collection of emotional…

音频与语音处理 · 电气工程与系统科学 2019-07-08 Noé Tits

Text-based video segmentation aims to segment the target object in a video based on a describing sentence. Incorporating motion information from optical flow maps with appearance and linguistic modalities is crucial yet has been largely…

计算机视觉与模式识别 · 计算机科学 2022-04-07 Wangbo Zhao , Kai Wang , Xiangxiang Chu , Fuzhao Xue , Xinchao Wang , Yang You

Interaction-aware Autonomous Driving (IAAD) is a rapidly growing field of research that focuses on the development of autonomous vehicles (AVs) that are capable of interacting safely and efficiently with human road users. This is a…

人机交互 · 计算机科学 2023-11-01 Luca Crosato , Kai Tian , Hubert P. H Shum , Edmond S. L. Ho , Yafei Wang , Chongfeng Wei

In autonomous driving tasks, scene understanding is the first step towards predicting the future behavior of the surrounding traffic participants. Yet, how to represent a given scene and extract its features are still open research…

计算机视觉与模式识别 · 计算机科学 2023-09-14 Ali Keysan , Andreas Look , Eitan Kosman , Gonca Gürsun , Jörg Wagner , Yu Yao , Barbara Rakitsch

Expected to provide higher transportation efficiency and security, autonomous driving has attracted substantial attentions from both industry and academia. Meanwhile, the emergence of edge intelligence has further introduced significant…

信号处理 · 电气工程与系统科学 2024-09-25 Yunqi Feng , Hesheng Shen , Zhendong Shan , Qianqian Yang , Xiufang Shi

Traffic signal control is an important and challenging real-world problem, which aims to minimize the travel time of vehicles by coordinating their movements at the road intersections. Current traffic signal control systems in use still…

机器学习 · 计算机科学 2020-01-17 Hua Wei , Guanjie Zheng , Vikash Gayah , Zhenhui Li

Learning contextual and spatial environmental representations enhances autonomous vehicle's hazard anticipation and decision-making in complex scenarios. Recent perception systems enhance spatial understanding with sensor fusion but often…

机器人学 · 计算机科学 2024-01-18 Shoaib Azam , Farzeen Munir , Ville Kyrki , Moongu Jeon , Witold Pedrycz

We present Flex, an efficient and effective scene encoder that addresses the computational bottleneck of processing high-volume multi-camera data in end-to-end autonomous driving. Flex employs a small set of learnable scene tokens to…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Jiawei Yang , Ziyu Chen , Yurong You , Yan Wang , Yiming Li , Yuxiao Chen , Boyi Li , Boris Ivanovic , Marco Pavone , Yue Wang

A scene text spotter is composed of text detection and recognition modules. Many studies have been conducted to unify these modules into an end-to-end trainable model to achieve better performance. A typical architecture places detection…

计算机视觉与模式识别 · 计算机科学 2020-07-21 Youngmin Baek , Seung Shin , Jeonghun Baek , Sungrae Park , Junyeop Lee , Daehyun Nam , Hwalsuk Lee

Recent advances in machine learning, particularly deep learning, have enabled autonomous systems to perceive and comprehend objects and their environments in a perceptual subsymbolic manner. These systems can now perform object detection,…

人工智能 · 计算机科学 2023-09-13 Amr Gomaa , Michael Feld

The Bluetooth protocol can be used for intervehicle communication equipped with Bluetooth devices. This work investigates the challenges and feasibility of developing intelligent driving system providing timesensitive information about…

网络与互联网体系结构 · 计算机科学 2014-05-27 Nevin Vunka Jungum , Razvi M. Doomun , Soulakshmee D. Ghurbhurrun , Sameerchand Pudaruth

This software project based paper is for a vision of the near future in which computer interaction is characterized by natural face-to-face conversations with lifelike characters that speak, emote, and gesture. The first step is speech. The…

人机交互 · 计算机科学 2013-05-10 Urmila Shrawankar , Anjali Mahajan

The aim of this paper is to develop a flexible framework capable of automatically recognizing phonetic units present in a speech utterance of any language spoken in any mode. In this study, we considered two modes of speech: conversation,…

音频与语音处理 · 电气工程与系统科学 2019-08-27 Kumud Tripathi , M. Kiran Reddy , K. Sreenivasa Rao

The establishment of fast and reliable communication technologies, such as 5G, is enabling the evolution of a new generation of connected ADAS. This work aims to develop a traffic light advisory system, Multiple Traffic Light Advisor…

系统与控制 · 电气工程与系统科学 2023-01-10 Michael Khayyat , Alberto Gabriele , Francesca Mancini , Stefano Arrigoni , Francesco Braghin

Vision-language models enable the understanding and reasoning of complex traffic scenarios through multi-source information fusion, establishing it as a core technology for autonomous driving. However, existing vision-language models are…

计算机视觉与模式识别 · 计算机科学 2025-12-17 Minghui Hou , Wei-Hsing Huang , Shaofeng Liang , Daizong Liu , Tai-Hao Wen , Gang Wang , Runwei Guan , Weiping Ding

While several datasets for autonomous navigation have become available in recent years, they tend to focus on structured driving environments. This usually corresponds to well-delineated infrastructure such as lanes, a small number of…

计算机视觉与模式识别 · 计算机科学 2018-11-27 Girish Varma , Anbumani Subramanian , Anoop Namboodiri , Manmohan Chandraker , C V Jawahar

Autonomous driving presents a complex challenge, which is usually addressed with artificial intelligence models that are end-to-end or modular in nature. Within the landscape of modular approaches, a bio-inspired neural circuit policy model…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Anass Bairouk , Mirjana Maras , Simon Herlin , Alexander Amini , Marc Blanchon , Ramin Hasani , Patrick Chareyre , Daniela Rus

Vehicle-to-vehicle communications can change the driving behavior of drivers significantly by providing them rich information on downstream traffic flow conditions. This study seeks to model the varying car-following behaviors involving…

系统与控制 · 计算机科学 2018-09-18 Lin Liu , Chunyuan Li , Yongfu Li , Srinivas Peeta , Lei Lin

Audio-Visual Scene-Aware Dialog (AVSD) is a task to generate responses when chatting about a given video, which is organized as a track of the 8th Dialog System Technology Challenge (DSTC8). To solve the task, we propose a universal…

计算与语言 · 计算机科学 2020-02-04 Zekang Li , Zongjia Li , Jinchao Zhang , Yang Feng , Cheng Niu , Jie Zhou

Traffic signal control has long been considered as a critical topic in intelligent transportation systems. Most existing learning methods mainly focus on isolated intersections and suffer from inefficient training. This paper aims at the…

机器学习 · 计算机科学 2019-10-01 Yusen Huo , Qinghua Tao , Jianming Hu