中文
相关论文

相关论文: Vehicular Visible Light Communications Noise Analy…

200 篇论文

The framework of variational autoencoders (VAEs) provides a principled method for jointly learning latent-variable models and corresponding inference models. However, the main drawback of this approach is the blurriness of the generated…

机器学习 · 计算机科学 2020-07-01 Ioannis Gatopoulos , Maarten Stol , Jakub M. Tomczak

Images captured from the real world are often affected by different types of noise, which can significantly impact the performance of Computer Vision systems and the quality of visual data. This study presents a novel approach for defect…

计算机视觉与模式识别 · 计算机科学 2024-05-14 Mohsen Hami , Mahdi JameBozorg

Humans are adept at leveraging visual cues from lip movements for recognizing speech in adverse listening conditions. Audio-Visual Speech Recognition (AVSR) models follow similar approach to achieve robust speech recognition in noisy…

音频与语音处理 · 电气工程与系统科学 2024-05-24 Maxime Burchi , Krishna C. Puvvada , Jagadeesh Balam , Boris Ginsburg , Radu Timofte

Audio-Visual Source Localization (AVSL) aims to localize the source of sound within a video. In this paper, we identify a significant issue in existing benchmarks: the sounding objects are often easily recognized based solely on visual…

多媒体 · 计算机科学 2024-09-12 Liangyu Chen , Zihao Yue , Boshen Xu , Qin Jin

The success of VLMs often relies on the dynamic high-resolution schema that adaptively augments the input images to multiple crops, so that the details of the images can be retained. However, such approaches result in a large number of…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Jiayi Han , Liang Du , Yiwen Wu , Xiangguo Zhou , Hongwei Du , Weibo Zheng

Real-world low-light images suffer from two main degradations, namely, inevitable noise and poor visibility. Since the noise exhibits different levels, its estimation has been implemented in recent works when enhancing low-light images from…

图像与视频处理 · 电气工程与系统科学 2021-10-08 Chuanjun Zheng , Daming Shi , Wentian Shi

Vehicle Ad-hoc Networks (VANETs) act as the core of vehicular communications and provide the fundamental wireless communication architecture to support both vehicle-to-vehicle (V2V) and vehicle-to-infrastructure (V2I) communication.…

网络与互联网体系结构 · 计算机科学 2022-09-15 Mao Ye , Nicolette Formosa , Mohammed Quddus

The visible light communication (VLC) by LED is one of the important communication methods because LED can work as high speed and VLC sends the information by high flushing LED. We use the pulse wave modulation for the VLC with LED because…

网络与互联网体系结构 · 计算机科学 2021-06-08 Wataru Uemura , Yasuhiro Fukumori , Takato Hayama

Variational encoder-decoders (VEDs) have shown promising results in dialogue generation. However, the latent variable distributions are usually approximated by a much simpler model than the powerful RNN structure used for encoding and…

计算与语言 · 计算机科学 2018-02-07 Xiaoyu Shen , Hui Su , Shuzi Niu , Vera Demberg

In this paper, an unmanned aerial vehicle (UAVs)-assisted visible light communication (VLC) has been considered which has two tiers: UAV-to-centroid and device-to-device (D2D). In the UAV-to-centroid tier, each UAV can simultaneously…

信号处理 · 电气工程与系统科学 2023-06-05 Alireza Qazavi , Foroogh S Tabataba , Mehdi Naderi Soorki

Previous visual object tracking methods employ image-feature regression models or coordinate autoregression models for bounding box prediction. Image-feature regression methods heavily depend on matching results and do not utilize…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Xinyu Zhou , Jinglun Li , Lingyi Hong , Kaixun Jiang , Pinxue Guo , Weifeng Ge , Wenqiang Zhang

This paper proposes a cascaded multiwire-power line communication (PLC)/multiple-visible light communication (VLC) system. This hybrid architecture offers low installation cost, enhanced performance, practical feasibility, and a wide range…

Recently, User-Generated Content (UGC) videos have gained popularity in our daily lives. However, UGC videos often suffer from poor exposure due to the limitations of photographic equipment and techniques. Therefore, Video Exposure…

计算机视觉与模式识别 · 计算机科学 2024-05-15 Xunchu Zhou , Xiaohong Liu , Yunlong Dong , Tengchuan Kou , Yixuan Gao , Zicheng Zhang , Chunyi Li , Haoning Wu , Guangtao Zhai

With the rapid development of communications and computing, the concept of connected vehicles emerges to improve driving safety, traffic efficiency and infotainment experience. Due to the limited capabilities of sensors and information…

信息论 · 计算机科学 2019-04-04 Haojun Yang , Kan Zheng , Kuan Zhang , Jie Mei , Yi Qian

We propose noise-robust voice conversion (VC) which takes into account the recording quality and environment of noisy source speech. Conventional denoising training improves the noise robustness of a VC model by learning noisy-to-clean VC…

Recent research in the design of end to end communication system using deep learning has produced models which can outperform traditional communication schemes. Most of these architectures leveraged autoencoders to design the encoder at the…

信息论 · 计算机科学 2020-01-28 Vishnu Raj , Sheetal Kalyani

Visible Light Communication (VLC) using light emitting diodes (LEDs) has been gaining increasing attention in recent years as it is appealing for a wide range of applications such as indoor positioning. Orthogonal frequency division…

信息论 · 计算机科学 2015-06-26 Mohammadreza Aminikashani , Wenjun Gu , Mohsen Kavehrad

Secure and reliable communications are crucial for Intelligent Transportation Systems (ITSs), where Vehicle-to-Infrastructure (V2I) communication plays a key role in enabling mobility-enhancing and safety-critical services. Current V2I…

In a conventional voice conversion (VC) framework, a VC model is often trained with a clean dataset consisting of speech data carefully recorded and selected by minimizing background interference. However, collecting such a high-quality…

声音 · 计算机科学 2021-09-23 Chao Xie , Yi-Chiao Wu , Patrick Lumban Tobing , Wen-Chin Huang , Tomoki Toda

We study the use of image-based Vision-Language Models (VLMs) for open-vocabulary segmentation of lidar scans in driving settings. Classically, image semantics can be back-projected onto 3D point clouds. Yet, resulting point labels are…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Nermin Samet , Gilles Puy , Renaud Marlet
‹ 上一页 1 8 9 10 下一页 ›