中文
相关论文

相关论文: An Enhanced Interleaving Frame Loss Concealment Me…

200 篇论文

Large Language Models (LLMs) can achieve near-optimal lossless compression by acting as powerful probability models. We investigate their use in the lossy domain, where reconstruction fidelity is traded for higher compression ratios. This…

机器学习 · 计算机科学 2025-10-28 Nnamdi Aghanya , Jun Li , Kewei Wang

Sending compressed video data in error-prone environments (like the Internet and wireless networks) might cause data degradation. Error concealment techniques try to conceal the received data in the decoder side. In this paper, an adaptive…

多媒体 · 计算机科学 2019-11-26 Seyed Mojtaba Marvasti-Zadeh , Hossein Ghanei-Yakhdan , Shohreh Kasaei

Despite recent advancements in packet loss concealment (PLC) using deep learning techniques, packet loss remains a significant challenge in real-time speech communication. Redundancy has been used in the past to recover the missing…

音频与语音处理 · 电气工程与系统科学 2024-10-28 Jean-Marc Valin , Jan Büthe , Ahmed Mustafa , Michael Klingbeil

Packet loss is a major cause of voice quality degradation in VoIP transmissions with serious impact on intelligibility and user experience. This paper describes a system based on a generative adversarial approach, which aims to repair the…

音频与语音处理 · 电气工程与系统科学 2023-07-31 Carlo Aironi , Samuele Cornell , Luca Serafini , Stefano Squartini

Efficient point cloud compression is essential for applications like virtual and mixed reality, autonomous driving, and cultural heritage. This paper proposes a deep learning-based inter-frame encoding scheme for dynamic point cloud…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Anique Akhtar , Zhu Li , Geert Van der Auwera

Ultra-reliable and low-latency communications (URLLC) partakes a major role in 5G networks for mission-critical applications. Sparse vector coding (SVC) appears as a strong candidate for future URLLC networks by enabling superior…

信号处理 · 电气工程与系统科学 2020-04-20 Emre Arslan , Ali Tugberk Dogukan , Ertugrul Basar

This paper investigates covert prompt transmission for secure and efficient large language model (LLM) services over wireless networks. We formulate a latency minimization problem under fidelity and detectability constraints to ensure…

网络与互联网体系结构 · 计算机科学 2025-05-01 Ruichen Zhang , Yinqiu Liu , Shunpu Tang , Jiacheng Wang , Dusit Niyato , Geng Sun , Yonghui Li , Sumei Sun

A novel inter-frame coding approach to the problem of varying channel-state conditions in broadcast wireless communication is developed in this paper; this problem causes the appropriate code-rate to vary across different transmitted frames…

信息论 · 计算机科学 2015-11-11 Hady Zeineddine , Mohammad M. Mansour

We propose a novel error concealment algorithm to be used at the receiver side of a lossy image transmission system. Our algorithm involves hiding the edge map of the original image at the transmitter within itself using a robust…

多媒体 · 计算机科学 2010-04-27 Shabnam Sodagari , Peyman Hesami , Alireza Nasiri Avanaki

While existing speech audio codecs designed for compression exploit limited forms of temporal redundancy and allow for multi-scale representations, they tend to represent all features of audio in the same way. In contrast, generative voice…

声音 · 计算机科学 2025-09-22 Ryan Collette , Ross Greenwood , Serena Nicoll

Audio packet loss is an inevitable problem in real-time speech communication. A band-split packet loss concealment network (BS-PLCNet) targeting full-band signals was recently proposed. Although it performs superiorly in the ICASSP 2024 PLC…

音频与语音处理 · 电气工程与系统科学 2024-06-11 Zihan Zhang , Xianjun Xia , Chuanzeng Huang , Yijian Xiao , Lei Xie

Exposing the rate information of wireless transmission enables highly efficient attacks that can severely degrade the performance of a network at very low cost. In this paper, we introduce an integrated solution to conceal the rate…

密码学与安全 · 计算机科学 2014-11-20 Triet D. Vo-Huu , Guevara Noubir

Redundant information of low-bit-rate speech is extremely small, thus it's very difficult to implement large capacity steganography on the low-bit-rate speech. Based on multiple vector quantization characteristics of the Line Spectrum Pair…

密码学与安全 · 计算机科学 2018-09-11 Zhongliang Yang , Xueshun Peng , Yongfeng Huang , Chinchen Chang

Discrete speech representations have garnered recent attention for their efficacy in training transformer-based models for various speech-related tasks such as automatic speech recognition (ASR), translation, speaker verification, and joint…

音频与语音处理 · 电气工程与系统科学 2024-09-26 Kunal Dhawan , Nithin Rao Koluguri , Ante Jukić , Ryan Langman , Jagadeesh Balam , Boris Ginsburg

This paper considers a source, which employs random linear coding (RLC) to encode a message, a legitimate destination, which can recover the message if it gathers a sufficient number of coded packets, and an eavesdropper. The probability of…

信息论 · 计算机科学 2022-03-24 Ioannis Chatzigeorgiou

If digital video data is transmitted over unreliable channels such as the internet or wireless terminals, the risk of severe image distortion due to transmission errors is ubiquitous. To cope with this, error concealment can be applied on…

图像与视频处理 · 电气工程与系统科学 2022-07-11 Jürgen Seiler , André Kaup

Dithering is a technique commonly used to improve the perceptual quality of lossy data compression. In this work, we analytically and experimentally justify the use of dithering for ASR input compression. We formalize an understanding of…

音频与语音处理 · 电气工程与系统科学 2025-12-12 Ellison Murray , Morriel Kasher , Predrag Spasojevic

In bandwidth-constrained communication such as satellite and underwater channels, speech must often be transmitted at ultra-low bitrates where intelligibility is the primary objective. At such extreme compression levels, codecs trained with…

声音 · 计算机科学 2026-04-21 Junyi Wang , Chi Zhang , Jing Qian , Haifeng Luo , Hao Wang , Zengrui Jin , Chao Zhang

Although the semantic communication with joint semantic-channel coding design has shown promising performance in transmitting data of different modalities over physical layer channels, the synchronization and packet-level forward error…

图像与视频处理 · 电气工程与系统科学 2024-08-13 Yun Tian , Jingkai Ying , Zhijin Qin , Ye Jin , Xiaoming Tao

Recent literature has shown that a learned front end with multi-channel audio input can outperform traditional beam-forming algorithms for automatic speech recognition (ASR). In this paper, we present our study on multi-channel acoustic…

声音 · 计算机科学 2020-02-04 Aparna Khare , Shiva Sundaram , Minhua Wu