English
Related papers

Related papers: Privacy-Preserving End-to-End Full-Duplex Speech D…

200 papers

In this paper, we consider a two-way wiretap Multi-Input Multi-Output Multi-antenna Eve (MIMOME) channel, where both nodes (Alice and Bob) transmit and receive in an in-band full-duplex (IBFD) manner. For this system with keyless security,…

Information Theory · Computer Science 2024-03-12 Navneet Garg , Haifeng Luo , Tharmalingam Ratnarajah

We investigate the physical layer security of wireless single-input single-output orthogonal-division multiplexing (OFDM) when a transmitter, which we refer to as Alice, sends her information to a receiver, which we refer to as Bob, in the…

Information Theory · Computer Science 2018-03-02 Mohamed F. Marzban , Ahmed El Shafie , Rakan Chabaan , Naofal Al-Dhahir

Speaker segmentation consists in partitioning a conversation between one or more speakers into speaker turns. Usually addressed as the late combination of three sub-tasks (voice activity detection, speaker change detection, and overlapped…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-11 Hervé Bredin , Antoine Laurent

Recordings in everyday life require privacy preservation of the speech content and speaker identity. This contribution explores the influence of noise and reverberation on the trade-off between privacy and utility for low-cost…

Audio and Speech Processing · Electrical Eng. & Systems 2026-02-04 Jule Pohlhausen , Francesco Nespoli , Joerg Bitzer

This paper proposes an end-to-end approach for single-channel speaker-independent multi-speaker speech separation, where time-frequency (T-F) masking, the short-time Fourier transform (STFT), and its inverse are represented as layers within…

Sound · Computer Science 2018-04-30 Zhong-Qiu Wang , Jonathan Le Roux , DeLiang Wang , John R. Hershey

The abundance of data collected by sensors in Internet of Things (IoT) devices, and the success of deep neural networks in uncovering hidden patterns in time series data have led to mounting privacy concerns. This is because private and…

Machine Learning · Computer Science 2022-06-02 Omid Hajihassani , Omid Ardakanian , Hamzeh Khazaei

Speech signals contain a lot of sensitive information, such as the speaker's identity, which raises privacy concerns when speech data get collected. Speaker anonymization aims to transform a speech signal to remove the source speaker's…

Sound · Computer Science 2023-01-16 Pierre Champion , Denis Jouvet , Anthony Larcher

As users increasingly rely on cloud-based computing services, it is important to ensure that uploaded speech data remains private. Existing solutions rely either on server-side methods or focus on hiding speaker identity. While these…

Audio and Speech Processing · Electrical Eng. & Systems 2021-10-26 Peter Wu , Paul Pu Liang , Jiatong Shi , Ruslan Salakhutdinov , Shinji Watanabe , Louis-Philippe Morency

The recently proposed x-vector based anonymization scheme converts any input voice into that of a random pseudo-speaker. In this paper, we present a flexible pseudo-speaker selection technique as a baseline for the first VoicePrivacy…

Audio and Speech Processing · Electrical Eng. & Systems 2020-05-19 Brij Mohan Lal Srivastava , Natalia Tomashenko , Xin Wang , Emmanuel Vincent , Junichi Yamagishi , Mohamed Maouche , Aurélien Bellet , Marc Tommasi

Deep neural networks are inherently opaque and challenging to interpret. Unlike hand-crafted feature-based models, we struggle to comprehend the concepts learned and how they interact within these models. This understanding is crucial not…

Computation and Language · Computer Science 2023-07-12 Shammur Absar Chowdhury , Nadir Durrani , Ahmed Ali

We assume a full-duplex (FD) cooperative network subject to hostile attacks and undergoing composite fading channels. We focus on two scenarios: \textit{a)} the transmitter has full CSI, for which we derive closed-form expressions for the…

Information Theory · Computer Science 2015-06-23 Hirley Alves , Glauber Brante , Richard D. Souza , Daniel B. da Costa , Matti Latva-aho

Target confusion, defined as occasional switching to non-target speakers, poses a key challenge for end-to-end speaker extraction (E2E-SE) systems. We argue that this problem is largely caused by the lack of generalizability and…

Sound · Computer Science 2025-05-29 Zhenghai You , Zhenyu Zhou , Lantian Li , Dong Wang

In this paper, we investigate the physical layer security of a full-duplex base station (BS) aided system in the worst case, where an uplink transmitter (UT) and a downlink receiver (DR) are both equipped with a single antenna, while a…

Information Theory · Computer Science 2019-04-23 Zhengmin Kong , Shaoshi Yang , Die Wang , Lajos Hanzo

In this paper, the physical layer security of a dual-hop underlay uplink cognitive radio network is investigated over Nakagami-m fading channels. Specifically, multiple secondary sources are taking turns in accessing the licensed spectrum…

Information Theory · Computer Science 2019-12-02 Mounia Bouabdellah , Faissal El Bouanani , Mohamed-Slim Alouini

The popularity and projected growth of in-home smart speaker assistants, such as Amazon's Echo, has raised privacy concerns among consumers and privacy advocates. Notable questions regarding the collection and storage of user data by…

Cryptography and Security · Computer Science 2020-07-21 Ilesanmi Olade , Christopher Champion , Haining Liang , Charles Fleming

Albeit recent progress in speaker verification generates powerful models, malicious attacks in the form of spoofed speech, are generally not coped with. Recent results in ASVSpoof2015 and BTAS2016 challenges indicate that spoof-aware…

Audio and Speech Processing · Electrical Eng. & Systems 2020-07-28 Heinrich Dinkel , Nanxin Chen , Yanmin Qian , Kai Yu

The degrees of freedom (DoF) of the two-user Gaussian multiple-input and multiple-output (MIMO) broadcast channel with confidential message (BCC) is studied under the assumption that delayed channel state information (CSI) is available at…

Information Theory · Computer Science 2011-12-13 Sheng Yang , Mari Kobayashi , Pablo Piantanida , Shlomo Shamai

Large language models (LLMs) have demonstrated exceptional capabilities in text understanding and generation, and they are increasingly being utilized across various domains to enhance productivity. However, due to the high costs of…

Cryptography and Security · Computer Science 2024-11-05 Yu Mao , Xueping Liao , Wei Liu , Anjia Yang

While Audio Large Language Models (ALLMs) have achieved remarkable progress in understanding and generation, their potential privacy implications remain largely unexplored. This paper takes the first step to investigate whether ALLMs…

Computation and Language · Computer Science 2026-01-08 Jin Wang , Liang Lin , Kaiwen Luo , Weiliu Wang , Yitian Chen , Moayad Aloqaily , Xuehai Tang , Zhenhong Zhou , Kun Wang , Li Sun , Qingsong Wen
‹ Prev 1 8 9 10 Next ›