English
Related papers

Related papers: Your Microphone Array Retains Your Identity: A Rob…

200 papers

Voice assistants like Amazon's Alexa, Google's Assistant, or Apple's Siri, have become the primary (voice) interface in smart speakers that can be found in millions of households. For privacy reasons, these speakers analyze every sound in…

Cryptography and Security · Computer Science 2020-08-04 Lea Schönherr , Maximilian Golla , Thorsten Eisenhofer , Jan Wiele , Dorothea Kolossa , Thorsten Holz

The rapid advancement of voice generation technologies has enabled the synthesis of speech that is perceptually indistinguishable from genuine human voices. While these innovations facilitate beneficial applications such as personalized…

Cryptography and Security · Computer Science 2025-07-30 Aditya Pujari , Ajita Rattani

Prevailing user authentication schemes on smartphones rely on explicit user interaction, where a user types in a passcode or presents a biometric cue such as face, fingerprint, or iris. In addition to being cumbersome and obtrusive to the…

Computer Vision and Pattern Recognition · Computer Science 2019-01-17 Debayan Deb , Arun Ross , Anil K. Jain , Kwaku Prakah-Asante , K. Venkatesh Prasad

We propose a system that gives a mobile robot the ability to separate simultaneous sound sources. A microphone array is used along with a real-time dedicated implementation of Geometric Source Separation and a post-filter that gives us a…

Robotics · Computer Science 2016-03-09 Jean-Marc Valin , Jean Rouat , François Michaud

The International Fingerprint Liveness Detection Competition is an international biennial competition open to academia and industry with the aim to assess and report advances in Fingerprint Presentation Attack Detection. The proposed…

Computer Vision and Pattern Recognition · Computer Science 2021-08-24 Roberto Casula , Marco Micheletto , Giulia Orrù , Rita Delussu , Sara Concas , Andrea Panzino , Gian Luca Marcialis

Replay speech attacks pose a significant threat to voice-controlled systems, especially in smart environments where voice assistants are widely deployed. While multi-channel audio offers spatial cues that can enhance replay detection…

Audio and Speech Processing · Electrical Eng. & Systems 2025-09-19 Michael Neri , Tuomas Virtanen

Parametric arrays (PA) offer exceptional directivity and compactness compared to conventional loudspeakers, facilitating various acoustic applications. However, accurate measurement of audio signals generated by PA remains challenging due…

Sound · Computer Science 2025-09-29 Woongji Kim , Beomseok Oh , Junsuk Rho , Wonkyu Moon

In recent years, fingerprint recognition systems have made remarkable advancements in the field of biometric security as it plays an important role in personal, national and global security. In spite of all these notable advancements, the…

Computer Vision and Pattern Recognition · Computer Science 2020-02-20 B. V. S Anusha , Sayan Banerjee , Subhasis Chaudhuri

Eavesdropping on voice conversations presents a growing threat to personal privacy and information security. In this paper, we present RadEar, a novel RF backscatter-based system designed to enable covert voice eavesdropping through walls.…

Networking and Internet Architecture · Computer Science 2026-03-16 Qijun Wang , Peihao Yan , Chunqi Qian , Huacheng Zeng

Artificial intelligence (AI) is increasingly being explored in health and social care to reduce administrative workload and allow staff to spend more time on patient care. This paper evaluates a voice-enabled Care Home Smart Speaker…

We present a cost-effective two-step authentication system that integrates face identification and speaker verification using only a camera and microphone available on common devices. The pipeline first performs face recognition to identify…

Computer Vision and Pattern Recognition · Computer Science 2026-01-13 Kuan Wei Chen , Ting Yi Lin , Wen Ren Yang , Aryan Kesarwani , Riya Singh

The development of privacy-preserving automatic speaker verification systems has been the focus of a number of studies with the intent of allowing users to authenticate themselves without risking the privacy of their voice. However, current…

Audio and Speech Processing · Electrical Eng. & Systems 2022-10-28 Francisco Teixeira , Alberto Abad , Bhiksha Raj , Isabel Trancoso

In this paper, we present a passive method to detect face presentation attack a.k.a face liveness detection using an ensemble deep learning technique. Face liveness detection is one of the key steps involved in user identity verification of…

Computer Vision and Pattern Recognition · Computer Science 2022-01-25 Shashank Shekhar , Avinash Patel , Mrinal Haloi , Asif Salim

This paper presents an experimental study on radio frequency (RF) fingerprinting of Bluetooth Classic devices. Our research aims to provide a practical evaluation of the possibilities for RF fingerprinting of everyday Bluetooth connected…

Signal Processing · Electrical Eng. & Systems 2024-02-12 Artis Rušiņš , Krišjānis Nesenbergs , Deniss Tiščenko , Pēteris Paikens

Successful active speaker detection requires a three-stage pipeline: (i) audio-visual encoding for all speakers in the clip, (ii) inter-speaker relation modeling between a reference speaker and the background speakers within each frame, and…

Computer Vision and Pattern Recognition · Computer Science 2021-09-08 Okan Köpüklü , Maja Taseska , Gerhard Rigoll

Voice Activity Detection (VAD) refers to the problem of distinguishing speech segments from background noise. Numerous approaches have been proposed for this purpose. Some are based on features derived from the power spectral density,…

Sound · Computer Science 2019-03-08 Thomas Drugman , Yannis Stylianou , Yusuke Kida , Masami Akamine

Recognition systems are commonly designed to authenticate users at the access control levels of a system. A number of voice recognition methods have been developed using a pitch estimation process which are very vulnerable in low Signal to…

Sound · Computer Science 2020-09-08 Aman Chadha , Divya Jyoti , M. Mani Roja

Recent breakthroughs in zero-shot voice synthesis have enabled imitating a speaker's voice using just a few seconds of recording while maintaining a high level of realism. Alongside its potential benefits, this powerful technology…

Sound · Computer Science 2024-01-09 Guangyu Chen , Yu Wu , Shujie Liu , Tao Liu , Xiaoyong Du , Furu Wei

Voice-based interfaces rely on a wake-up word mechanism to initiate communication with devices. However, achieving a robust, energy-efficient, and fast detection remains a challenge. This paper addresses these real production needs by…

Sound · Computer Science 2023-10-18 Fernando López , Jordi Luque , Carlos Segura , Pablo Gómez

Replay attacks remain a critical vulnerability for automatic speaker verification systems, particularly in real-time voice assistant applications. In this work, we propose acoustic maps as a novel spatial feature representation for replay…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-21 Michael Neri , Tuomas Virtanen