English
Related papers

Related papers: Sensing the Breath: A Multimodal Singing Tutoring …

200 papers

Breath with nose sound features has been shown as a potential biometric in personal identification and verification. In this paper, we show that information that comes from other modalities captured by motion sensors on the chest in…

Machine Learning · Computer Science 2021-11-01 Manh-Ha Bui , Viet-Anh Tran , Cuong Pham

Since the vocal component plays a crucial role in popular music, singing voice detection has been an active research topic in music information retrieval. Although several proposed algorithms have shown high performances, we argue that…

Sound · Computer Science 2018-06-05 Kyungyun Lee , Keunwoo Choi , Juhan Nam

Digitally presenting physiological signals as biofeedback to users raises awareness of both body and mind. This paper describes the effectiveness of conveying a physiological signal often overlooked for communication: breathing. We present…

Human-Computer Interaction · Computer Science 2018-06-25 Jérémy Frey , May Grabli , Ronit Slyper , Jessica Cauchard

This paper introduces a vibrotactile belt for interpersonal synchronization of breath. It can synchronize the breathing tempo of two people by transferring breathing rhythm of one user to vibration signals of another belt, where the depth…

Human-Computer Interaction · Computer Science 2024-11-11 Xilai Tan , Yan Zhang , Bin Zhao , Xiaolu Nan , Yuru Zhang , Dangxiao Wang

Developing Text-to-Speech (TTS) systems that can synthesize natural breath is essential for human-like voice agents but requires extensive manual annotation of breath positions in training data. To this end, we propose a self-training…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-17 Dong Yang , Tomoki Koriyama , Yuki Saito

We present a system that raises awareness about users' inner state. Di\v{s}imo is a multimodal ambient display that provides feedback about one's stress level, which is assessed through heart rate monitoring. Upon detecting a low heart rate…

Human-Computer Interaction · Computer Science 2018-06-24 Jelena Mladenovic , Jérémy Frey , Jessica Cauchard

Automatic Singing Assessment and Singing Information Processing have evolved over the past three decades to support singing pedagogy, performance analysis, and vocal training. While the first approach objectively evaluates a singer's…

Audio and Speech Processing · Electrical Eng. & Systems 2026-01-21 Arthur N. dos Santos , Bruno S. Masiero

Breathing is a spontaneous but controllable body function that can be used for hands-free interaction. Our work introduces "iBreath", a novel system to detect breathing gestures similar to clicks using bio-impedance. We evaluated iBreath's…

Human-Computer Interaction · Computer Science 2025-07-08 Mengxi Liu , Daniel Geißler , Deepika Gurung , Hymalai Bello , Bo Zhou , Sizhen Bian , Paul Lukowicz , Passant Elagroudy

Breathing is an essential part of human survival, which carries information about a person's physiological and psychological state. Generally, breath boundaries are marked by experts before using for any task. An unsupervised algorithm for…

Audio and Speech Processing · Electrical Eng. & Systems 2023-04-10 Shivani Yadav , Dipanjan Gope , Uma Maheswari K. , Prasanta Kumar Ghosh

Separating a song into vocal and accompaniment components is an active research topic, and recent years witnessed an increased performance from supervised training using deep learning techniques. We propose to apply the visual information…

Sound · Computer Science 2021-07-02 Bochen Li , Yuxuan Wang , Zhiyao Duan

Respiratory chest belt sensor can be used to measure the respiratory rate and other respiratory health parameters. Virtual Respiratory Belt, VRB, algorithms estimate the belt sensor waveform from speech audio. In this paper we compare the…

This paper presents an unobtrusive solution that can automatically identify deep breath when a person is walking past the global depth camera. Existing non-contact breath assessments achieve satisfactory results under restricted conditions…

Computer Vision and Pattern Recognition · Computer Science 2020-10-23 Yunlu Wang , Cheng Yang , Menghan Hu , Jian Zhang , Qingli Li , Guangtao Zhai , Xiao-Ping Zhang

Mindfulness training is widely recognized for its benefits in reducing depression, anxiety, and loneliness. With the rise of smartphone-based mindfulness apps, digital meditation has become more accessible, but sustaining long-term user…

Despite renewed awareness of the importance of articulation, it remains a challenge for instructors to handle the pronunciation needs of language learners. There are relatively scarce pedagogical tools for pronunciation teaching and…

Computer Vision and Pattern Recognition · Computer Science 2020-05-15 M. Hamed Mozaffari , Won-Sook Lee

Previous approaches in singer identification have used one of monophonic vocal tracks or mixed tracks containing multiple instruments, leaving a semantic gap between these two domains of audio. In this paper, we present a system to learn a…

Sound · Computer Science 2019-06-27 Kyungyun Lee , Juhan Nam

Singing voice beat and downbeat tracking posses several applications in automatic music production, analysis and manipulation. Among them, some require real-time processing, such as live performance processing and auto-accompaniment for…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-06 Mojtaba Heydari , Ju-Chiang Wang , Zhiyao Duan

Multimodal speech emotion recognition aims to detect speakers' emotions from audio and text. Prior works mainly focus on exploiting advanced networks to model and fuse different modality information to facilitate performance, while…

Computation and Language · Computer Science 2023-04-11 Zhen Wu , Yizhe Lu , Xinyu Dai

Tracking beats of singing voices without the presence of musical accompaniment can find many applications in music production, automatic song arrangement, and social media interaction. Its main challenge is the lack of strong rhythmic and…

Audio and Speech Processing · Electrical Eng. & Systems 2022-09-01 Mojtaba Heydari , Zhiyao Duan

Respiratory sound classification (RSC) is challenging due to varied acoustic signatures, primarily influenced by patient demographics and recording environments. To address this issue, we introduce a text-audio multimodal model that…

Sound · Computer Science 2024-06-17 June-Woo Kim , Miika Toikkanen , Yera Choi , Seoung-Eun Moon , Ho-Young Jung

Vocal training is difficult because the muscles that control pitch, resonance, and phonation are internal and invisible to learners. This paper investigates how Electromyography (EMG) and ultrasonic imaging (UI) can make these muscles…

Human-Computer Interaction · Computer Science 2026-03-23 Kanyu Chen , Rebecca Panskus , Erwin Wu , Yichen Peng , Daichi Saito , Emiko Kamiyama , Ruiteng Li , Chen-Chieh Liao , Karola Marky , Kato Akira , Hideki Koike , Kai Kunze
‹ Prev 1 2 3 10 Next ›