English
Related papers

Related papers: Creating Aesthetic Sonifications on the Web with S…

200 papers

Apparatus and methods are disclosed for performing object-based audio rendering on a plurality of audio objects which define a sound scene, each audio object comprising at least one audio signal and associated metadata. The apparatus…

Sinusoidal neural networks (SIRENs) are powerful implicit neural representations (INRs) for low-dimensional signals in vision and graphics. By encoding input coordinates with sinusoidal functions, they enable high-frequency image and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Haoan Feng , Diana Aldana , Tiago Novello , Leila De Floriani

Augmented listening devices, such as hearing aids and augmented reality headsets, enhance human perception by changing the sounds that we hear. Microphone arrays can improve the performance of listening systems in noisy environments, but…

Audio and Speech Processing · Electrical Eng. & Systems 2020-04-28 Ryan M. Corey , Andrew C. Singer

The visualization of publication and citation data is popular in bibliometrics. Although less common, the representation of empirical data as sound is an alternative form of presentation (in other fields than bibliometrics). In this…

Digital Libraries · Computer Science 2024-06-11 Lutz Bornmann , Rouven Lazlo Haegner

Web agents struggle to adapt to new websites due to the scarcity of environment specific tasks and demonstrations. Recent works have explored synthetic data generation to address this challenge, however, they suffer from data quality issues…

With rapid advances in generative artificial intelligence, the text-to-music synthesis task has emerged as a promising direction for music generation. Nevertheless, achieving precise control over multi-track generation remains an open…

Sound · Computer Science 2024-12-18 Yao Yao , Peike Li , Boyu Chen , Alex Wang

Neural audio synthesis methods now allow specifying ideas in natural language. However, these methods produce results that cannot be easily tweaked, as they are based on large latent spaces and up to billions of uninterpretable parameters.…

Sound · Computer Science 2024-06-04 Manuel Cherep , Nikhil Singh , Jessica Shand

For people with noise sensitivity, everyday soundscapes can be overwhelming. Existing tools such as active noise cancellation reduce discomfort by suppressing the entire acoustic environment, often at the cost of awareness of surrounding…

Sound · Computer Science 2026-04-02 Jeremy Zhengqi Huang , Emani Hicks , Sidharth , Gillian R. Hayes , Dhruv Jain

After decades of research, sonification is still rarely adopted in consumer electronics, software and user interfaces. Outside the science and arts scenes the term sonification seems not well known to the public. As a means of science…

Human-Computer Interaction · Computer Science 2024-10-16 Tim Ziemer , Navid Mirzayousef Jadid

In the art of video editing, sound helps add character to an object and immerse the viewer within a space. Through formative interviews with professional editors (N=10), we found that the task of adding sounds to video can be challenging.…

Learning to program an audio production VST synthesizer is a time consuming process, usually obtained through inefficient trial and error and only mastered after years of experience. As an educational and creative tool for sound designers,…

Sound · Computer Science 2021-04-12 Christopher Mitcheltree , Hideki Koike

Web service composition is the process of synthesizing a new composite service using a set of available Web services in order to satisfy a client request that cannot be treated by any available Web services. The Web services space is a…

Artificial Intelligence · Computer Science 2013-05-02 Chantal Cherifi , Yvan Rivierre , Jean-Francois Santucci

In daily life, we encounter a variety of sounds, both desirable and undesirable, with limited control over their presence and volume. Our work introduces "Listen, Chat, and Remix" (LCR), a novel multimodal sound remixer that controls each…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-12 Xilin Jiang , Cong Han , Yinghao Aaron Li , Nima Mesgarani

The large size of nowadays' online multimedia databases makes retrieving their content a difficult and time-consuming task. Users of online sound collections typically submit search queries that express a broad intent, often making the…

Information Retrieval · Computer Science 2020-06-16 Xavier Favory , Frederic Font , Xavier Serra

This paper describes a system that generates speaker-annotated transcripts of meetings by using a microphone array and a 360-degree camera. The hallmark of the system is its ability to handle overlapped speech, which has been an unsolved…

In this work, we introduce VERSA, a unified and standardized evaluation toolkit designed for various speech, audio, and music signals. The toolkit features a Pythonic interface with flexible configuration and dependency control, making it…

Augmented reality (AR) has shown promise for supporting Deaf and hard-of-hearing (DHH) individuals by captioning speech and visualizing environmental sounds, yet existing systems do not allow users to create personalized sound…

Human-Computer Interaction · Computer Science 2025-11-25 Jaewook Lee , Davin Win Kyi , Leejun Kim , Jenny Peng , Gagyeom Lim , Jeremy Zhengqi Huang , Dhruv Jain , Jon E. Froehlich

Systematic evaluation of speech separation and enhancement models under moving sound source conditions requires extensive and diverse data. However, real-world datasets often lack sufficient data for training and evaluation, and synthetic…

Sound · Computer Science 2025-03-07 Kai Li , Wendi Sang , Chang Zeng , Runxuan Yang , Guo Chen , Xiaolin Hu

The inclusion of voice persona in synthesized voice can be significant in a broad range of human-computer-interaction (HCI) applications, including augmentative and assistive communication (AAC), artistic performance, and design of virtual…

Sound · Computer Science 2022-11-01 Camille Noufi , Lloyd May , Jonathan Berger

In the Western music tradition, chords are the main constituent components of harmony, a fundamental dimension of music. Despite its relevance for several Music Information Retrieval (MIR) tasks, chord-annotated audio datasets are limited…

Sound · Computer Science 2024-08-02 Andrea Poltronieri , Valentina Presutti , Martín Rocamora