Related papers: The Tibetan Singing Bowl
Existing singing voice synthesis (SVS) models largely rely on fine-grained, phoneme-level durations, which limits their practical application. These methods overlook the complementary role of visual information in duration prediction.To…
Previous observations of the nontransient oscillations of rising bubbles and falling spheres in wormlike micellar fluids were limited to a single surfactant system. We present an extensive survey of rising bubbles in another system, an…
Scholars in the humanities rely heavily on ancient manuscripts to study history, religion, and socio-political structures in the past. Many efforts have been devoted to digitizing these precious manuscripts using OCR technology, but most…
Recent advancements in generative models have significantly enhanced talking face video generation, yet singing video generation remains underexplored. The differences between human talking and singing limit the performance of existing…
Our fluid dynamics video shows the response of a layer of viscoelastic fluid to an array of four-roll mills steadily rotating underneath. When the relaxation time of the fluid is sufficiently long, the fluid divides into "cells" with a…
This paper proposes a controllable singing voice synthesis system capable of generating expressive singing voice with two novel methodologies. First, a local style token module, which predicts frame-level style tokens from an input pitch…
This paper introduces GlOttal-flow LPC Filter (GOLF), a novel method for singing voice synthesis (SVS) that exploits the physical characteristics of the human voice using differentiable digital signal processing. GOLF employs a glottal…
Wood is one of the main materials used for making musical instruments due to its outstanding acoustical properties. Despite such unique properties, its inferior mechanical properties, moisture sensitivity, and time- and cost-consuming…
In this paper, we address the problem of lip-voice synchronisation in videos containing human face and voice. Our approach is based on determining if the lips motion and the voice in a video are synchronised or not, depending on their…
The rendering of Sanskrit poetry from text to speech is a problem that has not been solved before. One reason may be the complications in the language itself. We present unique algorithms based on extensive empirical analysis, to synthesize…
We present FlueBricks, a construction kit for acoustic reasoning via building and customizing flute-like instruments. By assembling generator, resonator, and connector modules that embody various aeroacoustic properties, users gain deeper…
Sound can complement vision in ball sports by providing subtle cues about contact dynamics. In table tennis, the brief, high-frequency sounds produced during racket-ball impacts carry information about the racket type, the surface…
Social interaction begins with the other person's attention, but it is difficult for a d/Deaf or hard-of-hearing (DHH) person to notice the initial conversation cues. Wearable or visual devices have been proposed previously. However, these…
Humpback whales can generate intricate bubbly regions, called bubble nets, via their blowholes. They appear to exploit these bubble nets for feeding via loud vocalizations. A fully-coupled phase-averaging approach is used to model the flow,…
An air bubble trapped in water by an oscillating acoustic field undergoes either radial or nonspherical pulsations depending on the strength of the forcing pressure. Two different instability mechanisms (the Rayleigh--Taylor instability and…
In this paper, we propose a model which can generate a singing voice from normal speech utterance by harnessing zero-shot, many-to-many style transfer learning. Our goal is to give anyone the opportunity to sing any song in a timely manner.…
The source of the hummingbirds distinctive hum is not well understood, but there are clues to its origin in the acoustic nearfield and farfield. To unravel this mystery, we recorded the acoustic nearfield generated by six freely hovering…
The classical elastic mechanics shows that the fundamental frequency of a sand grain chain is similar to the typical frequency of acoustic emission generated by the booming dunes. The "song of dunes" is therefore considered to originate…
We study the behavior of cylindrical objects as they sink into a dry granular bed fluidized due to lateral oscillations, in order to shed light on human constructions and other objects. Somewhat unexpectedly, we have found that, within a…
Cicadas (Homoptera:Cicadidae) are insects able to produce loudly songs and it is known that the mechanism to produce sound of tymballing cicadas works as a Helmholtz resonator. In this work we offer evidence on the participation of the…