English
Related papers

Related papers: A comparative study of two-dimensional vocal tract…

200 papers

Accurate modeling of the vocal tract is necessary to construct articulatory representations for interpretable speech processing and linguistics. However, vocal tract modeling is challenging because many internal articulators are occluded…

Computer Vision and Pattern Recognition · Computer Science 2024-06-25 Rishi Jain , Bohan Yu , Peter Wu , Tejas Prabhune , Gopala Anumanchipalli

The Finite Difference Time Domain (FDTD) method is a widely used numerical technique for solving Maxwell's equations, particularly in computational electromagnetics and photonics. It enables accurate modeling of wave propagation in complex…

Computation and Language · Computer Science 2025-04-15 Yifei He , Måns I. Andersson , Stefano Markidis

The focus of this article is on shape and topology optimization of transient vibroacoustic problems. The main contribution is a transient problem formulation that enables optimization over wide ranges of frequencies with complex signals,…

Optimization and Control · Mathematics 2023-06-28 Cetin B. Dilgen , Niels Aage

Radial Basis Function-generated Finite Differences (RBF-FD) is a popular variant of local strong-form meshless methods that do not require a predefined connection between the nodes, making it easier to adapt node-distribution to the problem…

Computational Engineering, Finance, and Science · Computer Science 2021-06-01 Jure Močnik - Berljavac , Pankaj K Mishra , Jure Slak , Gregor Kosec

This paper aims to present an innovative and cost-effective design for Acoustic Soft Tactile (AST) Skin, with the primary goal of significantly enhancing the accuracy of 2-D tactile feature estimation. The existing challenge lies in…

Robotics · Computer Science 2024-03-01 Vishnu Rajendran , Simon Parsons , Amir Ghalamzan E

The Finite-Difference Time-Domain (FDTD) method is a well-known technique for the analysis of quantum devices. It solves a discretized Schrodinger equation in an explicitly iterative process. However, the method requires the spatial grid…

Quantum Physics · Physics 2012-12-05 Frederick Ira Moxley , Weizhong Dai

The performance of speech enhancement and separation systems in anechoic environments has been significantly advanced with the recent progress in end-to-end neural network architectures. However, the performance of such systems in…

Audio and Speech Processing · Electrical Eng. & Systems 2020-11-17 Yi Luo , Cong Han , Nima Mesgarani

We present a finite difference time domain (FDTD) model for computation of A line scans in time domain optical coherence tomography (OCT). By simulating only the end of the two arms of the interferometer and computing the interference…

Medical Physics · Physics 2018-07-12 F. Troiani , K. Nikolic , T. G. Constandinou

Segmenting vocal tract articulators in real-time MRI (rtMRI) is a challenging dynamic image segmentation problem characterized by low contrast, rapid motion, and limited spatial resolution. However, while rtMRI acquisitions may provide…

Recent high-performance transformer-based speech enhancement models demonstrate that time domain methods could achieve similar performance as time-frequency domain methods. However, time-domain speech enhancement systems typically receive…

Sound · Computer Science 2023-10-31 Junhui Li , Pu Wang , Jialu Li , Xinzhe Wang , Youshan Zhang

It is imperative to ensure the robustness of deep learning models in critical applications such as, healthcare. While recent advances in deep learning have improved the performance of volumetric medical image segmentation models, these…

Image and Video Processing · Electrical Eng. & Systems 2023-07-21 Asif Hanif , Muzammal Naseer , Salman Khan , Mubarak Shah , Fahad Shahbaz Khan

Structured Finite Element Methods (FEMs) based on low-rank approximation in the form of the so-called Quantized Tensor Train (QTT) decomposition (QTT-FEM) have been proposed and extensively studied in the case of elliptic equations. In this…

Numerical Analysis · Mathematics 2024-11-19 Sara Fraschini , Vladimir Kazeev , Ilaria Perugia

Open-vocabulary panoptic reconstruction is crucial for advanced robotics and simulation. However, existing 3D reconstruction methods, such as NeRF or Gaussian Splatting variants, often struggle to achieve the real-time inference frequency…

Robotics · Computer Science 2026-04-14 Xuan Yu , Yuxuan Xie , Shichao Zhai , Shuhao Ye , Rong Xiong , Yue Wang

Audio-visual speech separation methods aim to integrate different modalities to generate high-quality separated speech, thereby enhancing the performance of downstream tasks such as speech recognition. Most existing state-of-the-art (SOTA)…

Sound · Computer Science 2024-03-22 Samuel Pegg , Kai Li , Xiaolin Hu

The automatic classification of 3D medical data is memory-intensive. Also, variations in the number of slices between samples is common. Na\"ive solutions such as subsampling can solve these problems, but at the cost of potentially…

Computer Vision and Pattern Recognition · Computer Science 2023-07-24 Marzieh Oghbaie , Teresa Araujo , Taha Emre , Ursula Schmidt-Erfurth , Hrvoje Bogunovic

This paper proposes a novel bidirectional neural vocoder, named BiVocoder, capable both of feature extraction and reverse waveform generation within the short-time Fourier transform (STFT) domain. For feature extraction, the BiVocoder takes…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-05 Hui-Peng Du , Ye-Xin Lu , Yang Ai , Zhen-Hua Ling

An alternative way of visualizing electromagnetic waves in matter and of deriving the Finite Difference Time Domain method (FDTD) for simulating Maxwell's equations for one dimensional systems is presented. The method uses d'Alembert's…

Optics · Physics 2021-02-24 Ross Hyman , Nathaniel Stern , Allen Taflove

Style transfer is a technique for combining two images based on the activations and feature statistics in a deep learning neural network architecture. This paper studies the analogous task in the audio domain and takes a critical look at…

Sound · Computer Science 2020-08-10 M. Huzaifah , L. Wyse

An efficient finite-difference time-domain (FDTD) algorithm is built to solve the transverse electric 2D Maxwell's equations with inhomogeneous dielectric media where the electric fields are discontinuous across the dielectric interface.…

Computational Physics · Physics 2023-07-14 Timothy Meagher , Bin Jiang , Peng Jiang

Audio Word2Vec offers vector representations of fixed dimensionality for variable-length audio segments using Sequence-to-sequence Autoencoder (SA). These vector representations are shown to describe the sequential phonetic structures of…

Computation and Language · Computer Science 2018-02-20 Chia-Hao Shen , Janet Y. Sung , Hung-Yi Lee