English
Related papers

Related papers: Joint Scattering for Automatic Chick Call Recognit…

200 papers

ChatMOF is an autonomous Artificial Intelligence (AI) system that is built to predict and generate metal-organic frameworks (MOFs). By leveraging a large-scale language model (GPT-4 and GPT-3.5-turbo), ChatMOF extracts key details from…

Computation and Language · Computer Science 2023-08-28 Yeonghun Kang , Jihan Kim

Spectrograms have been widely used in Convolutional Neural Networks based schemes for acoustic scene classification, such as the STFT spectrogram and the MFCC spectrogram, etc. They have different time-frequency characteristics,…

Computer Vision and Pattern Recognition · Computer Science 2018-09-06 Weiping Zheng , Zhenyao Mo , Xiaotao Xing , Gansen Zhao

Monitoring calf behaviour continuously would be beneficial to identify routine practices (e.g., weaning, dehorning, etc.) that impact calf welfare in dairy farms. In that regard, accelerometer data collected from neck collars can be used…

Machine Learning · Computer Science 2024-05-01 Oshana Dissanayake , Sarah E. McPherson , Joseph Allyndree , Emer Kennedy , Padraig Cunningham , Lucile Riaboff

Intent detection and slot filling are two main tasks in natural language understanding and play an essential role in task-oriented dialogue systems. The joint learning of both tasks can improve inference accuracy and is popular in recent…

Computation and Language · Computer Science 2022-05-17 Liang Huang , Senjie Liang , Feiyang Ye , Nan Gao

This paper investigates the temporal excitation patterns of creaky voice. Creaky voice is a voice quality frequently used as a phrase-boundary marker, but also as a means of portraying attitude, affective states and even social status.…

Audio and Speech Processing · Electrical Eng. & Systems 2020-06-02 Thomas Drugman , John Kane , Christer Gobl

In this study, we propose a framework for chirp-based communications by exploiting discrete Fourier transform-spread orthogonal frequency division multiplexing (DFT-s-OFDM). We show that a well-designed frequency-domain spectral shaping…

Signal Processing · Electrical Eng. & Systems 2020-11-24 Alphan Sahin , Nozhan Hosseini , Hosseinali Jamal , Safi Shams Muhtasimul Hoque , David W. Matolak

Speaker verification is the process by which a speakers claim of identity is tested against a claimed speaker by his or her voice. Speaker verification is done by the use of some parameters (features) from the speakers voice which can be…

Sound · Computer Science 2019-08-16 Bhavana V. S , Pradip K. Das

An adaptive time-frequency representation (TFR) with higher energy concentration usually requires higher complexity. Recently, a low-complexity adaptive short-time Fourier transform (ASTFT) based on the chirp rate has been proposed. To…

Information Theory · Computer Science 2017-05-26 Soo-Chang Pei , Shih-Gu Huang

Onsets are a key factor to split audio into several notes. In this paper, we ensemble multiple temporal convolution network (TCN) based model and utilize a restricted frequency range spectrogram to achieve more robust onset detection.…

Sound · Computer Science 2023-06-09 Yu Cheng Hung , Jian-Jiun Ding

Connectionist temporal classification (CTC) -based models are attractive because of their fast inference in automatic speech recognition (ASR). Language model (LM) integration approaches such as shallow fusion and rescoring can improve the…

Computation and Language · Computer Science 2022-09-07 Hayato Futami , Hirofumi Inaguma , Masato Mimura , Shinsuke Sakai , Tatsuya Kawahara

For high mobility communication scenario, the recently emerged orthogonal time frequency space (OTFS) modulation introduces a new delay-Doppler domain signal space, and can provide better communication performance than traditional…

Signal Processing · Electrical Eng. & Systems 2022-08-16 Muye Li , Shun Zhang , Yao Ge , Feifei Gao , Pingzhi Fan

The performance of automatic speech recognition systems degrades with increasing mismatch between the training and testing scenarios. Differences in speaker accents are a significant source of such mismatch. The traditional approach to deal…

Computation and Language · Computer Science 2018-02-09 Xuesong Yang , Kartik Audhkhasi , Andrew Rosenberg , Samuel Thomas , Bhuvana Ramabhadran , Mark Hasegawa-Johnson

Research in semantic communication has garnered considerable attention, particularly in the area of image transmission, where joint source-channel coding (JSCC)-based neural network (NN) modules are frequently employed. However, these…

Signal Processing · Electrical Eng. & Systems 2025-08-05 Yoon Huh , Bumjun Kim , Wan Choi

We describe here an experimental technique based on the acoustic scattering phenomenon allowing the direct probing of the vorticity field in a turbulent flow. Using time-frequency distributions, recently introduced in signal analysis…

chao-dyn · Physics 2009-10-31 Christophe Baudet , Olivier Michel , William J. Williams

This paper studies the detection of bird calls in audio segments using stacked convolutional and recurrent neural networks. Data augmentation by blocks mixing and domain adaptation using a novel method of test mixing are proposed and…

Sound · Computer Science 2017-06-08 Sharath Adavanne , Konstantinos Drossos , Emre Çakır , Tuomas Virtanen

In this study we argue that integrating ChatGPT into the data processing pipeline of automated sensors in precision agriculture has the potential to bring several benefits and enhance various aspects of modern farming practices. Policy…

Artificial Intelligence · Computer Science 2023-11-14 Ilyas Potamitis

Automatic detection and classification of animal sounds has many applications in biodiversity monitoring and animal behaviour. In the past twenty years, the volume of digitised wildlife sound available has massively increased, and automatic…

In this work, we explore the constant-Q transform (CQT) for speech emotion recognition (SER). The CQT-based time-frequency analysis provides variable spectro-temporal resolution with higher frequency resolution at lower frequencies. Since…

Audio and Speech Processing · Electrical Eng. & Systems 2021-02-09 Premjeet Singh , Goutam Saha , Md Sahidullah

Motivated by future automotive applications, we study the joint target detection and parameter estimation problem using orthogonal time frequency space (OTFS), a digital modulation format robust to time-frequency selective channels.…

Signal Processing · Electrical Eng. & Systems 2020-04-24 Lorenzo Gaudio , Mari Kobayashi , Giuseppe Caire , Giulio Colavolpe

The J-PET scanner, which allows for single bed imaging of the whole human body, is currently under development at the Jagiellonian University. The dis- cussed detector offers improvement of the Time of Flight (TOF) resolution due to the use…