中文
相关论文

相关论文: Joint Scattering for Automatic Chick Call Recognit…

200 篇论文

ChatMOF is an autonomous Artificial Intelligence (AI) system that is built to predict and generate metal-organic frameworks (MOFs). By leveraging a large-scale language model (GPT-4 and GPT-3.5-turbo), ChatMOF extracts key details from…

计算与语言 · 计算机科学 2023-08-28 Yeonghun Kang , Jihan Kim

Spectrograms have been widely used in Convolutional Neural Networks based schemes for acoustic scene classification, such as the STFT spectrogram and the MFCC spectrogram, etc. They have different time-frequency characteristics,…

计算机视觉与模式识别 · 计算机科学 2018-09-06 Weiping Zheng , Zhenyao Mo , Xiaotao Xing , Gansen Zhao

Monitoring calf behaviour continuously would be beneficial to identify routine practices (e.g., weaning, dehorning, etc.) that impact calf welfare in dairy farms. In that regard, accelerometer data collected from neck collars can be used…

Intent detection and slot filling are two main tasks in natural language understanding and play an essential role in task-oriented dialogue systems. The joint learning of both tasks can improve inference accuracy and is popular in recent…

计算与语言 · 计算机科学 2022-05-17 Liang Huang , Senjie Liang , Feiyang Ye , Nan Gao

This paper investigates the temporal excitation patterns of creaky voice. Creaky voice is a voice quality frequently used as a phrase-boundary marker, but also as a means of portraying attitude, affective states and even social status.…

音频与语音处理 · 电气工程与系统科学 2020-06-02 Thomas Drugman , John Kane , Christer Gobl

In this study, we propose a framework for chirp-based communications by exploiting discrete Fourier transform-spread orthogonal frequency division multiplexing (DFT-s-OFDM). We show that a well-designed frequency-domain spectral shaping…

信号处理 · 电气工程与系统科学 2020-11-24 Alphan Sahin , Nozhan Hosseini , Hosseinali Jamal , Safi Shams Muhtasimul Hoque , David W. Matolak

Speaker verification is the process by which a speakers claim of identity is tested against a claimed speaker by his or her voice. Speaker verification is done by the use of some parameters (features) from the speakers voice which can be…

声音 · 计算机科学 2019-08-16 Bhavana V. S , Pradip K. Das

An adaptive time-frequency representation (TFR) with higher energy concentration usually requires higher complexity. Recently, a low-complexity adaptive short-time Fourier transform (ASTFT) based on the chirp rate has been proposed. To…

信息论 · 计算机科学 2017-05-26 Soo-Chang Pei , Shih-Gu Huang

Onsets are a key factor to split audio into several notes. In this paper, we ensemble multiple temporal convolution network (TCN) based model and utilize a restricted frequency range spectrogram to achieve more robust onset detection.…

声音 · 计算机科学 2023-06-09 Yu Cheng Hung , Jian-Jiun Ding

Connectionist temporal classification (CTC) -based models are attractive because of their fast inference in automatic speech recognition (ASR). Language model (LM) integration approaches such as shallow fusion and rescoring can improve the…

计算与语言 · 计算机科学 2022-09-07 Hayato Futami , Hirofumi Inaguma , Masato Mimura , Shinsuke Sakai , Tatsuya Kawahara

For high mobility communication scenario, the recently emerged orthogonal time frequency space (OTFS) modulation introduces a new delay-Doppler domain signal space, and can provide better communication performance than traditional…

信号处理 · 电气工程与系统科学 2022-08-16 Muye Li , Shun Zhang , Yao Ge , Feifei Gao , Pingzhi Fan

The performance of automatic speech recognition systems degrades with increasing mismatch between the training and testing scenarios. Differences in speaker accents are a significant source of such mismatch. The traditional approach to deal…

Research in semantic communication has garnered considerable attention, particularly in the area of image transmission, where joint source-channel coding (JSCC)-based neural network (NN) modules are frequently employed. However, these…

信号处理 · 电气工程与系统科学 2025-08-05 Yoon Huh , Bumjun Kim , Wan Choi

We describe here an experimental technique based on the acoustic scattering phenomenon allowing the direct probing of the vorticity field in a turbulent flow. Using time-frequency distributions, recently introduced in signal analysis…

chao-dyn · 物理学 2009-10-31 Christophe Baudet , Olivier Michel , William J. Williams

This paper studies the detection of bird calls in audio segments using stacked convolutional and recurrent neural networks. Data augmentation by blocks mixing and domain adaptation using a novel method of test mixing are proposed and…

声音 · 计算机科学 2017-06-08 Sharath Adavanne , Konstantinos Drossos , Emre Çakır , Tuomas Virtanen

In this study we argue that integrating ChatGPT into the data processing pipeline of automated sensors in precision agriculture has the potential to bring several benefits and enhance various aspects of modern farming practices. Policy…

人工智能 · 计算机科学 2023-11-14 Ilyas Potamitis

Automatic detection and classification of animal sounds has many applications in biodiversity monitoring and animal behaviour. In the past twenty years, the volume of digitised wildlife sound available has massively increased, and automatic…

In this work, we explore the constant-Q transform (CQT) for speech emotion recognition (SER). The CQT-based time-frequency analysis provides variable spectro-temporal resolution with higher frequency resolution at lower frequencies. Since…

音频与语音处理 · 电气工程与系统科学 2021-02-09 Premjeet Singh , Goutam Saha , Md Sahidullah

Motivated by future automotive applications, we study the joint target detection and parameter estimation problem using orthogonal time frequency space (OTFS), a digital modulation format robust to time-frequency selective channels.…

信号处理 · 电气工程与系统科学 2020-04-24 Lorenzo Gaudio , Mari Kobayashi , Giuseppe Caire , Giulio Colavolpe

The J-PET scanner, which allows for single bed imaging of the whole human body, is currently under development at the Jagiellonian University. The dis- cussed detector offers improvement of the Time of Flight (TOF) resolution due to the use…