中文
相关论文

相关论文: ASGIR: Audio Spectrogram Transformer Guided Classi…

200 篇论文

Audio recognition in specialized areas such as birdsong and submarine acoustics faces challenges in large-scale pre-training due to the limitations in available samples imposed by sampling environments and specificity requirements. While…

声音 · 计算机科学 2023-09-26 Xiang Li , Junhao Chen , Chao Li , Hongwu Lv

Automatic Speech Recognition (ASR) for air traffic control is generally trained by pooling Air Traffic Controller (ATCO) and pilot data into one set. This is motivated by the fact that pilot's voice communications are more scarce than…

Open-vocabulary species recognition is a major challenge in computer vision, particularly in ornithology, where new taxa are continually discovered. While benchmarks like CUB-200-2011 and Birdsnap have advanced fine-grained recognition…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Faizan Farooq Khan , Jun Chen , Youssef Mohamed , Chun-Mei Feng , Mohamed Elhoseiny

We present a robust classification approach for avian vocalization in complex and diverse soundscapes, achieving second place in the BirdCLEF2021 challenge. We illustrate how to make full use of pre-trained convolutional neural networks, by…

声音 · 计算机科学 2021-07-19 Christof Henkel , Pascal Pfeiffer , Philipp Singer

Open audio databases such as Xeno-Canto are widely used to build datasets to explore bird song repertoire or to train models for automatic bird sound classification by deep learning algorithms. However, such databases suffer from the fact…

机器学习 · 计算机科学 2023-02-16 Félix Michaud , Jérôme Sueur , Maxime Le Cesne , Sylvain Haupert

As the technology is advancing, audio recognition in machine learning is improved as well. Research in audio recognition has traditionally focused on speech. Living creatures (especially the small ones) are part of the whole ecosystem,…

声音 · 计算机科学 2018-10-23 Siddhardha Balemarthy , Atul Sajjanhar , James Xi Zheng

Automatic Speech Recognition (ASR) can be used as the assistance of speech communication between pilots and air-traffic controllers. Its application can significantly reduce the complexity of the task and increase the reliability of…

计算与语言 · 计算机科学 2021-08-30 Iuliia Nigmatulina , Rudolf Braun , Juan Zuluaga-Gomez , Petr Motlicek

This project proposes the development of a comprehensive real-time biodiversity monitoring system that harnesses sound data through a network of acoustic sensors and advanced artificial intelligence algorithms. The system analyzes sound…

音频与语音处理 · 电气工程与系统科学 2024-10-18 Kumar Srinivas Bobba , Kartheeban K , Vamsi Krishna Sai , Dinesh Bugga , Vijaya Mani Surendra Bolla

Voice information retrieval is a technique that provides Information Retrieval System with the capacity to transcribe spoken queries and use the text output for information search. CIS is a field of research that involves studying the…

信息检索 · 计算机科学 2021-10-06 Sulaiman Adesegun Kukoyi , O. F. W Onifade , Kamorudeen A. Amuda

Automated bioacoustic analysis aids understanding and protection of both marine and terrestrial animals and their habitats across extensive spatiotemporal scales, and typically involves analyzing vast collections of acoustic data. With the…

音频与语音处理 · 电气工程与系统科学 2023-12-22 Burooj Ghani , Tom Denton , Stefan Kahl , Holger Klinck

Based on the transfer learning, we design a bird species identification model that uses the VGG-16 model (pretrained on ImageNet) for feature extraction, then a classifier consisting of two fully-connected hidden layers and a Softmax layer…

声音 · 计算机科学 2018-03-06 Jiang-jian Xie , Chang-qing Ding , Wen-bin Li , Cheng-hao Cai

Respiratory disease, the third leading cause of deaths globally, is considered a high-priority ailment requiring significant research on identification and treatment. Stethoscope-recorded lung sounds and artificial intelligence-powered…

声音 · 计算机科学 2024-05-15 Whenty Ariyanti , Kai-Chun Liu , Kuan-Yu Chen , Yu Tsao

Birds produce multiple types of vocalizations that, together, constitute a vocal repertoire. For some species, the repertoire size is of importance because it informs us about their brain capacity, territory size or social behaviour.…

定量方法 · 定量生物学 2023-03-21 Joachim Poutaraud

Deep learning models have significantly advanced acoustic bird monitoring by being able to recognize numerous bird species based on their vocalizations. However, traditional deep learning models are black boxes that provide no insight into…

机器学习 · 计算机科学 2024-11-14 René Heinrich , Lukas Rauch , Bernhard Sick , Christoph Scholz

Bioacoustics, the study of sounds produced by living organisms, plays a vital role in conservation, biodiversity monitoring, and behavioral studies. Many tasks in this field, such as species, individual, and behavior classification and…

Insects rely on their hearing in order to communicate, identify and locate potential mates, and avoid predators. Due to their small sizes, many insect species are not able to utilize the interaural time and intensity differences employed by…

Audio denoising has been explored for decades using both traditional and deep learning-based methods. However, these methods are still limited to either manually added artificial noise or lower denoised audio quality. To overcome these…

声音 · 计算机科学 2022-10-20 Youshan Zhang , Jialu Li

Passive acoustic monitoring (PAM) has shown great promise in helping ecologists understand the health of animal populations and ecosystems. However, extracting insights from millions of hours of audio recordings requires the development of…

Dialect variation hampers automatic recognition of bird calls collected by passive acoustic monitoring. We address the problem on DB3V, a three-region, ten-species corpus of 8-s clips, and propose a deployable framework built on Time-Delay…

声音 · 计算机科学 2025-09-29 Jiani Ding , Qiyang Sun , Alican Akman , Björn W. Schuller

Recording and analysing environmental audio recordings has become a common approach for monitoring the environment. A current problem with performing analyses of environmental recordings is interference from noise that can mask sounds of…

声音 · 计算机科学 2018-04-17 Alexander Brown , Saurabh Garg , James Montgomery