English
Related papers

Related papers: SonoTraceLab -- A Raytracing-Based Acoustic Modell…

200 papers

In this paper, we study the use of soft labels to train a system for sound event detection (SED). Soft labels can result from annotations which account for human uncertainty about categories, or emerge as a natural representation of…

Audio and Speech Processing · Electrical Eng. & Systems 2023-03-01 Irene Martín-Morató , Manu Harju , Paul Ahokas , Annamaria Mesaros

Acoustic sonar imaging systems are widely used for underwater surveillance in both civilian and military sectors. However, acquiring high-quality sonar datasets for training Artificial Intelligence (AI) models confronts challenges such as…

Computer Vision and Pattern Recognition · Computer Science 2024-08-26 Kamal Basha S , Athira Nambiar

Rectilinear forms of snake-like robotic locomotion are anticipated to be an advantage in obstacle-strewn scenarios characterizing urban disaster zones, subterranean collapses, and other natural environments. The elongated, laterally-narrow…

Robotics · Computer Science 2019-08-21 Alexander H. Chang , Shiyu Feng , Yipu Zhao , Justin S. Smith , Patricio A. Vela

We introduce EgoSonics, a method to generate semantically meaningful and synchronized audio tracks conditioned on silent egocentric videos. Generating audio for silent egocentric videos could open new applications in virtual reality,…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Aashish Rai , Srinath Sridhar

Sparse Autoencoders (SAEs) are powerful tools for interpreting neural representations, yet their use in audio remains underexplored. We train SAEs across all encoder layers of Whisper and HuBERT, provide an extensive evaluation of their…

In this work, we present the development of a new database, namely Sound Localization and Classification (SLoClas) corpus, for studying and analyzing sound localization and classification. The corpus contains a total of 23.27 hours of data…

Sound · Computer Science 2021-08-06 Xinyuan Qian , Bidisha Sharma , Amine El Abridi , Haizhou Li

Voice plays an important role in our lives by facilitating communication, conveying emotions, and indicating health. Therefore, tracking vocal interactions can provide valuable insight into many aspects of our lives. This paper presents our…

Sound · Computer Science 2023-12-19 Sheen An Goh , Manoj Gulati , Ambuj Varshney

Conventional approaches to sound localization and separation are based on microphone arrays in artificial systems. Inspired by the selective perception of human auditory system, we design a multi-source listening system which can separate…

Sound · Computer Science 2019-11-11 Xuecong Sun , Han Jia , Zhe Zhang , Yuzhen Yang , Zhaoyong Sun , Jun Yang

Prior work in computational bioacoustics has mostly focused on the detection of animal presence in a particular habitat. However, animal sounds contain much richer information than mere presence; among others, they encapsulate the…

Imitation learning in robots, also called programing by demonstration, has made important advances in recent years, allowing humans to teach context dependant motor skills/tasks to robots. We propose to extend the usual contexts…

Artificial Intelligence · Computer Science 2012-03-13 Thomas Cederborg , Pierre-Yves Oudeyer

The development and application of modern technology is an essential basis for the efficient monitoring of species in natural habitats and landscapes to trace the development of ecosystems, species communities, and populations, and to…

Computer Vision and Pattern Recognition · Computer Science 2022-11-24 Timm Haucke , Hjalmar S. Kühl , Volker Steinhage

Localizing the bronchoscope in real time is essential for ensuring intervention quality. However, most existing methods struggle to balance between speed and generalization. To address these challenges, we present BronchoTrack, an…

Computer Vision and Pattern Recognition · Computer Science 2024-02-21 Qingyao Tian , Huai Liao , Xinyan Huang , Bingyu Yang , Jinlin Wu , Jian Chen , Lujie Li , Hongbin Liu

Spectrum sensing is one of the means of utilizing the scarce source of wireless spectrum efficiently. In this paper, a convolutional neural network (CNN) model employing spectral correlation function which is an effective characterization…

Signal Processing · Electrical Eng. & Systems 2021-04-29 Kürşat Tekbıyık , Özkan Akbunar , Ali Rıza Ekti , Ali Görçin , Güneş Karabulut Kurt , Khalid A. Qaraqe

This paper presents MicroRoboScope, a portable, compact, and versatile microrobotic experimentation platform designed for real-time, closed-loop control of both magnetic and acoustic microrobots. The system integrates an embedded computer,…

Robotics · Computer Science 2025-10-14 Max Sokolich , Yanda Yang , Subrahmanyam Cherukumilli , Fatma Ceren Kirmizitas , Sambeeta Das

Precision livestock farming requires advanced monitoring tools to meet the increasing management needs of the industry. Computer vision systems capable of long-term multi-animal tracking (MAT) are essential for continuous behavioral…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Anne Marthe Sophie Ngo Bibinbe , Patrick Gagnon , Jamie Ahloy-Dallaire , Eric R. Paquet

Automatic speaker verification, like every other biometric system, is vulnerable to spoofing attacks. Using only a few minutes of recorded voice of a genuine client of a speaker verification system, attackers can develop a variety of…

Sound · Computer Science 2019-06-20 Balamurali BT , Kin Wah Edward Lin , Simon Lui , Jer-Ming Chen , Dorien Herremans

Electroencephalogram (EEG) is a non-invasive technique to record bioelectrical signals. Integrating supervised deep learning techniques with EEG signals has recently facilitated automatic analysis across diverse EEG-based tasks. However,…

Signal Processing · Electrical Eng. & Systems 2024-01-12 Weining Weng , Yang Gu , Shuai Guo , Yuan Ma , Zhaohua Yang , Yuchen Liu , Yiqiang Chen

Sound Event Localization and Detection (SELD) is crucial in spatial audio processing, enabling systems to detect sound events and estimate their 3D directions. Existing SELD methods use single- or dual-branch architectures: single-branch…

Sound · Computer Science 2025-07-31 Hogeon Yu

Realistic recordings of soundscapes often have multiple sound events co-occurring, such as car horns, engine and human voices. Sound event retrieval is a type of content-based search aiming at finding audio samples, similar to an audio…

Audio and Speech Processing · Electrical Eng. & Systems 2020-02-24 Jianyu Fan , Eric Nichols , Daniel Tompkins , Ana Elisa Mendez Mendez , Benjamin Elizalde , Philippe Pasquier

Generalization in audio deepfake detection presents a significant challenge, with models trained on specific datasets often struggling to detect deepfakes generated under varying conditions and unknown algorithms. While collectively…

‹ Prev 1 8 9 10 Next ›