中文
相关论文

相关论文: Deep learning classification system for coconut ma…

200 篇论文

Deep learning has emerged as a powerful alternative to hand-crafted methods for emotion recognition on combined acoustic and text modalities. Baseline systems model emotion information in text and acoustic modes independently using Deep…

音频与语音处理 · 电气工程与系统科学 2020-10-13 Darshana Priyasad , Tharindu Fernando , Simon Denman , Clinton Fookes , Sridha Sridharan

Sound Event Detection and Audio Classification tasks are traditionally addressed through time-frequency representations of audio signals such as spectrograms. However, the emergence of deep neural networks as efficient feature extractors…

Robustness of Deep Neural Networks (DNNs) is an important aspect to consider for their clinical applications. This work examined robustness issue for a DNN-based multi-class classification model via comprehensive experimental and simulation…

医学物理 · 物理学 2023-03-07 Yuting Peng , Chenyang Shen , Yesenia Gonzalez , Yin Gao , Xun Jia

We study large-scale kernel methods for acoustic modeling and compare to DNNs on performance metrics related to both acoustic modeling and recognition. Measuring perplexity and frame-level classification accuracy, kernel-based acoustic…

This study explores the field of audio classification from raw waveform using Convolutional Neural Networks (CNNs), a method that eliminates the need for extracting specialised features in the pre-processing step. Unlike recent trends in…

声音 · 计算机科学 2024-12-03 Kazi Nazmul Haque , Rajib Rana , Tasnim Jarin , Bjorn W. Schuller

Many automated operations in agriculture, such as weeding and plant counting, require robust and accurate object detectors. Robotic fruit harvesting is one of these, and is an important technology to address the increasing labour shortages…

计算机视觉与模式识别 · 计算机科学 2021-05-11 Jasper Brown , Salah Sukkarieh

Deep learning (DL) has emerged as a powerful subset of machine learning (ML) and artificial intelligence (AI), outperforming traditional ML methods, especially in handling unstructured and large datasets. Its impact spans across various…

机器学习 · 计算机科学 2025-03-18 Farhad Mortezapour Shiri , Thinagaran Perumal , Norwati Mustapha , Raihani Mohamed

Ear recognition as a biometric modality is becoming increasingly popular, with promising broader application areas. While current applications involve adults, one of the challenges in ear recognition for children is the rapid structural…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Afzal Hossain , Tipu Sultan , Stephanie Schuckers

Deep learning techniques have shown promising results in the automatic classification of respiratory sounds. However, accurately distinguishing these sounds in real-world noisy conditions remains challenging for clinical deployment. In…

音频与语音处理 · 电气工程与系统科学 2025-05-01 Jing-Tong Tzeng , Jeng-Lin Li , Huan-Yu Chen , Chun-Hsiang Huang , Chi-Hsin Chen , Cheng-Yi Fan , Edward Pei-Chuan Huang , Chi-Chun Lee

We conduct an in depth study on the performance of deep learning based radio signal classification for radio communications signals. We consider a rigorous baseline method using higher order moments and strong boosted gradient tree…

机器学习 · 计算机科学 2018-03-14 Timothy J. O'Shea , Tamoghna Roy , T. Charles Clancy

This article focuses on signal classification for deep-sea acoustic neutrino detection. In the deep sea, the background of transient signals is very diverse. Approaches like matched filtering are not sufficient to distinguish between…

天体物理仪器与方法 · 物理学 2011-04-19 M. Neff , G. Anton , A. Enzenhöfer , K. Graf , J. Hößl , U. Katz , R. Lahmann , C. Richardt

Given the recent surge in developments of deep learning, this article provides a review of the state-of-the-art deep learning techniques for audio signal processing. Speech, music, and environmental sound processing are considered…

声音 · 计算机科学 2019-05-28 Hendrik Purwins , Bo Li , Tuomas Virtanen , Jan Schlüter , Shuo-yiin Chang , Tara Sainath

This study aimed to develop a deep learning model for the classification of bearing faults in wind turbine generators from acoustic signals. A convolutional LSTM model was successfully constructed and trained by using audio data from five…

声音 · 计算机科学 2024-03-15 Zhao Wang , Xiaomeng Li , Na Li , Longlong Shu

While both the data volume and heterogeneity of the digital music content is huge, it has become increasingly important and convenient to build a recommendation or search system to facilitate surfacing these content to the user or consumer…

Tomatoes are a major crop worldwide, and accurately classifying their maturity is important for many agricultural applications, such as harvesting, grading, and quality control. In this paper, the authors propose a novel method for tomato…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Asim Khan , Taimur Hassan , Muhammad Shafay , Israa Fahmy , Naoufel Werghi , Lakmal Seneviratne , Irfan Hussain

This dissertation presents several novel deep-learning (DL)-based approaches for classifying digitally modulated signals, one method of which involves the use of capsule networks (CAPs) together with cyclic cumulant (CC) features of the…

信号处理 · 电气工程与系统科学 2025-03-27 John A. Snoap

Automatic music genre classification is a long-standing challenge in Music Information Retrieval (MIR); work on non-Western music traditions remains scarce. Nepali music encompasses culturally rich and acoustically diverse genres--from the…

声音 · 计算机科学 2026-03-17 Sachin Prajuli , Abhishek Karna , OmPrakash Dhakl

The success of modern farming and plant breeding relies on accurate and efficient collection of data. For a commercial organization that manages large amounts of crops, collecting accurate and consistent data is a bottleneck. Due to limited…

计算机视觉与模式识别 · 计算机科学 2021-02-26 Saeed Khaki , Hieu Pham , Ye Han , Andy Kuhl , Wade Kent , Lizhi Wang

Tree fruit breeding is a long-term activity involving repeated measurements of various fruit quality traits on a large number of samples. These traits are traditionally measured by manually counting the fruits, weighing to indirectly…

计算机视觉与模式识别 · 计算机科学 2023-02-15 Ritayu Nagpal , Sam Long , Shahid Jahagirdar , Weiwei Liu , Scott Fazackerley , Ramon Lawrence , Amritpal Singh

We introduce a new audio processing technique that increases the sampling rate of signals such as speech or music using deep convolutional neural networks. Our model is trained on pairs of low and high-quality audio examples; at test-time,…

声音 · 计算机科学 2017-08-03 Volodymyr Kuleshov , S. Zayd Enam , Stefano Ermon