中文
相关论文

相关论文: Unsupervised Welding Defect Detection Using Audio …

200 篇论文

The nature of the atomic defects on the hydrogen passivated Si (100) surface is analyzed using deep learning and scanning tunneling microscopy (STM). A robust deep learning framework capable of identifying atomic species, defects, in the…

材料科学 · 物理学 2020-02-19 Maxim Ziatdinov , Udi Fuchs , James H. G. Owen , John N. Randall , Sergei V. Kalinin

This paper presents the development of a multi-sensor user interface to facilitate the instruction of arc welding tasks. Traditional methods to acquire hand-eye coordination skills are typically conducted through one-to-one instruction…

人机交互 · 计算机科学 2022-06-29 Hoi-Yin Lee , Peng Zhou , Anqing Duan , Jiangliu Wang , Victor Wu , David Navarro-Alarcon

Material defects (MD) represent a primary challenge affecting product performance and giving rise to safety issues in related products. The rapid and accurate identification and localization of MD constitute crucial research endeavors in…

计算机视觉与模式识别 · 计算机科学 2025-05-06 Jun Bai , Di Wu , Tristan Shelley , Peter Schubel , David Twine , John Russell , Xuesen Zeng , Ji Zhang

Detecting facial action units (AU) is one of the fundamental steps in automatic recognition of facial expression of emotions and cognitive states. Though there have been a variety of approaches proposed for this task, most of these models…

计算机视觉与模式识别 · 计算机科学 2019-11-28 Mihee Lee , Ognjen Rudovic , Vladimir Pavlovic , Maja Pantic

Deep Learning refers to a set of machine learning techniques that utilize neural networks with many hidden layers for tasks, such as image classification, speech recognition, language understanding. Deep learning has been proven to be very…

机器学习 · 计算机科学 2017-05-02 Andre Luckow , Matthew Cook , Nathan Ashcraft , Edwin Weill , Emil Djerekarov , Bennie Vorster

Wind turbines are subjected to continuous rotational stresses and unusual external forces such as storms, lightning, strikes by flying objects, etc., which may cause defects in turbine blades. Hence, it requires a periodical inspection to…

计算机视觉与模式识别 · 计算机科学 2024-01-02 Md Fazle Rabbi , Solayman Hossain Emon , Ehtesham Mahmud Nishat , Tzu-Liang , Tseng , Atira Ferdoushi , Chun-Che Huang , Md Fashiar Rahman

Medical Imagings are considered one of the crucial diagnostic tools for different bones-related diseases, especially bones fractures. This paper investigates the robustness of pre-trained deep learning models for classifying bone fractures…

图像与视频处理 · 电气工程与系统科学 2025-07-15 Robby Hoover , Nelly Elsayed , Zag ElSayed , Chengcheng Li

As the labor force decreases, the demand for labor-saving automatic anomalous sound detection technology that conducts maintenance of industrial equipment has grown. Conventional approaches detect anomalies based on the reconstruction…

音频与语音处理 · 电气工程与系统科学 2020-05-20 Kaori Suefusa , Tomoya Nishida , Harsh Purohit , Ryo Tanabe , Takashi Endo , Yohei Kawaguchi

In this paper we present a research on identification of audio recording devices from background noise, thus providing a method for forensics. The audio signal is the sum of speech signal and noise signal. Usually, people pay more attention…

声音 · 计算机科学 2016-04-28 Simeng Qi , Zheng Huang , Yan Li , Shaopei Shi

Artificial Intelligence has gained a lot of attention recently, it has been utilized in several fields ranging from daily life activities, such as responding to emails and scheduling appointments, to manufacturing and automating work…

软件工程 · 计算机科学 2026-02-02 Mohammed O. Alannsary

Self-supervised representations excel at many vision and speech tasks, but their potential for audio-visual deepfake detection remains underexplored. Unlike prior work that uses these features in isolation or buried within complex…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Dragos-Alexandru Boldisor , Stefan Smeu , Dan Oneata , Elisabeta Oneata

With the prevalence of artificial intelligence (AI)-generated content, such as audio deepfakes, a large body of recent work has focused on developing deepfake detection techniques. However, most models are evaluated on a narrow set of…

音频与语音处理 · 电气工程与系统科学 2025-09-29 Yi Zhu , Heitor R. Guimarães , Arthur Pimentel , Tiago Falk

In this paper, a novel optical inspection system is presented that is directly suitable for Industry 4.0 and the implementation on IoT-devices controlling the manufacturing process. The proposed system is capable of distinguishing between…

图像与视频处理 · 电气工程与系统科学 2022-09-29 Andreas Spruck , Jürgen Seiler , Michael Roll , Thomas Dudziak , Jürgen Eckstein , André Kaup

We present an AI-assisted Augmented Reality assembly workflow that uses deep learning-based object recognition to identify different assembly components and display step-by-step instructions. For each assembly step, the system displays a…

计算机视觉与模式识别 · 计算机科学 2025-11-17 Alexander Htet Kyaw , Haotian Ma , Sasa Zivkovic , Jenny Sabin

Existing methods on audio-visual deepfake detection mainly focus on high-level features for modeling inconsistencies between audio and visual data. As a result, these approaches usually overlook finer audio-visual artifacts, which are…

计算机视觉与模式识别 · 计算机科学 2024-10-15 Marcella Astrid , Enjie Ghorbel , Djamila Aouada

Face manipulation technology is advancing very rapidly, and new methods are being proposed day by day. The aim of this work is to propose a deepfake detector that can cope with the wide variety of manipulation methods and scenarios…

计算机视觉与模式识别 · 计算机科学 2023-05-19 Davide Cozzolino , Alessandro Pianese , Matthias Nießner , Luisa Verdoliva

Real-time detection of irregularities in visual data is very invaluable and useful in many prospective applications including surveillance, patient monitoring systems, etc. With the surge of deep learning methods in the recent years,…

计算机视觉与模式识别 · 计算机科学 2018-07-19 Mohammad Sabokrou , Masoud Pourreza , Mohsen Fayyaz , Rahim Entezari , Mahmood Fathy , Jürgen Gall , Ehsan Adeli

This paper describes sound event localization and detection (SELD) for spatial audio recordings captured by firstorder ambisonics (FOA) microphones. In this task, one may train a deep neural network (DNN) using FOA data annotated with the…

声音 · 计算机科学 2024-10-31 Yoto Fujita , Yoshiaki Bando , Keisuke Imoto , Masaki Onishi , Kazuyoshi Yoshii

Understanding human behavior and activity facilitates advancement of numerous real-world applications, and is critical for video analysis. Despite the progress of action recognition algorithms in trimmed videos, the majority of real-world…

计算机视觉与模式识别 · 计算机科学 2021-10-04 Elahe Vahdani , Yingli Tian

This work focuses on reliable detection of bird sound emissions as recorded in the open field. Acoustic detection of avian sounds can be used for the automatized monitoring of multiple bird taxa and querying in long-term recordings for…

声音 · 计算机科学 2016-09-28 Ilyas Potamitis