中文
相关论文

相关论文: Automatic Assessment of Dysarthria Using Audio-vis…

200 篇论文

Speech classifiers of paralinguistic traits traditionally learn from diverse hand-crafted low-level features, by selecting the relevant information for the task at hand. We explore an alternative to this selection, by learning jointly the…

计算与语言 · 计算机科学 2019-01-09 Juliette Millet , Neil Zeghidour

How can a machine learn to recognize visual attributes emerging out of online community without a definitive supervised dataset? This paper proposes an automatic approach to discover and analyze visual attributes from a noisy collection of…

计算机视觉与模式识别 · 计算机科学 2016-07-26 Sirion Vittayakorn , Takayuki Umeda , Kazuhiko Murasaki , Kyoko Sudo , Takayuki Okatani , Kota Yamaguchi

Depression and Attention Deficit Hyperactivity Disorder (ADHD) stand out as the common mental health challenges today. In affective computing, speech signals serve as effective biomarkers for mental disorder assessment. Current research,…

音频与语音处理 · 电气工程与系统科学 2025-03-05 Shuanglin Li , Siyang Song , Rajesh Nair , Syed Mohsen Naqvi

Learning disabilities, which primarily interfere with the basic learning skills such as reading, writing and math, are known to affect around 10% of children in the world. The poor motor skills and motor coordination as part of the…

机器学习 · 计算机科学 2022-06-28 Jayakanth Kunhoth , Somaya Al-Maadeed , Suchithra Kunhoth , Younus Akbari

Dysarthria is a speech disorder that hinders communication due to difficulties in articulating words. Detection of dysarthria is important for several reasons as it can be used to develop a treatment plan and help improve a person's quality…

Dysgraphia, a handwriting learning disability, has a serious negative impact on children's academic results, daily life and overall wellbeing. Early detection of dysgraphia allows for an early start of a targeted intervention. Several…

Visual explanation is an approach for visualizing the grounds of judgment by deep learning, and it is possible to visually interpret the grounds of a judgment for a certain input by visualizing an attention map. As for deep-learning models…

人工智能 · 计算机科学 2023-06-06 Kohei Hattori , Tsubasa Hirakawa , Takayoshi Yamashita , Hironobu Fujiyoshi

Hypernasality is a common characteristic symptom across many motor-speech disorders. For voiced sounds, hypernasality introduces an additional resonance in the lower frequencies and, for unvoiced sounds, there is reduced articulatory…

音频与语音处理 · 电气工程与系统科学 2020-09-14 Michael Saxon , Ayush Tripathi , Yishan Jiao , Julie Liss , Visar Berisha

To facilitate diagnosis on cardiac ultrasound (US), clinical practice has established several standard views of the heart, which serve as reference points for diagnostic measurements and define viewports from which images are acquired.…

图像与视频处理 · 电气工程与系统科学 2024-03-04 Sarina Thomas , Cristiana Tiago , Børge Solli Andreassen , Svein Arne Aase , Jurica Šprem , Erik Steen , Anne Solberg , Guy Ben-Yosef

In recent years, deep learning models have been applied to neuroimaging data for early diagnosis of Alzheimer's disease (AD). Structural magnetic resonance imaging (sMRI) and positron emission tomography (PET) images provide structural and…

图像与视频处理 · 电气工程与系统科学 2023-08-01 Yanteng Zhanga , Xiaohai He , Yi Hao Chan , Qizhi Teng , Jagath C. Rajapakse

Dysarthric speech recognition (DSR) research has witnessed remarkable progress in recent years, evolving from the basic understanding of individual words to the intricate comprehension of sentence-level expressions, all driven by the…

声音 · 计算机科学 2025-10-21 Shiyao Wang , Shiwan Zhao , Jiaming Zhou , Yong Qin

In recent years, deep learning has achieved remarkable success in the field of image restoration. However, most convolutional neural network-based methods typically focus on a single scale, neglecting the incorporation of multi-scale…

图像与视频处理 · 电气工程与系统科学 2025-02-27 Jiatao Jiang , Zhen Cui , Chunyan Xu , Jian Yang

We propose a fully automated algorithm based on a deep learning framework enabling screening of a coronary computed tomography angiography (CCTA) examination for confident detection of the presence or absence of coronary artery…

图像与视频处理 · 电气工程与系统科学 2020-06-09 Sema Candemir , Richard D. White , Mutlu Demirer , Vikash Gupta , Matthew T. Bigelow , Luciano M. Prevedello , Barbaros S. Erdal

As the first diagnostic imaging modality of avascular necrosis of the femoral head (AVNFH), accurately staging AVNFH from a plain radiograph is critical yet challenging for orthopedists. Thus, we propose a deep learning-based AVNFH…

图像与视频处理 · 电气工程与系统科学 2020-11-11 Yang Li , Yan Li , Hua Tian

Despite the rapid progress of automatic speech recognition (ASR) technologies targeting normal speech in recent decades, accurate recognition of dysarthric and elderly speech remains highly challenging tasks to date. Sources of…

音频与语音处理 · 电气工程与系统科学 2022-03-18 Mengzhe Geng , Xurong Xie , Zi Ye , Tianzi Wang , Guinan Li , Shujie Hu , Xunying Liu , Helen Meng

This study investigates deep learning methods for automated classification of dental conditions in panoramic X-ray images. A dataset of 1,512 radiographs with 11,137 expert-verified annotations across four conditions fillings, cavities,…

计算机视觉与模式识别 · 计算机科学 2025-09-01 Alireza Golkarieh , Kiana Kiashemshaki , Sajjad Rezvani Boroujeni

Dysarthria is a condition which hampers the ability of an individual to control the muscles that play a major role in speech delivery. The loss of fine control over muscles that assist the movement of lips, vocal chords, tongue and…

声音 · 计算机科学 2021-03-11 Ayush Tripathi , Swapnil Bhosale , Sunil Kumar Kopparapu

Dysarthric speech severity assessment typically requires trained clinicians or supervised models built from labelled pathological speech, limiting scalability across languages and clinical settings. We present a training-free method that…

计算与语言 · 计算机科学 2026-04-14 Bernard Muller , Antonio Armando Ortiz Barrañón , LaVonne Roberts

Dysarthric speech recognition faces challenges from severity variations and disparities relative to normal speech. Conventional approaches individually fine-tune ASR models pre-trained on normal speech per patient to prevent feature…

声音 · 计算机科学 2025-08-27 Qing Xiao , Yingshan Peng , PeiPei Zhang

State-of-the-art automatic speech recognition (ASR) systems perform well on healthy speech. However, the performance on impaired speech still remains an issue. The current study explores the usefulness of using Wav2Vec self-supervised…