中文
相关论文

相关论文: GeHirNet: A Gender-Aware Hierarchical Model for Vo…

200 篇论文

Deep learning models designed for visual classification tasks on natural images have become prevalent in medical image analysis. However, medical images differ from typical natural images in many ways, such as significantly higher…

机器学习 · 计算机科学 2019-08-21 Yiqiu Shen , Nan Wu , Jason Phang , Jungkyu Park , Gene Kim , Linda Moy , Kyunghyun Cho , Krzysztof J. Geras

Microscopic histology image analysis is a cornerstone in early detection of breast cancer. However these images are very large and manual analysis is error prone and very time consuming. Thus automating this process is in high demand. We…

计算机视觉与模式识别 · 计算机科学 2018-10-23 Ismaël Koné , Lahsen Boulmane

Breast cancer is one of the leading causes of mortality in women. Early detection and treatment are imperative for improving survival rates, which have steadily increased in recent years as a result of more sophisticated…

计算机视觉与模式识别 · 计算机科学 2018-02-27 Sulaiman Vesal , Nishant Ravikumar , AmirAbbas Davari , Stephan Ellmann , Andreas Maier

Automated content analysis increasingly supports communication research, yet scaling manual coding into computational pipelines raises concerns about measurement reliability and validity. We introduce a Hierarchical Error Correction (HEC)…

计算与语言 · 计算机科学 2025-10-27 Zhilong Zhao , Yindi Liu

Semantic information refers to the meaning conveyed through words, phrases, and contextual relationships within a given linguistic structure. Humans can leverage semantic information, such as familiar linguistic patterns and contextual…

音频与语音处理 · 电气工程与系统科学 2025-02-06 Jixun Yao , Hexin Liu , Chen Chen , Yuchen Hu , EngSiong Chng , Lei Xie

Paralinguistic properties of speech are essential in analyzing and choosing optimal treatment options for patients with speech disorders. However, automatic modeling of these characteristics is difficult due to the lack of labeled speech…

音频与语音处理 · 电气工程与系统科学 2025-09-25 Jenthe Thienpondt , Geoffroy Vanderreydt , Abdessalem Hammami , Kris Demuynck

Accurate classification of laryngeal vascular as benign or malignant is crucial for early detection of laryngeal cancer. However, organizations with limited access to laryngeal vascular images face challenges due to the lack of large and…

计算机视觉与模式识别 · 计算机科学 2025-01-17 Xinyi Fang , Xu Yang , Chak Fong Chong , Kei Long Wong , Yapeng Wang , Tiankui Zhang , Sio-Kei Im

Surgical phase recognition is a challenging and necessary task for the development of context-aware intelligent systems that can support medical personnel for better patient care and effective operating room management. In this paper, we…

人机交互 · 计算机科学 2023-12-12 Kubilay Can Demir , Tobias Weise , Matthias May , Axel Schmid , Andreas Maier , Seung Hee Yang

One of the interests of modern poultry farming is the vocalization of laying hens which contain very useful information on health behavior. This information is used as health and well-being indicators that help breeders better monitor…

声音 · 计算机科学 2024-01-19 Fréjus A. A. Laleye , Mikaël A. Mousse

One way to extract patterns from clinical records is to consider each patient record as a bag with various number of instances in the form of symptoms. Medical diagnosis is to discover informative ones first and then map them to one or more…

机器学习 · 计算机科学 2019-04-10 Zeyuan Wang , Josiah Poon , Shiding Sun , Simon Poon

The widespread of powerful personal devices capable of collecting voice of their users has opened the opportunity to build speaker adapted speech recognition system (ASR) or to participate to collaborative learning of ASR. In both cases,…

计算与语言 · 计算机科学 2021-11-09 Salima Mdhaffar , Jean-François Bonastre , Marc Tommasi , Natalia Tomashenko , Yannick Estève

Motivation: Disease diagnosis oriented dialogue system models the interactive consultation procedure as Markov Decision Process and reinforcement learning algorithms are used to solve the problem. Existing approaches usually employ a flat…

人工智能 · 计算机科学 2023-11-08 Cheng Zhong , Kangenbei Liao , Wei Chen , Qianlong Liu , Baolin Peng , Xuanjing Huang , Jiajie Peng , Zhongyu Wei

Early diagnosis of critical diseases can significantly improve patient survival and reduce treatment costs. However, existing diagnostic techniques are often costly, invasive, and inaccessible in low-resource regions. This paper presents a…

计算机视觉与模式识别 · 计算机科学 2025-10-30 Manisha More , Kavya Bhand , Kaustubh Mukdam , Kavya Sharma , Manas Kawtikwar , Hridayansh Kaware , Prajwal Kavhar

Human gender classification based on biometric features is a major concern for computer vision due to its vast variety of applications. The human ear is popular among researchers as a soft biometric trait, because it is less affected by age…

计算机视觉与模式识别 · 计算机科学 2023-08-21 Ritwiz Singh , Keshav Kashyap , Rajesh Mukherjee , Asish Bera , Mamata Dalui Chakraborty

This work presents a novel framework based on feed-forward neural network for text-independent speaker classification and verification, two related systems of speaker recognition. With optimized features and model training, it achieves 100%…

声音 · 计算机科学 2017-03-20 Zhenhao Ge , Ananth N. Iyer , Srinath Cheluvaraja , Ram Sundaram , Aravind Ganapathiraju

There are growing implications surrounding generative AI in the speech domain that enable voice cloning and real-time voice conversion from one individual to another. This technology poses a significant ethical threat and could lead to…

声音 · 计算机科学 2023-08-25 Jordan J. Bird , Ahmad Lotfi

The massive spread of hate speech, hateful content targeted at specific subpopulations, is a problem of critical social importance. Automated methods of hate speech detection typically employ state-of-the-art deep learning (DL)-based text…

计算与语言 · 计算机科学 2022-05-23 Tomer Wullach , Amir Adler , Einat Minkov

It is well known that emotion recognition performance is not ideal. The work of this research is devoted to improving emotion recognition performance by employing a two-stage recognizer that combines and integrates gender recognizer and…

声音 · 计算机科学 2018-01-23 Ismail Shahin

The performance of most emotion recognition systems degrades in real-life situations ('in the wild' scenarios) where the audio is contaminated by reverberation. Our study explores new methods to alleviate the performance degradation of SER…

音频与语音处理 · 电气工程与系统科学 2024-09-17 Ohad Cohen , Gershon Hazan , Sharon Gannot

We conducted a comprehensive analysis of an Automatic Voice Disorders Detection (AVDD) system using existing voice disorder datasets with available demographic metadata. The study involved analysing system performance across various…

音频与语音处理 · 电气工程与系统科学 2025-04-15 Mariel Estevez , Cyntia Bonomi , Dayana Ribas , Alfonso Ortega , Luciana Ferrer