English
Related papers

Related papers: F0 analysis of Ghanaian pop singing reveals progre…

200 papers

This study examines the pitch content in traditional Ghanaian seperewa (Akan harp-lute) songs, utilizing a unique dataset from field recordings of the mid-twentieth century. We selected 71 songs and used Demucs to isolate vocals from…

Sound · Computer Science 2024-11-14 Kelvin L Walls , Iran R Roman , Kelsey Van Ert , Colter Harper , Leila Adu-Gilmore

Large language models (LLMs) have demonstrated impressive multilingual capabilities for well-resourced languages, yet their performance on low-resource African languages remains poorly understood and largely unevaluated. This paper presents…

Large language models (LLMs) are increasingly deployed in applications with societal impact, raising concerns about the cultural biases they encode. We probe these representations by evaluating whether LLMs can perform author profiling from…

Computation and Language · Computer Science 2026-03-20 Valentin Lafargue , Ariel Guerra-Adames , Emmanuelle Claeys , Elouan Vuichard , Jean-Michel Loubes

In this paper, we propose GaMMA, a state-of-the-art (SoTA) large multimodal model (LMM) designed to achieve comprehensive musical content understanding. GaMMA inherits the streamlined encoder-decoder design of LLaVA, enabling effective…

Sound · Computer Science 2026-05-04 Zuyao You , Zhesong Yu , Mingyu Liu , Bilei Zhu , Yuan Wan , Zuxuan Wu

Scales, sets of discrete pitches that form the basis of melodies, are thought to be one of the most universal hallmarks of music. But we know relatively little about cross-cultural diversity of scales or how they evolved. To remedy this, we…

Physics and Society · Physics 2023-05-03 John M McBride , Sam Passmore , Tsvi Tlusty

This study investigates the use of self-supervised learning embeddings, particularly BYOL-A, in conjunction with a deep neural network classifier for Music Genre Classification. Our experiments demonstrate that BYOL-A embeddings outperform…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-17 Kashish Rai , Mrinmoy Bhattacharjee

We approach the singing phrase audio to score matching problem by using phonetic and duration information - with a focus on studying the jingju a cappella singing case. We argue that, due to the existence of a basic melodic contour for each…

Sound · Computer Science 2017-07-13 Rong Gong , Jordi Pons , Xavier Serra

Multimodal learning integrates diverse modalities but suffers from modality imbalance, where dominant modalities suppress weaker ones due to inconsistent convergence rates. Existing methods predominantly rely on static modulation or…

Machine Learning · Computer Science 2026-02-11 Zhaocheng Liu , Zhiwen Yu , Xiaoqing Liu

A recitation is a way of combining the words together so that they have a sense of rhythm and thus an emotional content is imbibed within. In this study we envisaged to answer these questions in a scientific manner taking into consideration…

Audio and Speech Processing · Electrical Eng. & Systems 2020-08-06 Chirayata Bhattacharyya , Sourya Sengupta , Sayan Nag , Shankha Sanyal , Archi Banerjee , Ranjan Sengupta , Dipak Ghosh

Current approaches to music emotion annotation remain heavily reliant on manual labelling, a process that imposes significant resource and labour burdens, severely limiting the scale of available annotated data. This study examines the…

Sound · Computer Science 2025-08-19 Meng Yang , Jon McCormack , Maria Teresa Llano , Wanchao Su

Large Language Models (LLMs) have shown remarkable performance across various tasks, yet significant disparities remain for non-English languages, and especially native African languages. This paper addresses these disparities by creating…

Computation and Language · Computer Science 2024-12-18 Tuka Alhanai , Adam Kasumovic , Mohammad Ghassemi , Aven Zitzelberger , Jessica Lundin , Guillaume Chabot-Couture

Warning: Contains harmful model outputs. Despite significant advancements, the propensity of Large Language Models (LLMs) to generate harmful and unethical content poses critical challenges. Measuring value alignment of LLMs becomes crucial…

Computation and Language · Computer Science 2025-06-12 Han Jiang , Xiaoyuan Yi , Zhihua Wei , Ziang Xiao , Shu Wang , Xing Xie

The influence of Large Language Models (LLMs) is rapidly growing, automating more jobs over time. Assessing the fairness of LLMs is crucial due to their expanding impact. Studies reveal the reflection of societal norms and biases in LLMs,…

Computation and Language · Computer Science 2024-07-10 Jayanta Sadhu , Maneesha Rani Saha , Rifat Shahriyar

An estimation problem of fundamental interest is that of phase synchronization, in which the goal is to recover a collection of phases using noisy measurements of relative phases. It is known that in the Gaussian noise setting, the maximum…

Optimization and Control · Mathematics 2016-11-02 Huikang Liu , Man-Chung Yue , Anthony Man-Cho So

We consider constraints on generalized tachyon field (GTF) models from latest observational data (including 182 gold SNIa data, the shift parameter, and the acoustic scale). We obtain at 68.3% confidence level $\Omega_{\rm m}=0.37\pm0.01$,…

Astrophysics · Physics 2009-06-25 Rong-Jia Yang , Shuang Nan Zhang , Yuan Liu

One of the fundamental questions of cultural evolutionary research is how individual-level processes scale up to generate population-level patterns. Previous studies in music have revealed that frequency-based bias (e.g. conformity and…

Applications · Statistics 2019-07-01 Mason Youngblood

Large Language Models (LLMs), such as ChatGPT, are widely used to generate content for various purposes and audiences. However, these models may not reflect the cultural and emotional diversity of their users, especially for low-resource…

Computation and Language · Computer Science 2024-07-01 Ibrahim Said Ahmad , Shiran Dudy , Resmi Ramachandranpillai , Kenneth Church

General speech restoration demands techniques that can interpret complex speech structures under various distortions. While State-Space Models like SEMamba have advanced the state-of-the-art in speech denoising, they are not inherently…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-13 Yongjoon Lee , Jung-Woo Choi

Deep learning has achieved strong performance for electrocardiogram (ECG) classification within individual datasets, yet dependable generalization across heterogeneous acquisition settings remains a major obstacle to clinical deployment and…

Machine Learning · Computer Science 2025-12-30 Hai Duong Nguyen , Xuan-The Tran

Pitch variability in rap vocals is overlooked in favor of the genre's uniquely dynamic rhythmic properties. We present an analysis of fundamental frequency (F0) variation in rap vocals over the past 14 years, focusing on song examples that…

Sound · Computer Science 2023-12-22 Kelvin L Walls , Iran R Roman , Bea Steers , Elena Georgieva
‹ Prev 1 2 3 10 Next ›