English
Related papers

Related papers: ArEEG_Chars: Dataset for Envisioned Speech Recogni…

200 papers

Electroencephalography (EEG) is a complex signal and can require several years of training to be correctly interpreted. Recently, deep learning (DL) has shown great promise in helping make sense of EEG signals due to its capacity to learn…

Machine Learning · Computer Science 2019-01-23 Yannick Roy , Hubert Banville , Isabela Albuquerque , Alexandre Gramfort , Tiago H. Falk , Jocelyn Faubert

Electroencephalography (EEG) is a non-invasive technique for recording brain electrical activity, widely used in brain-computer interface (BCI) and healthcare. Recent EEG foundation models trained on large-scale datasets have shown improved…

Machine Learning · Computer Science 2025-09-29 Yi Ding , Muyun Jiang , Weibang Jiang , Shuailei Zhang , Xinliang Zhou , Chenyu Liu , Shanglin Li , Yong Li , Cuntai Guan

ArzEn-MultiGenre is a parallel dataset of Egyptian Arabic song lyrics, novels, and TV show subtitles that are manually translated and aligned with their English counterparts. The dataset contains 25,557 segment pairs that can be used to…

Computation and Language · Computer Science 2025-08-05 Rania Al-Sabbagh

Recognizing human non-speech vocalizations is an important task and has broad applications such as automatic sound transcription and health condition monitoring. However, existing datasets have a relatively small number of vocal sound…

Sound · Computer Science 2022-06-22 Yuan Gong , Jin Yu , James Glass

Decoding language representations directly from the brain can enable new Brain-Computer Interfaces (BCI) for high bandwidth human-human and human-machine communication. Clinically, such technologies can restore communication in people with…

Machine Learning · Computer Science 2019-09-05 Pengfei Sun , Gopala K. Anumanchipalli , Edward F. Chang

This study presents EgyBERT, an Arabic language model pretrained on 10.4 GB of Egyptian dialectal texts. We evaluated EgyBERT's performance by comparing it with five other multidialect Arabic language models across 10 evaluation datasets.…

Computation and Language · Computer Science 2024-08-08 Faisal Qarah

Brain-computer interfaces (BCIs) offer a pathway to restore communication for individuals with severe motor or speech impairments. Imagined handwriting provides an intuitive paradigm for character-level neural decoding, bridging the gap…

Signal Processing · Electrical Eng. & Systems 2025-10-24 Ovishake Sen , Raghav Soni , Darpan Virmani , Akshar Parekh , Patrick Lehman , Sarthak Jena , Adithi Katikhaneni , Adam Khalifa , Baibhab Chatterjee

Recently, there have been tremendous research outcomes in the fields of speech recognition and natural language processing. This is due to the well-developed multi-layers deep learning paradigms such as wav2vec2.0, Wav2vecU, WavBERT, and…

Computer Vision and Pattern Recognition · Computer Science 2021-10-12 Omar Mohamed , Salah A. Aly

Brain-computer interface (BCI) research, while promising, has largely been confined to static and fixed environments, limiting real-world applicability. To move towards practical BCI, we introduce a real-time wireless imagined speech…

Artificial Intelligence · Computer Science 2025-11-12 Ji-Ha Park , Heon-Gyu Kwak , Gi-Hwan Shin , Yoo-In Jeon , Sun-Min Park , Ji-Yeon Hwang , Seong-Whan Lee

Handwriting imagery has emerged as a promising paradigm for brain-computer interfaces (BCIs) aimed at translating brain activity into text output. Compared with invasively recorded electroencephalography (EEG), non-invasive recording offers…

Signal Processing · Electrical Eng. & Systems 2025-09-04 Hao Yang , Guang Ouyang

There are many difficulties facing a handwritten Arabic recognition system such as unlimited variation in human handwriting, similarities of distinct character shapes, interconnections of neighbouring characters and their position in the…

Computer Vision and Pattern Recognition · Computer Science 2014-02-27 Ahmed Sahlol , Cheng Suen

Current electroencephalogram (EEG) decoding models are typically trained on small numbers of subjects performing a single task. Here, we introduce a large-scale, code-submission-based competition comprising two challenges. First, the…

Ramsa is a developing 41-hour speech corpus of Emirati Arabic designed to support sociolinguistic research and low-resource language technologies. It contains recordings from structured interviews with native speakers and episodes from…

Computation and Language · Computer Science 2026-03-10 Rania Al-Sabbagh

Cybersickness poses a serious challenge for users of virtual reality (VR) technology. Consequently, there has been significant effort to track its occurrence during VR use with passive measures like brain activity recorded through…

Human-Computer Interaction · Computer Science 2026-03-30 Jacqueline Yau , Katherine J. Mimnaugh , Evan G. Center , Timo Ojala , Steven M. LaValle , Wenzhen Yuan , Nancy Amato , Minje Kim , Kara D. Federmeier

Deep learning has recently enabled the decoding of language from the neural activity of a few participants with electrodes implanted inside their brain. However, reliably decoding words from non-invasive recordings remains an open…

Signal Processing · Electrical Eng. & Systems 2024-12-25 Stéphane d'Ascoli , Corentin Bel , Jérémy Rapin , Hubert Banville , Yohann Benchetrit , Christophe Pallier , Jean-Rémi King

This paper delineates AISHELL-5, the first open-source in-car multi-channel multi-speaker Mandarin automatic speech recognition (ASR) dataset. AISHLL-5 includes two parts: (1) over 100 hours of multi-channel speech data recorded in an…

Sound · Computer Science 2025-05-30 Yuhang Dai , He Wang , Xingchen Li , Zihan Zhang , Shuiyuan Wang , Lei Xie , Xin Xu , Hongxiao Guo , Shaoji Zhang , Hui Bu , Wei Chen

Speech Emotion Recognition (SER) is one of the essential perceptual methods of humans in understanding the situation and how to interact with others, therefore, in recent years, it has been tried to add the ability to recognize emotions to…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-21 Ali Yazdani , Yasser Shekofteh

Auditory attention to natural speech is a complex brain process. Its quantification from physiological signals can be valuable to improving and widening the range of applications of current brain-computer-interface systems, however it…

Human-Computer Interaction · Computer Science 2020-05-26 Nikesh Bajaj , Jesús Requena Carrión , Francesco Bellotti

Despite major advancements in Automatic Speech Recognition (ASR), the state-of-the-art ASR systems struggle to deal with impaired speech even with high-resource languages. In Arabic, this challenge gets amplified, with added complexities in…

Sound · Computer Science 2023-06-08 Massa Baali , Ibrahim Almakky , Shady Shehata , Fakhri Karray

End-to-end speech Named Entity Recognition (NER) aims to directly extract entities from speech. Prior work has shown that end-to-end (E2E) approaches can outperform cascaded pipelines for English, French, and Chinese, but Arabic remains…

Computation and Language · Computer Science 2026-04-03 Youssef Saidi , Haroun Elleuch , Fethi Bougares