中文
相关论文

相关论文: Toward Imagined Speech based Smart Communication S…

200 篇论文

An electroencephalography (EEG) based Brain Computer Interface (BCI) enables people to communicate with the outside world by interpreting the EEG signals of their brains to interact with devices such as wheelchairs and intelligent robots.…

人机交互 · 计算机科学 2017-09-27 Xiang Zhang , Lina Yao , Quan Z. Sheng , Salil S. Kanhere , Tao Gu , Dalin Zhang

As the Metaverse continues to grow, the need for efficient communication and intelligent content generation becomes increasingly important. Semantic communication focuses on conveying meaning and understanding from user inputs, while…

人机交互 · 计算机科学 2023-07-25 Yijing Lin , Zhipeng Gao , Hongyang Du , Dusit Niyato , Jiawen Kang , Abbas Jamalipour , Xuemin Sherman Shen

Creating a digital world that closely mimics the real world with its many complex interactions and outcomes is possible today through advanced emulation software and ubiquitous computing power. Such a software-based emulation of an entity…

信号处理 · 电气工程与系统科学 2023-05-18 Batool Salehi , Utku Demir , Debashri Roy , Suyash Pradhan , Jennifer Dy , Stratis Ioannidis , Kaushik Chowdhury

Semantic communication, which focuses on conveying the meaning of information rather than exact bit reconstruction, has gained considerable attention in recent years. Meanwhile, reconfigurable intelligent surface (RIS) is a promising…

信息论 · 计算机科学 2023-06-30 Jiajia Shi , Tse-Tin Chan , Haoyuan Pan , Tat-Ming Lok

This paper reports on a study in which a novel virtual moving sound-based spatial auditory brain-computer interface (BCI) paradigm is developed. Classic auditory BCIs rely on spatially static stimuli, which are often boring and difficult to…

神经元与认知 · 定量生物学 2014-01-08 Yohann Lelievre , Tomasz M. Rutkowski

The way we engage with digital spaces and the digital world has undergone rapid changes in recent years, largely due to the emergence of the Metaverse. As technology continues to advance, the demand for sophisticated and immersive…

人机交互 · 计算机科学 2024-09-04 Senthil Kumar Jagatheesaperumal , Praveen Sathikumar , Harikrishnan Rajan

Studies have shown that in noisy acoustic environments, providing binaural signals to the user of an assistive listening device may improve speech intelligibility and spatial awareness. This paper presents a binaural speech enhancement…

音频与语音处理 · 电气工程与系统科学 2024-03-11 Vikas Tokala , Eric Grinstein , Mike Brookes , Simon Doclo , Jesper Jensen , Patrick A. Naylor

Novel text-to-speech systems can generate entirely new voices that were not seen during training. However, it remains a difficult task to efficiently create personalized voices from a high-dimensional speaker space. In this work, we use…

BCI systems are able to communicate directly between the brain and computer using neural activity measurements without the involvement of muscle movements. For BCI systems to be widely used by people with severe disabilities, long-term…

人机交互 · 计算机科学 2023-05-31 Krishna Pai , Rakhee Kallimani , Sridhar Iyer , B. Uma Maheswari , Rajashri Khanai , Dattaprasad Torse

Task-oriented dialogue systems (TDSs) are assessed mainly in an offline setting or through human evaluation. The evaluation is often limited to single-turn or is very time-intensive. As an alternative, user simulators that mimic user…

计算与语言 · 计算机科学 2023-11-06 Weiwei Sun , Shuyu Guo , Shuo Zhang , Pengjie Ren , Zhumin Chen , Maarten de Rijke , Zhaochun Ren

Noninvasive brain-computer interface (BCI) is widely used to recognize users' intentions. Especially, BCI related to tactile and sensation decoding could provide various effects on many industrial fields such as manufacturing advanced touch…

人机交互 · 计算机科学 2020-12-22 Jeong-Hyun Cho , Ji-Hoon Jeong , Myoung-Ki Kim , Seong-Whan Lee

Controlling the style and characteristics of speech synthesis is crucial for adapting the output to specific contexts and user requirements. Previous Text-to-speech (TTS) works have focused primarily on the technical aspects of producing…

声音 · 计算机科学 2025-09-04 Jiawei Zhang , Tian-Hao Zhang , Jun Wang , Jiaran Gao , Xinyuan Qian , Xu-Cheng Yin

Text-to-speech (TTS) has advanced from generating natural-sounding speech to enabling fine-grained control over attributes like emotion, timbre, and style. Driven by rising industrial demand and breakthroughs in deep learning, e.g.,…

计算与语言 · 计算机科学 2025-08-26 Tianxin Xie , Yan Rong , Pengfei Zhang , Wenwu Wang , Li Liu

Expressive text-to-speech has shown improved performance in recent years. However, the style control of synthetic speech is often restricted to discrete emotion categories and requires training data recorded by the target speaker in the…

计算与语言 · 计算机科学 2022-07-14 Yookyung Shin , Younggun Lee , Suhee Jo , Yeongtae Hwang , Taesu Kim

The cross-speaker emotion transfer task in text-to-speech (TTS) synthesis particularly aims to synthesize speech for a target speaker with the emotion transferred from reference speech recorded by another (source) speaker. During the…

声音 · 计算机科学 2022-04-11 Tao Li , Xinsheng Wang , Qicong Xie , Zhichao Wang , Lei Xie

The Metaverse represents a transformative shift beyond traditional mobile Internet, creating an immersive, persistent digital ecosystem where users can interact, socialize, and work within 3D virtual environments. Powered by large models…

计算机与社会 · 计算机科学 2025-08-12 Yuntao Wang , Qinnan Hu , Zhou Su , Linkang Du , Qichao Xu , Weiwei Li

Synthesized speech is common today due to the prevalence of virtual assistants, easy-to-use tools for generating and modifying speech signals, and remote work practices. Synthesized speech can also be used for nefarious purposes, including…

声音 · 计算机科学 2022-05-05 Emily R. Bartusiak , Edward J. Delp

One of the current AI issues depicted in popular culture is the fear of conscious super AIs that try to take control over humanity. And as computational power goes upwards and that turns more and more into a reality, understanding…

神经元与认知 · 定量生物学 2023-05-22 Daniel Lopes

Chat GPT belongs to the category of Generative Pre-trained Transformer (GPT) language models, which have received specialized training to produce text based on natural language inputs. Its purpose is to imitate human-like conversation and…

人机交互 · 计算机科学 2023-11-03 Wei Zhou

We present a meta-learning approach for adaptive text-to-speech (TTS) with few data. During training, we learn a multi-speaker model using a shared conditional WaveNet core and independent learned embeddings for each speaker. The aim of…

‹ 上一页 1 8 9 10 下一页 ›