中文
相关论文

相关论文: Mind Your Vision: Multimodal Estimation of Refract…

200 篇论文

Electrooculography (EOG) is an electrophysiological signal that determines the human eye orientation and is therefore widely used in Human Tracking Interfaces (HCI). The purpose of this project is to develop a communication method for…

信号处理 · 电气工程与系统科学 2025-01-07 Aniket Raj , Amit Kumar

In this study, we introduce an innovative EEG signal reconstruction sub-module designed to enhance the performance of deep learning models on EEG eye-tracking tasks. This sub-module can integrate with all Encoder-Classifier-based deep…

人机交互 · 计算机科学 2024-08-13 Weigeng Li , Neng Zhou , Xiaodong Qu

This work investigates the predictive potential of bipolar electroencephalogram (EEG) recordings towards efficient prediction of poor neurological outcomes. A retrospective design using a hybrid deep learning approach is utilized to…

信号处理 · 电气工程与系统科学 2023-10-09 Hemin Ali Qadir , Naimahmed Nesaragi , Per Steiner Halvorsen , Ilangko Balasingham

Eye movement prediction is a promising area of research with the potential to improve performance and the user experience of systems based on eye-tracking technology. In this study, we analyze individual differences in gaze prediction…

人机交互 · 计算机科学 2025-01-30 Kateryna Melnyk , Lee Friedman , Dmytro Katrychuk , Oleg Komogortsev

In many visual systems, visual tracking often bases on RGB image sequences, in which some targets are invalid in low-light conditions, and tracking performance is thus affected significantly. Introducing other modalities such as depth and…

计算机视觉与模式识别 · 计算机科学 2021-11-12 Chenglong Li , Tianhao Zhu , Lei Liu , Xiaonan Si , Zilin Fan , Sulan Zhai

Studies on emotion recognition (ER) show that combining lexical and acoustic information results in more robust and accurate models. The majority of the studies focus on settings where both modalities are available in training and…

计算与语言 · 计算机科学 2019-06-26 Gustavo Aguilar , Viktor Rozgić , Weiran Wang , Chao Wang

This work presents a new multimodal system for remote attention level estimation based on multimodal face analysis. Our multimodal approach uses different parameters and signals obtained from the behavior and physiological processes that…

计算机视觉与模式识别 · 计算机科学 2023-01-24 Roberto Daza , Luis F. Gomez , Aythami Morales , Julian Fierrez , Ruben Tolosana , Ruth Cobos , Javier Ortega-Garcia

Parkinson's Disease (PD) often results in motor and cognitive impairments, including gait dysfunction, particularly in patients with freezing of gait (FOG). Current detection methods are either subjective or reliant on specialized gait…

Reading ability detection is important in modern educational field. In this paper, a method of predicting scores of reading ability is proposed, using the eye-tracking data of a few subjects (e.g., 68 subjects). The proposed method built a…

人机交互 · 计算机科学 2024-09-16 Nanxi Li , Hongjiang Wang , Zehui Zhan

Emotion recognition is essential for applications in affective computing and behavioral prediction, but conventional systems relying on single-modality data often fail to capture the complexity of affective states. To address this…

多媒体 · 计算机科学 2025-09-08 Jianlu Wang , Yanan Wang , Tong Liu

Mental fatigue is a leading cause of motor vehicle accidents, medical errors, loss of workplace productivity, and student disengagements in e-learning environment. Development of sensors and systems that can reliably track mental fatigue…

人机交互 · 计算机科学 2023-09-12 Prabin Sharma , Joanna C. Justus , Megha Thapa , Govinda R. Poudel

Decoding images from non-invasive electroencephalographic (EEG) signals has been a grand challenge in understanding how the human brain process visual information in real-world scenarios. To cope with the issues of signal-to-noise ratio and…

信号处理 · 电气工程与系统科学 2024-06-26 Chi-Sheng Chen , Chun-Shu Wei

Object pose estimation is a long-standing problem in computer vision. Recently, attention-based vision transformer models have achieved state-of-the-art results in many computer vision applications. Exploiting the permutation-invariant…

计算机视觉与模式识别 · 计算机科学 2023-12-14 Arul Selvam Periyasamy , Vladimir Tsaturyan , Sven Behnke

Each year, multi-modal interaction continues to grow within both industry and academia. However, researchers have yet to fully explore the impact of multi-modal systems on learning and memory retention. This research investigates how…

人机交互 · 计算机科学 2025-09-09 Omar Elgohary , Zhu-Tien

We present an on-line 3D visual object tracking framework for monocular cameras by incorporating spatial knowledge and uncertainty from semantic mapping along with high frequency measurements from visual odometry. Using a combination of…

计算机视觉与模式识别 · 计算机科学 2016-03-15 Prateek Singhal , Ruffin White , Henrik Christensen

Understanding the correlation between EEG features and cognitive tasks is crucial for elucidating brain function. Brain activity synchronizes during speaking and listening tasks. However, it is challenging to estimate task-dependent brain…

神经元与认知 · 定量生物学 2024-10-01 Dai Shimizu , Ko Watanabe , Andreas Dengel

Electrooculography (EOG) is widely used for gaze tracking in Human-Robot Collaboration (HRC). However, baseline drift caused by low-frequency noise significantly impacts the accuracy of EOG signals, creating challenges for further sensor…

信号处理 · 电气工程与系统科学 2025-09-10 Lianming Hu , Xiaotong Zhang , Kamal Youcef-Toumi

Light field data has been demonstrated to facilitate the depth estimation task. Most learning-based methods estimate the depth infor-mation from EPI or sub-aperture images, while less methods pay attention to the focal stack. Existing…

计算机视觉与模式识别 · 计算机科学 2021-04-14 Yongri Piao , Xinxin Ji , Miao Zhang , Yukun Zhang

Assistive visual navigation systems for visually impaired individuals have become increasingly popular thanks to the rise of mobile computing. Most of these devices work by translating visual information into voice commands. In complex…

计算机视觉与模式识别 · 计算机科学 2024-12-18 Hao Wang , Jiayou Qin , Xiwen Chen , Ashish Bastola , John Suchanek , Zihao Gong , Abolfazl Razi

Self-supervised learning has greatly facilitated medical image analysis by suppressing the training data requirement for real-world applications. Current paradigms predominantly rely on self-supervision within uni-modal image data, thereby…

计算机视觉与模式识别 · 计算机科学 2025-03-31 Shaohao Rui , Lingzhi Chen , Zhenyu Tang , Lilong Wang , Mianxin Liu , Shaoting Zhang , Xiaosong Wang