中文
相关论文

相关论文: OpenEDS2020: Open Eyes Dataset

200 篇论文

Surgical phase recognition has gained significant attention due to its potential to offer solutions to numerous demands of the modern operating room. However, most existing methods concentrate on minimally invasive surgery (MIS), leaving…

计算机视觉与模式识别 · 计算机科学 2024-11-28 Ryo Fujii , Masashi Hatano , Hideo Saito , Hiroki Kajita

We propose a novel neural pipeline, MSGazeNet, that learns gaze representations by taking advantage of the eye anatomy information through a multistream framework. Our proposed solution comprises two components, first a network for…

计算机视觉与模式识别 · 计算机科学 2024-03-04 Zunayed Mahmud , Paul Hungler , Ali Etemad

Data size is the bottleneck for developing deep saliency models, because collecting eye-movement data is very time consuming and expensive. Most of current studies on human attention and saliency modeling have used high quality stereotype…

计算机视觉与模式识别 · 计算机科学 2019-11-20 Zhaohui Che , Ali Borji , Guangtao Zhai , Xiongkuo Min , Guodong Guo , Patrick Le Callet

In this paper, we evaluate a synthetic framework to be used in the field of gaze estimation employing deep learning techniques. The lack of sufficient annotated data could be overcome by the utilization of a synthetic evaluation framework…

计算机视觉与模式识别 · 计算机科学 2020-06-15 Gonzalo Garde , Andoni Larumbe-Bergera , Benoît Bossavit , Rafael Cabeza , Sonia Porta , Arantxa Villanueva

The volumetric representation of human interactions is one of the fundamental domains in the development of immersive media productions and telecommunication applications. Particularly in the context of the rapid advancement of Extended…

计算机视觉与模式识别 · 计算机科学 2024-02-15 Fatemeh Ghorbani Lohesara , Davi Rabbouni Freitas , Christine Guillemot , Karen Eguiazarian , Sebastian Knorr

Eye tracking is a key technology for gaze-based interactions in Extended Reality (XR), but traditional frame-based systems struggle to meet XR's demands for high accuracy, low latency, and power efficiency. Event cameras offer a promising…

计算机视觉与模式识别 · 计算机科学 2024-09-25 Junyuan Ding , Ziteng Wang , Chang Gao , Min Liu , Qinyu Chen

Despite progress in vision-based inspection algorithms, real-world industrial challenges -- specifically in data availability, quality, and complex production requirements -- often remain under-addressed. We introduce the VISION Datasets, a…

计算机视觉与模式识别 · 计算机科学 2023-06-21 Haoping Bai , Shancong Mou , Tatiana Likhomanenko , Ramazan Gokberk Cinbis , Oncel Tuzel , Ping Huang , Jiulong Shan , Jianjun Shi , Meng Cao

The recent breakthroughs in computer vision have benefited from the availability of large representative datasets (e.g. ImageNet and COCO) for training. Yet, robotic vision poses unique challenges for applying visual algorithms developed…

计算机视觉与模式识别 · 计算机科学 2020-03-09 Qi She , Fan Feng , Xinyue Hao , Qihan Yang , Chuanlin Lan , Vincenzo Lomonaco , Xuesong Shi , Zhengwei Wang , Yao Guo , Yimin Zhang , Fei Qiao , Rosa H. M. Chan

This paper addresses the challenging problem of estimating the general visual attention of people in images. Our proposed method is designed to work across multiple naturalistic social scenarios and provides a full picture of the subject's…

计算机视觉与模式识别 · 计算机科学 2018-07-30 Eunji Chong , Nataniel Ruiz , Yongxin Wang , Yun Zhang , Agata Rozga , James Rehg

Eye movements can provide informative cues to understand human visual scan/search behavior and cognitive load during varying tasks. Visualizations of real-time gaze measures during tasks, provide an understanding of human behavior as the…

人机交互 · 计算机科学 2024-09-11 Gavindya Jayawardena , Vikas Ashok , Sampath Jayarathna

Interpretation of giga-pixel whole-slide images (WSIs) is an important but difficult task for pathologists. Their diagnostic accuracy is estimated to average around 70%. Adding a second pathologist does not substantially improve decision…

计算机视觉与模式识别 · 计算机科学 2025-10-29 Veronica Thai , Rui Li , Meng Ling , Shuning Jiang , Jeremy Wolfe , Raghu Machiraju , Yan Hu , Zaibo Li , Anil Parwani , Jian Chen

As robots become more present in open human environments, it will become crucial for robotic systems to understand and predict human motion. Such capabilities depend heavily on the quality and availability of motion capture data. However,…

Human eye contact is a form of non-verbal communication and can have a great influence on social behavior. Since the location and size of the eye contact targets vary across different videos, learning a generic video-independent eye contact…

计算机视觉与模式识别 · 计算机科学 2022-10-06 Tianyi Wu , Yusuke Sugano

Eye movements provide a window into human behaviour, attention, and interaction dynamics. Challenges in real-world, multi-person environments have, however, restrained eye-tracking research predominantly to single-person, in-lab settings.…

Event-based vision revolutionizes traditional image sensing by capturing asynchronous intensity variations rather than static frames, enabling ultrafast temporal resolution, sparse data encoding, and enhanced motion perception. While this…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Joey Mulé , Dhandeep Challagundla , Rachit Saini , Riadul Islam

The interaction between the vestibular and ocular system has primarily been studied in controlled environments. Consequently, off-the shelf tools for categorization of gaze events (e.g. fixations, pursuits, saccade) fail when head movements…

计算机视觉与模式识别 · 计算机科学 2020-06-11 Rakshit Kothari , Zhizhuo Yang , Christopher Kanan , Reynold Bailey , Jeff Pelz , Gabriel Diaz

Deep neural networks for video-based eye tracking have demonstrated resilience to noisy environments, stray reflections, and low resolution. However, to train these networks, a large number of manually annotated images are required. To…

计算机视觉与模式识别 · 计算机科学 2020-12-22 Nitinraj Nair , Rakshit Kothari , Aayush K. Chaudhary , Zhizhuo Yang , Gabriel J. Diaz , Jeff B. Pelz , Reynold J. Bailey

Video object segmentation (VOS) aims to segment specified target objects throughout a video. Although state-of-the-art methods have achieved impressive performance (e.g., 90+% J&F) on benchmarks such as DAVIS and YouTube-VOS, these datasets…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Henghui Ding , Kaining Ying , Chang Liu , Shuting He , Xudong Jiang , Yu-Gang Jiang , Philip H. S. Torr , Song Bai

Existing datasets for autonomous driving (AD) often lack diversity and long-range capabilities, focusing instead on 360{\deg} perception and temporal reasoning. To address this gap, we introduce Zenseact Open Dataset (ZOD), a large-scale…

Whole-eye optical coherence tomography (WEOCT) has emerged as a transformative imaging modality capable of simultaneously capturing the anterior and posterior segments of the human eye. WEOCT enables comprehensive ocular biometry, which is…