中文
相关论文

相关论文: GazePrior: Zero-Shot AR/VR Eye Tracking via Learne…

200 篇论文

High-quality 3D urban reconstruction is essential for applications in urban planning, navigation, and AR/VR. However, capturing detailed ground-level data across cities is both labor-intensive and raises significant privacy concerns related…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Fuqiang Zhao , Yijing Guo , Siyuan Yang , Xi Chen , Luo Wang , Lan Xu , Yingliang Zhang , Yujiao Shi , Jingyi Yu

Gaze prediction plays a critical role in Virtual Reality (VR) applications by reducing sensor-induced latency and enabling computationally demanding techniques such as foveated rendering, which rely on anticipating user attention. However,…

计算机视觉与模式识别 · 计算机科学 2026-01-27 Christos Petrou , Harris Partaourides , Athanasios Balomenos , Yannis Kopsinis , Sotirios Chatzis

Fast and accurate eye tracking in a virtual reality or augmented reality headset could lead to better display performance and enable novel methods of user interaction with the system. However, it remains a challenge for a system to combine…

3D reconstruction and view synthesis are foundational problems in computer vision, graphics, and immersive technologies such as augmented reality (AR), virtual reality (VR), and digital twins. Traditional methods rely on computationally…

Transparent object grasping remains a persistent challenge in robotics, largely due to the difficulty of acquiring precise 3D information. Conventional optical 3D sensors struggle to capture transparent objects, and machine learning methods…

机器人学 · 计算机科学 2025-04-15 Yi Han , Zixin Lin , Dongjie Li , Lvping Chen , Yongliang Shi , Gan Ma

3D and 2D gaze estimation share the fundamental objective of capturing eye movements but are traditionally treated as two distinct research domains. In this paper, we introduce a novel cross-task few-shot 2D gaze estimation approach, aiming…

计算机视觉与模式识别 · 计算机科学 2025-03-26 Yihua Cheng , Hengfei Wang , Zhongqun Zhang , Yang Yue , Bo Eun Kim , Feng Lu , Hyung Jin Chang

Vision-based Bird's Eye View (BEV) representation is an emerging perception formulation for autonomous driving. The core challenge is to construct BEV space with multi-camera features, which is a one-to-many ill-posed problem. Diving into…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Yiming Wu , Ruixiang Li , Zequn Qin , Xinhai Zhao , Xi Li

The latest developments in computer hardware, sensor technologies, and artificial intelligence can make virtual reality (VR) and virtual spaces an important part of human everyday life. Eye tracking offers not only a hands-free way of…

The main challenges of using electroencephalogram (EEG) signals to make eye-tracking (ET) predictions are the differences in distributional patterns between benchmark data and real-world data and the noise resulting from the unintended…

机器学习 · 计算机科学 2022-08-02 Brian Xiang , Abdelrahman Abdelmonsef

We propose a new dataset and a novel approach to learning hand-object interaction priors for hand and articulated object pose estimation. We first collect a dataset using visual teleoperation, where the human operator can directly play…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Zehao Zhu , Jiashun Wang , Yuzhe Qin , Deqing Sun , Varun Jampani , Xiaolong Wang

Research investigating cognitive aspects of information systems is often dependent on detail-rich data. Eye-trackers promise to provide respective data, but the associated costs are often beyond the researchers' budget. Recently,…

人机交互 · 计算机科学 2015-11-16 Stefan Zugal , Jakob Pinggera

Training perceptive humanoid locomotion policies that traverse complex terrains with natural gaits remains an open challenge, typically demanding multi-stage training pipelines, adversarial objectives, or extensive real-world calibration.…

机器人学 · 计算机科学 2026-03-20 Chenxi Han , Shilu He , Yi Cheng , Linqi Ye , Houde Liu

Augmented Reality (AR) headsets continuously sense their surroundings, capturing nearby bystanders and raising privacy risks. Visual bystander privacy-enhancing technologies (PETs) mitigate this risk by detecting bystanders in egocentric…

密码学与安全 · 计算机科学 2026-05-29 Syed Ibrahim Mustafa Shah Bukhari , Matthew Corbett , Bo Ji , Brendan David-John

Despite the recent development of learning-based gaze estimation methods, most methods require one or more eye or face region crops as inputs and produce a gaze direction vector as output. Cropping results in a higher resolution in the eye…

计算机视觉与模式识别 · 计算机科学 2023-05-10 Haldun Balim , Seonwook Park , Xi Wang , Xucong Zhang , Otmar Hilliges

The current apprenticeship model for surgical training requires a high level of supervision, which does not scale well to meet the growing need for more surgeons. Many endoscopic procedures are directly taught in the operating room (OR)…

人机交互 · 计算机科学 2025-06-30 Jumanh Atoum , Jinkyung Park , Mamtaj Akter , Nicholas Kavoussi , Pamela Wisniewski , Jie Ying Wu

Gaze-tracking is a novel way of interacting with computers which allows new scenarios, such as enabling people with motor-neuron disabilities to control their computers or doctors to interact with patient information without touching screen…

人工智能 · 计算机科学 2020-10-13 Jatin Sharma , Jon Campbell , Pete Ansell , Jay Beavers , Christopher O'Dowd

In this paper, we present GazeTrak, the first acoustic-based eye tracking system on glasses. Our system only needs one speaker and four microphones attached to each side of the glasses. These acoustic sensors capture the formations of the…

Various stuff and things in visual data possess specific traits, which can be learned by deep neural networks and are implicitly represented as the visual prior, e.g., object location and shape, in the model. Such prior potentially impacts…

计算机视觉与模式识别 · 计算机科学 2023-05-31 Jinheng Xie , Kai Ye , Yudong Li , Yuexiang Li , Kevin Qinghong Lin , Yefeng Zheng , Linlin Shen , Mike Zheng Shou

Eye tracking (ET) plays a critical role in augmented and virtual reality applications. However, rapidly deploying high-accuracy, on-device gaze estimation for new products remains challenging because hardware configurations (e.g., camera…

计算机视觉与模式识别 · 计算机科学 2026-04-06 Cheng Jiang , Jogendra Kundu , David Colmenares , Fengting Yang , Joseph Robinson , Yatong An , Ali Behrooz

Trajectory-Guided image-to-video (I2V) generation aims to synthesize videos that adhere to user-specified motion instructions. Existing methods typically rely on computationally expensive fine-tuning on scarce annotated datasets. Although…

计算机视觉与模式识别 · 计算机科学 2025-12-10 Ruicheng Zhang , Jun Zhou , Zunnan Xu , Zihao Liu , Jiehui Huang , Mingyang Zhang , Yu Sun , Xiu Li